Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Sergey Tulyakov — most-cited papers & profile · Multimodal
← authors
·
overview
Sergey Tulyakov
28
papers ·
80
citations ·
37
h-index
University of Santa Monica · Snap (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
InfinityGAN: Towards Infinite-Pixel Image Synthesis
2021 · 20 citations
Text2Tex: Text-driven Texture Synthesis via Diffusion Models
2023 · 13 citations
NeROIC: Neural Rendering of Objects from Online Image Collections
2022 · 12 citations
AutoDecoding Latent 3D Diffusion Models
2023 · 9 citations
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image Generation
2024 · 9 citations
Transformable Bottleneck Networks
2019 · 8 citations
In&Out : Diverse Image Outpainting via GAN Inversion
2021 · 4 citations
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models
2025 · 1 citations
E$^2$GAN: Efficient Training of Efficient GANs for Image-to-Image Translation
2024 · 1 citations
SF-V: Single Forward Video Generation Model
2024 · 1 citations
SF-V: Single Forward Video Generation Model
2024 · 1 citations
BitsFusion: 1.99 bits Weight Quantization of Diffusion Model
2024 · 1 citations
Improving the Diffusability of Autoencoders
2025
SnapGen-V: Generating a Five-Second Video within Five Seconds on a Mobile Device
2024
Visual Concept-driven Image Generation with Text-to-Image Diffusion Model
2024
Topics
Diffusion Models
Image Generation
3D & NeRF Generation
Training & Sampling
Text-to-Video
Text-to-Image
3D Vision
Conditioning & Control
Audio Generation
GANs