Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Brian Karrer — most-cited papers & profile · Multimodal
← authors
·
overview
Brian Karrer
9
papers ·
51
citations ·
26
h-index
Pennsylvania State University · Menlo School
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
2023 · 45 citations
Adjoint Matching: Fine-tuning Flow and Diffusion Generative Models with Memoryless Stochastic Optimal Control
2024 · 2 citations
Set Block Decoding is a Language Model Inference Accelerator
2025
Generator Matching: Generative modeling with arbitrary Markov processes
2024
Training-free Linear Image Inverses via Flows
2023
Training-free Linear Image Inverses via Flows
2023
Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction Tuning
2023
Topics
Diffusion Models
Training & Sampling
Conditioning & Control
Model Architecture
Efficiency
Fine-Tuning
Vision-Language
Training Techniques
Text-to-Speech
Audio Generation