Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hang Yan — most-cited papers & profile · Multimodal
← authors
·
overview
Hang Yan
25
papers ·
357
citations ·
0
h-index
Guangdong University of Technology · Hefei University of Technology · Hunan University · Ministry of Education · Guangdong Academy of Sciences
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
2024 · 230 citations
Identifying Semantic Induction Heads To Understand In-context Learning
2024 · 53 citations
Secrets of RLHF in Large Language Models Part I: PPO
2023 · 19 citations
Secrets of RLHF in Large Language Models Part II: Reward Modeling
2024 · 7 citations
InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD
2024 · 2 citations
Genius: A Generalizable and Purely Unsupervised Self-Training Framework For Advanced Reasoning
2025 · 1 citations
TIDE: Trajectory-based Diagnostic Evaluation of Test-Time Improvement in LLM Agents
2026
BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping
2025
What Makes a Good Speech Tokenizer for LLM-Centric Speech Generation? A Systematic Study
2025
BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping
2025
Nex-N1: Agentic Models Trained via a Unified Ecosystem for Large-Scale Environment Construction
2025
Secrets of RLHF in Large Language Models Part I: PPO
2023
Secrets of RLHF in Large Language Models Part II: Reward Modeling
2024
InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model
2024
InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning
2024
Topics
Training Techniques
Model Architecture
Reinforcement Learning
Evaluation
Code
Vision-Language
In-Context Learning
Safety & Alignment
Efficiency
Agentic