Awesome Computer Vision
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β authors
Β·
overview
Loading authorβ¦
π€
Ask AI
Hang Yan β most-cited papers & profile Β· Computer Vision
β authors
Β·
overview
Hang Yan
27
papers Β·
233
citations Β·
0
h-index
Jiangsu University
Google Scholar β
Semantic Scholar β
OpenAlex β
Most-cited papers
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
2024 Β· 230 citations
InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD
2024 Β· 2 citations
Genius: A Generalizable and Purely Unsupervised Self-Training Framework For Advanced Reasoning
2025 Β· 1 citations
TIDE: Trajectory-based Diagnostic Evaluation of Test-Time Improvement in LLM Agents
2026
Building Self-Evolving Agents via Experience-Driven Lifelong Learning: A Framework and Benchmark
2025
What Makes a Good Speech Tokenizer for LLM-Centric Speech Generation? A Systematic Study
2025
BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping
2025
Nex-N1: Agentic Models Trained via a Unified Ecosystem for Large-Scale Environment Construction
2025
CoLLiE: Collaborative Training of Large Language Models in an Efficient Way
2023
Identifying Semantic Induction Heads to Understand In-Context Learning
2024
Hint-before-Solving Prompting: Guiding LLMs to Effectively Utilize Encoded Knowledge
2024
InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output
2024
Secrets of RLHF in Large Language Models Part I: PPO
2023
Secrets of RLHF in Large Language Models Part II: Reward Modeling
2024
InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model
2024
Topics
Training Techniques
Code
Evaluation
Reinforcement Learning
Efficiency
Model Architecture
Fine-Tuning
In-Context Learning
Safety & Alignment
Vision-Language