Awesome Similarity Search
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Ran He — most-cited papers & profile · Similarity Search
← authors
·
overview
Ran He
21
papers ·
14
citations ·
63
h-index
University of Chinese Academy of Sciences
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
2024 · 11 citations
VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction
2025 · 3 citations
VideoDetective: Clue Hunting via both Extrinsic Query and Intrinsic Relevance for Long Video Understanding
2026
Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses
2026
Overconfident Errors Need Stronger Correction: Asymmetric Confidence Penalties for Reinforcement Learning
2026
FlashPrefill: Instantaneous Pattern Discovery and Thresholding for Ultra-Fast Long-Context Prefilling
2026
ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning
2026
Beyond GRPO and On-Policy Distillation: An Empirical Sparse-to-Dense Reward Principle for Language-Model Post-Training
2026
OmniUMI: Towards Physically Grounded Robot Learning via Human-Aligned Multimodal Interaction
2026
MixLM: High-Throughput and Effective LLM Ranking via Text-Embedding Mix-Interaction
2025
VITA-VLA: Efficiently Teaching Vision-Language Models to Act via Action Expert Distillation
2025
Cooperative Pseudo Labeling for Unsupervised Federated Classification
2025
Rethinking the Role of Prompting Strategies in LLM Test-Time Scaling: A Perspective of Probability Theory
2025
Marmot: Object-Level Self-Correction via Multi-Agent Reasoning
2025
Long-VITA: Scaling Large Multi-modal Models to 1 Million Tokens with Leading Short-Context Accuracy
2025
Topics
Training Techniques
Fine-Tuning
Video-Language
Visual QA & Reasoning
Benchmarks
Perception
Reinforcement Learning
In-Context Learning
Efficiency
Model Architecture