Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Xiangtai Li — most-cited papers & profile · Large Language Models
← authors
·
overview
Xiangtai Li
11
papers ·
186
citations ·
30
h-index
Nanyang Technological University · ETH Zurich
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
TransVOD: End-to-End Video Object Detection with Spatial-Temporal Transformers
2022 · 186 citations
Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diffusion Language Models
2026
Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future
2025
PairUni: Pairwise Training for Unified Multimodal Language Models
2025
Denseworld-1m: Towards Detailed Dense Grounded Caption In The Real World
2025
The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer
2025
PVUW 2025 Challenge Report: Advances in Pixel-level Understanding of Complex Videos in the Wild
2025
UMC: Unified Resilient Controller for Legged Robots with Joint Malfunctions
2025
Are They the Same? Exploring Visual Correspondence Shortcomings of Multimodal LLMs
2025
PolyphonicFormer: Unified Query Learning for Depth-aware Video Panoptic Segmentation
2021
Topics
Vision-Language Models
Benchmarks
Video Understanding
Control
Video-Language
Segmentation
Tracking
Code Models
Code Generation
Testing