Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Guo Chen — most-cited papers & profile · Large Language Models
← authors
·
overview
Guo Chen
20
papers ·
74
citations ·
14
h-index
Peking University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation
2023 · 32 citations
Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding
2024 · 19 citations
SPMamba: State-space model is all you need in speech separation
2024 · 8 citations
MRSN: Multi-Relation Support Network for Video Action Detection
2023 · 7 citations
AVSegFormer: Audio-Visual Segmentation with Transformer
2023 · 3 citations
Champion Solution for the WSDM2023 Toloka VQA Challenge
2023 · 2 citations
Champion Solution for the WSDM2023 Toloka VQA Challenge
2023 · 2 citations
Word-Free Spoken Language Understanding for Mandarin-Chinese
2021 · 1 citations
TactiDex: A Real-World Tactile-Guided Benchmark for Human-Like Dexterous Manipulation
2026
Fed-GAME: Personalized Federated Learning with Graph Attention Mixture-of-Experts For Time-Series Forecasting
2026
Enhancing Spectrogram Realism in Singing Voice Synthesis via Explicit Bandwidth Extension Prior to Vocoder
2025
Egoexobench: A Benchmark For First- And Third-person View Video Understanding In Mllms
2025
Videoitg: Multimodal Video Understanding With Instructed Temporal Grounding
2025
Time-Frequency-Based Attention Cache Memory Model for Real-Time Speech Separation
2025
Text2Schema: Filling the Gap in Designing Database Table Structures based on Natural Language
2025
Top co-authors
Dayiheng Liu
· 1
De-An Huang
· 1
Guilin Liu
· 1
Hongxu Yin
· 1
Jan Kautz
· 1
Jialin Liu
· 1
Jindong Jiang
· 1
Junhao Zheng
· 1
Kexin Yang
· 1
Kurt Keutzer
· 1
Linfeng Zhang
· 1
Muyang Li
· 1
Topics
Speech Recognition
Audio Understanding
Video-Language
Visual QA & Reasoning
Speech Translation
Speech Enhancement
Video Understanding
Vision-Language Models
Object Detection
Manipulation