Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Chong Luo — most-cited papers & profile · Multimodal
← authors
·
overview
Chong Luo
21
papers ·
303
citations ·
14
h-index
Sichuan University · China Tobacco · West China Hospital of Sichuan University · Microsoft Research Asia (China) · First Affiliated Hospital Zhejiang University · Zhejiang University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
2024 · 158 citations
A Twofold Siamese Network for Real-Time Object Tracking
2018 · 70 citations
PHASEN: A Phase-and-Harmonics-Aware Speech Enhancement Network
2019 · 26 citations
Peripheral Vision Transformer
2022 · 18 citations
SPM-Tracker: Series-Parallel Matching for Real-Time Visual Object Tracking
2019 · 17 citations
General-Purpose Speech Representation Learning through a Self-Supervised Multi-Granularity Framework
2021 · 8 citations
Towards a Better Match in Siamese Network Based Visual Object Tracker
2018 · 3 citations
Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
2025 · 2 citations
An Anchor-Free Detector for Continuous Speech Keyword Spotting
2022 · 1 citations
Lens: Rethinking Training Efficiency for Foundational Text-to-Image Models
2026
PACR: Progressively Ascending Confidence Reward for LLM Reasoning
2025
JointDiT: Enhancing RGB-Depth Joint Modeling with Diffusion Transformers
2025
Shorten After You're Right: Lazy Length Penalties for Reasoning RL
2025
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning
2025
Zero-Shot Text-to-Speech for Text-Based Insertion in Audio Narration
2021
Topics
Tracking
Object Detection
Model-Based RL
RLHF & Alignment
Image Generation
3D Vision
Speech Enhancement
Speech Recognition
Efficiency
Training Techniques