Awesome AI Agents
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Beidi Chen — most-cited papers & profile · AI Agents
← authors
·
overview
Beidi Chen
18
papers ·
1740
citations ·
1
h-index
Peking University · Peking University Third Hospital
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Efficient Streaming Language Models With Attention Sinks
2023 · 1654 citations
Prosperity before Collapse: How Far Can Off-Policy RL Reach with Stale Data on LLMs?
2025 · 29 citations
LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding
2024 · 26 citations
Prompt-prompted Adaptive Structured Pruning For Efficient LLM Generation
2024 · 21 citations
Speculative Prefill: Turbocharging TTFT With Lightweight And Training-free Token Importance Estimation
2025 · 10 citations
When "Correct" Is Not Safe: Can We Trust Functionally Correct Patches Generated by Code Agents?
2025
APE: Faster and Longer Context-Augmented Generation via Adaptive Parallel Encoding
2025
HeadInfer: Memory-Efficient LLM Inference by Head-wise Offloading
2025
When "Correct" Is Not Safe: Can We Trust Functionally Correct Patches Generated by Code Agents?
2025
On The Surprising Effectiveness Of Attention Transfer For Vision Transformers
2024
H_2O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models
2023
Deja Vu: Contextual Sparsity for Efficient LLMs at Inference Time
2023
GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection
2024
Megalodon: Efficient LLM Pretraining and Inference with Unlimited Context Length
2024
Nearest Neighbor Speculative Decoding for LLM Generation and Attribution
2024
Topics
Efficiency
Model Architecture
Training Techniques
In-Context Learning
Code Generation
RAG
Program Repair
Bug Detection
Software Engineering
attention