Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
collapse
loadingβ¦
π€
Ask AI
Awesome collapse β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
collapse
18 papers tagged collapse β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
18 papers Β· trending (default)
numbers = π₯ heat
RAGEN-2: Reasoning Collapse in Agentic RL
(2026)
Zihan Wang et al.
1.94
GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification
(2026)
Wangjie Gan et al.
1.94
Delta Attention Residuals
(2026)
Cheng Luo et al.
1.94
STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability
(2026)
Haipeng Luo et al.
1.94
EvoEmbedding: Evolvable Representations for Long-Context Retrieval and Agentic Memory
(2026)
Chang Nie et al.
1.94
DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimization
(2026)
Jian Mu et al.
1.89
Stabilizing Reinforcement Learning for Diffusion Language Models
(2026)
Jianyuan Zhong et al.
1.78
Online Causal Kalman Filtering for Stable and Effective Policy Optimization
(2026)
Shuo He et al.
1.72
ARM: Adaptive Reasoning Model
(2025)
Siye Wu et al.
1.28
CARFT: Boosting LLM Reasoning via Contrastive Learning with Annotated Chain-of-Thought-based Reinforced Fine-Tuning
(2025)
Wenqiao Zhu et al.
1.28
RewardDance: Reward Scaling in Visual Generation
(2025)
Jie Wu et al.
1.28
Benefits and Pitfalls of Reinforcement Learning for Language Model Planning: A Theoretical Perspective
(2025)
Siwei Wang et al.
1.28
Epistemic Diversity and Knowledge Collapse in Large Language Models
(2025)
Dustin Wright et al.
1.28
Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models
(2025)
Qizheng Zhang et al.
1.28
Detecting Data Contamination from Reinforcement Learning Post-training for Large Language Models
(2025)
Yongding Tao et al.
1.28
UI-Ins: Enhancing GUI Grounding with Multi-Perspective Instruction-as-Reasoning
(2025)
Liangyu Chen et al.
1.28
Hyper-Connections
(2024)
Defa Zhu et al.
β
How to Synthesize Text Data without Model Collapse?
(2024)
Xuekai Zhu et al.
β