Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
entropy
loadingβ¦
π€
Ask AI
Awesome entropy β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
entropy
22 papers tagged entropy β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
22 papers Β· trending (default)
numbers = π₯ heat
Reasoning with Exploration: An Entropy Perspective
(2025)
Daixuan Cheng et al.
2.87
Artificial Entanglement in the Fine-Tuning of Large Language Models
(2026)
Min Chen et al.
1.94
The Side Effects of Being Smart: Safety Risks in MLLMs' Multi-Image Reasoning
(2026)
Renmiao Chen et al.
1.94
RAGEN-2: Reasoning Collapse in Agentic RL
(2026)
Zihan Wang et al.
1.94
OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks
(2026)
Wenbo Hu et al.
1.94
GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification
(2026)
Wangjie Gan et al.
1.94
STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability
(2026)
Haipeng Luo et al.
1.94
The First Token Knows: Single-Decode Confidence for Hallucination Detection
(2026)
Mina Gabriel
1.89
Self-Improving Language Models with Bidirectional Evolutionary Search
(2026)
Guowei Xu et al.
1.89
TIP: Token Importance in On-Policy Distillation
(2026)
Yuanda Xu et al.
1.83
The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models
(2025)
Ganqu Cui et al.
1.28
SRFT: A Single-Stage Method with Supervised and Reinforcement Fine-Tuning for Reasoning
(2025)
Yuqian Fu et al.
1.28
The Invisible Leash: Why RLVR May Not Escape Its Origin
(2025)
Fang Wu et al.
1.28
Compressing Chain-of-Thought in LLMs via Step Entropy
(2025)
Zeju Li et al.
1.28
AMFT: Aligning LLM Reasoners by Meta-Learning the Optimal Imitation-Exploration Balance
(2025)
Lixuan He et al.
1.28
Toward Effective Tool-Integrated Reasoning via Self-Evolved Preference Learning
(2025)
Yifei Chen et al.
1.28
ExGRPO: Learning to Reason from Experience
(2025)
Runzhe Zhan et al.
1.28
Detecting Data Contamination from Reinforcement Learning Post-training for Large Language Models
(2025)
Yongding Tao et al.
1.28
EAGER: Entropy-Aware GEneRation for Adaptive Inference-Time Scaling
(2025)
Daniel Scalena et al.
1.28
Video Reasoning without Training
(2025)
Deepak Sridhar et al.
1.28
Multi-hop Reasoning via Early Knowledge Alignment
(2025)
Yuxin Wang et al.
1.28
Confidence Regulation Neurons in Language Models
(2024)
Alessandro Stolfo et al.
β