Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
(RLVR)
loadingβ¦
π€
Ask AI
Awesome (RLVR) β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
(RLVR)
5 papers tagged (RLVR) β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
5 papers Β· trending (default)
numbers = π₯ heat
RLBFF: Binary Flexible Feedback to bridge between Human Feedback & Verifiable Rewards
(2025)
Zhilin Wang et al.
1.28
Thinking-Free Policy Initialization Makes Distilled Reasoning Models More Effective and Efficient Reasoners
(2025)
Xin Xu et al.
1.28
The Best of N Worlds: Aligning Reinforcement Learning with Best-of-N Sampling via max@k Optimisation
(2025)
Farid Bagirov et al.
1.28
Long-horizon Reasoning Agent for Olympiad-Level Mathematical Problem Solving
(2025)
Songyang Gao et al.
1.28
OPV: Outcome-based Process Verifier for Efficient Long Chain-of-Thought Verification
(2025)
Zijian Wu et al.
1.28