Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
supervision
loadingβ¦
π€
Ask AI
Awesome supervision β curated papers, datasets & benchmarks Β· Awesome Large Language Models
β all topics
overview
supervision
21 papers tagged supervision β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
21 papers Β· trending (default)
numbers = π₯ heat
Privileged Information Distillation for Language Models
(2026)
Emiliano Penaloza et al.
1.94
Sci-CoE: Co-evolving Scientific Reasoning LLMs via Geometric Consensus with Sparse Supervision
(2026)
Xiaohan He et al.
1.94
Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision
(2026)
Yinghui He et al.
1.94
Dual-View Training for Instruction-Following Information Retrieval
(2026)
Qingcheng Zeng et al.
1.94
Delta-Adapter: Scalable Exemplar-Based Image Editing with Single-Pair Supervision
(2026)
Jiacheng Chen et al.
1.94
Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why
(2026)
Mohammadreza Armandpour et al.
1.94
SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer
(2026)
Haoyi Zhu et al.
1.94
LLaVA-OneVision-2: Towards Next-Generation Perceptual Intelligence
(2026)
Xiang An et al.
1.94
Not All Disagreement Is Learnable: Token Teachability in On-Policy Distillation
(2026)
Yuanyi Wang et al.
1.94
HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers
(2026)
Guozhen Zhang et al.
1.94
Beyond Scalar Distances: Semantic Attribute Gradients from Frozen MLLMs for Visual Embeddings
(2026)
Shubhang Bhatnagar et al.
1.94
TIP: Token Importance in On-Policy Distillation
(2026)
Yuanda Xu et al.
1.83
UI-Voyager: A Self-Evolving GUI Agent Learning via Failed Experience
(2026)
Zichuan Lin et al.
1.78
VIKI-R: Coordinating Embodied Multi-Agent Cooperation via Reinforcement Learning
(2025)
Li Kang et al.
1.28
AWorld: Dynamic Multi-Agent System with Stable Maneuvering for Robust GAIA Problem Solving
(2025)
Zhitian Xie et al.
1.28
Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation
(2025)
Junyan Ye et al.
1.28
Visual Representation Alignment for Multimodal Large Language Models
(2025)
Heeji Yoon et al.
1.28
V-Thinker: Interactive Thinking with Images
(2025)
Runqi Qiao et al.
1.28
Hybrid Attribution Priors for Explainable and Robust Model Training
(2025)
Zhuoran Zhang et al.
1.28
Cross-Lingual Supervision improves Large Language Models Pre-training
(2023)
Andrea Schioppa et al.
β
MATATA: a weak-supervised MAthematical Tool-Assisted reasoning for Tabular Applications
(2024)
Vishnou Vinayagame et al.
β