Awesome Multimodal
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β authors
Β·
overview
Loading authorβ¦
π€
Ask AI
Yaojie Lu β most-cited papers & profile Β· Multimodal
β authors
Β·
overview
Yaojie Lu
17
papers Β·
0
citations Β·
4
h-index
Google Scholar β
Semantic Scholar β
OpenAlex β
Most-cited papers
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation
2026
PraMem: Practice-derived Experiential Memory for Long-horizon Behavior Prediction
2026
DocOps: A Verifiable Benchmark for Autonomous Agents in Complex Document Operations
2026
Decoupling Reasoning and Confidence: Resurrecting Calibration in Reinforcement Learning from Verifiable Rewards
2026
Inside the Skill Market: From Software Engineering Activities to Reusable Agent Skills
2026
Before the Action: Benchmarking LLMs on Prospective Hypothesis Discovery
2026
P^2O: Joint Policy and Prompt Optimization
2026
P^2O: Joint Policy and Prompt Optimization
2026
MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End Reinforcement Learning
2025
RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback
2025
Across Programming Language Silos: A Study on Cross-Lingual Retrieval-augmented Code Generation
2025
DeepRAG: Thinking to Retrieval Step by Step for Large Language Models
2025
RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback
2025
A Unified View of Delta Parameter Editing in Post-Trained Large-Scale Models
2024
Aligning Large Language Models via Self-Steering Optimization
2024
Topics
Evaluation
Reinforcement Learning
Training Techniques
Benchmarks
Code Agents
Safety & Alignment
Fine-Tuning
Model-Based RL
RLHF & Alignment
Policy Gradient