Awesome AI Agents
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yuan Wang — most-cited papers & profile · AI Agents
← authors
·
overview
Yuan Wang
22
papers ·
260
citations ·
0
h-index
Inner Mongolia Chifeng Forestry Science Research Institute
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
2025 · 206 citations
Controllable Multi-Objective Re-ranking with Policy Hypernetworks
2023 · 32 citations
Golden Cudgel Network for Real-Time Semantic Segmentation
2025 · 18 citations
Data-Efficient RLVR via Off-Policy Influence Guidance
2025 · 3 citations
G-flocking: Flocking Model Optimization based on Genetic Framework
2019 · 1 citations
Thinking With Deltas: Incentivizing Reinforcement Learning Via Differential Visual Reasoning Policy
2026
Graph Self-Supervised Learning via Learnable View Augmentation for Recommender System
2026
Thinking in Text and Images: Interleaved Vision--Language Reasoning Traces for Long-Horizon Robot Manipulation
2026
Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation
2026
Photonics-Enhanced Graph Convolutional Networks
2025
From Imperative to Declarative: Towards LLM-friendly OS Interfaces for Boosted Computer-Use Agents
2025
Improving Learning of New Diseases through Knowledge-Enhanced Initialization for Federated Adapter Tuning
2025
Atom-Searcher: Enhancing Agentic Deep Research via Fine-Grained Atomic Thought Reward
2025
GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
2025
Med-U1: Incentivizing Unified Medical Reasoning in LLMs via Large-scale Reinforcement Learning
2025
Topics
Vision-Language Models
Visual QA & Reasoning
Model-Based RL
RLHF & Alignment
Policy Gradient
Control
Benchmarks
Multi-Agent
Value-Based
Manipulation