Awesome Recommender Systems
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hao Ma — most-cited papers & profile · Recommender Systems
← authors
·
overview
Hao Ma
21
papers ·
51
citations ·
21
h-index
University of Electronic Science and Technology of China · Northwestern Polytechnical University · Zhengzhou University of Aeronautics · China Aerodynamics Research and Development Center · Shanghai Institute of Geological Survey
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Coevolving With The Other You: Fine-tuning LLM With Sequential Cooperative Multi-agent Reinforcement Learning
2024 · 37 citations
Coevolving with the Other You: Fine-Tuning LLM with Sequential Cooperative Multi-Agent Reinforcement Learning
2024 · 4 citations
Vib2Mol: from vibrational spectra to molecular structures-a unified deep learning framework
2025 · 2 citations
Data-Efficient Online Learning of Ball Placement in Robot Table Tennis
2023 · 2 citations
The Perfect Blend: Redefining RLHF with Mixture of Judges
2024 · 2 citations
Efficient Soft Actor-Critic with LLM-Based Action-Level Guidance for Continuous Control
2026
Enhancing Reinforcement Learning Fine-Tuning with an Online Refiner
2026
Efficient Model-Based Reinforcement Learning for Robot Control via Online Optimization
2025
COPO: Consistency-Aware Policy Optimization
2025
COPO: Consistency-Aware Policy Optimization
2025
Constraint-Aware Diffusion Guidance for Robotics: Real-Time Obstacle Avoidance for Autonomous Racing
2025
Vision-Based Generic Potential Function for Policy Alignment in Multi-Agent Reinforcement Learning
2025
Causal Mean Field Multi-Agent Reinforcement Learning
2025
Stochastic Online Optimization For Cyber-physical And Robotic Systems
2024
Effective Long-Context Scaling of Foundation Models
2023
Topics
RLHF & Alignment
Policy Gradient
Control
Multi-Agent
Model-Based RL
Fine-Tuning
In-Context Learning
Training Techniques
Exploration
Manipulation