Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hao Hu — most-cited papers & profile · Multimodal
← authors
·
overview
Hao Hu
36
papers ·
123
citations ·
14
h-index
Shanghai Jiao Tong University · Wuchang University of Technology · Beijing Advanced Sciences and Innovation Center · Shanghai Ocean University · China Agricultural University · University of Science and Technology Beijing
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Reason For Future, Act For Now: A Principled Framework For Autonomous LLM Agents With Provable Sample Efficiency
2023 · 49 citations
Generalizable Episodic Memory for Deep Reinforcement Learning
2021 · 20 citations
Kimi k1.5: Scaling Reinforcement Learning with LLMs
2025 · 11 citations
Offline Reinforcement Learning with Value-based Episodic Memory
2021 · 10 citations
MetaCURE: Meta Reinforcement Learning with Empowerment-Driven Exploration
2020 · 9 citations
On the Role of Discount Factor in Offline Reinforcement Learning
2022 · 8 citations
Integrating One-shot View Planning With A Single Next-best View Via Long-tail Multiview Sampling
2023 · 7 citations
Global versus Localized Generative Adversarial Nets
2017 · 3 citations
Kimi K3: Open Frontier Intelligence
2026
CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents
2026
Enhancing Cloud Network Resilience via a Robust LLM-Empowered Multi-Agent Reinforcement Learning Framework
2026
T-GMP: Terrain-conditioned Generative Motion Priors for Versatile and Natural Humanoid Locomotion
2026
GuideWalk: Learning Unified Autonomous Navigation and Locomotion for Humanoid Robots across Versatile Terrains
2026
Kimi K3: Open Frontier Intelligence
2026
A Unified Knowledge Embedded Reinforcement Learning-based Framework for Generalized Capacitated Vehicle Routing Problems
2026
Topics
Offline RL
Exploration
Model-Based RL
cs.AI
Policy Gradient
Value-Based
cs.LG
Multi-Agent
Meta-RL
Safe RL