Awesome AI for Code
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Yining Ma — most-cited papers & profile · AI for Code
← authors
·
overview
Yining Ma
24
papers ·
279
citations ·
5
h-index
China Southern Power Grid (China)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Deep Reinforcement Learning for Solving the Heterogeneous Capacitated Vehicle Routing Problem
2021 · 187 citations
Deep Reinforcement Learning for Dynamic Algorithm Selection: A Proof-of-Principle Study on Differential Evolution
2024 · 34 citations
Fault-Tolerant Federated Reinforcement Learning with Theoretical Guarantee
2021 · 16 citations
Auto-configuring Exploration-Exploitation Tradeoff in Evolutionary Computation via Deep Reinforcement Learning
2024 · 9 citations
MetaBox: A Benchmark Platform for Meta-Black-Box Optimization with Reinforcement Learning
2023 · 5 citations
RL4CO: an Extensive Reinforcement Learning for Combinatorial Optimization Benchmark
2023 · 4 citations
Evolving Testing Scenario Generation Method and Intelligence Evaluation Framework for Automated Vehicles
2023 · 2 citations
Diversity Optimization for Travelling Salesman Problem via Deep Reinforcement Learning
2025 · 1 citations
Fault-Tolerant Federated Reinforcement Learning with Theoretical Guarantee
2021 · 1 citations
FedHQL: Federated Heterogeneous Q-Learning
2023 · 1 citations
FedHQL: Federated Heterogeneous Q-Learning
2023 · 1 citations
Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning
2026
Learning-guided Prioritized Planning for Lifelong Multi-Agent Path Finding in Warehouse Automation
2026
AlphaOPT: Formulating Optimization Programs with Self-Improving LLM Experience Library
2025
Decision-making with Speculative Opponent Models
2022
Topics
Policy Gradient
Multi-Agent
Model-Based RL
Meta-RL
Exploration
Safe RL
cs.AI
Federated Learning
Game AI
Value-Based