Awesome Reinforcement Learning
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β authors
Β·
overview
Loading authorβ¦
π€
Ask AI
Maksym Andriushchenko β most-cited papers & profile Β· Reinforcement Learning
β authors
Β·
overview
Maksym Andriushchenko
15
papers Β·
96
citations Β·
0
h-index
Google Scholar β
Semantic Scholar β
OpenAlex β
Most-cited papers
Logit Pairing Methods Can Fool Gradient-Based Attacks
2018 Β· 46 citations
Is In-context Learning Sufficient For Instruction Following In Llms?
2024 Β· 22 citations
JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models
2024 Β· 15 citations
Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks
2024 Β· 9 citations
HalluHard: A Hard Multi-Turn Hallucination Benchmark
2026
Skill-Inject: Measuring Agent Vulnerability to Skill File Attacks
2026
Monitoring Decomposition Attacks in LLMs with Lightweight Sequential Monitors
2025
Capability-Based Scaling Laws for LLM Red-Teaming
2025
Strategic Dishonesty Can Undermine AI Safety Evaluations of Frontier LLM
2025
Layer-wise Linear Mode Connectivity
2023
Topics
Adversarial ML
Safety & Alignment
Evaluation
LLM Security
Vulnerability Detection
Threat Intelligence
In-Context Learning
RAG
Agentic
Federated Learning