Awesome Reinforcement Learning
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β authors
Β·
overview
Loading authorβ¦
π€
Ask AI
Zonghao Ying β most-cited papers & profile Β· Reinforcement Learning
β authors
Β·
overview
Zonghao Ying
22
papers Β·
2
citations Β·
0
h-index
Beijing Academy of Artificial Intelligence Β· Beihang University
Google Scholar β
Semantic Scholar β
OpenAlex β
Most-cited papers
Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt
2024 Β· 1 citations
SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models
2024 Β· 1 citations
Technical Report on the CVPR 2026@AdvML Workshop Challenge
2026
DMN: A Compositional Framework for Jailbreaking Multimodal LLMs with Multi-Image Inputs
2026
SafeHarbor: Hierarchical Memory-Augmented Guardrail for LLM Agent Safety
2026
SafeHarbor: Defining Precise Decision Boundaries via Hierarchical Memory-Augmented Guardrail for LLM Agent Safety
2026
SafeHarbor: Hierarchical Memory-Augmented Guardrail for LLM Agent Safety
2026
DMN: A Compositional Framework for Jailbreaking Multimodal LLMs with Multi-Image Inputs
2026
Reading Between the Pixels: An Inscriptive Jailbreak Attack on Text-to-Image Models
2026
AgentVisor: Defending LLM Agents Against Prompt Injection via Semantic Virtualization
2026
Uncovering Security Threats and Architecting Defenses in Autonomous Agents: A Case Study of OpenClaw
2026
DIVER: Dynamic Iterative Visual Evidence Reasoning for Multimodal Fake News Detection
2026
Sequential Comics For Jailbreaking Multimodal Large Language Models Via Structured Visual Storytelling
2025
PRISM: Programmatic Reasoning with Image Sequence Manipulation for LVLM Jailbreaking
2025
Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025
2025
Topics
Vulnerability Detection
LLM Security
Adversarial ML
Vision-Language Models
Threat Intelligence
Benchmarks
Network Security
Visual QA & Reasoning
Safety
Evaluation