Awesome AI for Code
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Tianyu Pang — most-cited papers & profile · AI for Code
← authors
·
overview
Tianyu Pang
38
papers ·
130
citations ·
0
h-index
University of Hong Kong · South China University of Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Improved Few-shot Jailbreaking Can Circumvent Aligned Language Models And Their Defenses
2024 · 79 citations
Defeating the Training-Inference Mismatch via FP16
2025 · 34 citations
BAFFLE: A Baseline of Backpropagation-Free Federated Learning
2023 · 6 citations
Efficient Diffusion Policies for Offline Reinforcement Learning
2023 · 5 citations
Black-box Detection of Backdoor Attacks with Limited Information and Data
2021 · 3 citations
On Calibrating Diffusion Probabilistic Models
2023 · 2 citations
Denial-of-Service Poisoning Attacks against Large Language Models
2024 · 1 citations
Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models
2026
Active Beyond-Diagonal RIS Empowered Heterogeneous Edge Computing: A Distributional Reinforcement Learning Approach
2026
Rethinking the Divergence Regularization in LLM RL
2026
Rethinking the Trust Region in LLM Reinforcement Learning
2026
VerlTool: Towards Holistic Agentic Reinforcement Learning with Tool Use
2025
Language Models Can Learn from Verbal Feedback Without Scalar Rewards
2025
Why LLM Safety Guardrails Collapse After Fine-tuning: A Similarity Analysis Between Alignment and Fine-tuning Datasets
2025
SkyLadder: Better and Faster Pretraining via Context Window Scheduling
2025
Top co-authors
Haonan Wang
· 2
Xiangxin Zhou
· 2
Biqiang Li
· 1
Bowen Ping
· 1
Chenxin Li
· 1
Haitao Wu
· 1
Hengyu Liu
· 1
Jianghai Chen
· 1
Jiaqi Tang
· 1
Jiarui Yao
· 1
Jun Zhang
· 1
Kaituo Feng
· 1
Topics
Policy Gradient
RLHF & Alignment
Adversarial ML
Model-Based RL
cs.LG
LLM Security
Vulnerability Detection
Threat Intelligence
Safety & Alignment
Efficiency