Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Siyuan Liang — most-cited papers & profile · Multimodal
← authors
·
overview
Siyuan Liang
30
papers ·
36
citations ·
0
h-index
Beijing Institute of Technology · Beijing Electronic Science and Technology Institute · Beijing Research Institute of Mechanical and Electrical Technology
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Visual Adversarial Attack on Vision-Language Models for Autonomous Driving
2024 · 1 citations
Explaining multimodal LLMs via intra-modal token interactions
2025
Universal Camouflage Attack On Vision-language Models For Autonomous Driving
2025
Bridging The Task Gap: Multi-task Adversarial Transferability In CLIP And Its Derivatives
2025
Natural Reflection Backdoor Attack On Vision Language Model For Autonomous Driving
2025
Roboview-bias: Benchmarking Visual Bias In Embodied Agents For Robotic Manipulation
2025
Poison Once, Control Anywhere: Clean-text Visual Backdoors In Vlm-based Mobile Agents
2025
SRD: Reinforcement-learned Semantic Perturbation For Backdoor Defense In Vlms
2025
Robust Anti-backdoor Instruction Tuning In Lvlms
2025
Manipulating Multimodal Agents via Cross-Modal Prompt Injection
2025
Top co-authors
Aishan Liu
· 2
Shengshan Hu
· 2
Tianyuan Zhang
· 2
Xianglong Liu
· 2
Boyi Jia
· 1
Cheng Qian
· 1
Dehong Kong
· 1
Enguang Liu
· 1
Hongling Zheng
· 1
Jiawei Liang
· 1
Koushik Howlader
· 1
Kuanrong Liu
· 1
Topics
Vision-Language Models
Video-Language
Visual QA & Reasoning
Embodied & Agents
Instruction Tuning
Image-Text Retrieval
Benchmarks