Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Zhiyong Wu — most-cited papers & profile · Multimodal
← authors
·
overview
Zhiyong Wu
18
papers ·
108
citations ·
28
h-index
College of Medical Sciences · Universidad del Noreste
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
NEXT: A Neural Network Framework for Next POI Recommendation
2017 · 86 citations
TDAG: A Multi-agent Framework Based On Dynamic Task Decomposition And Agent Generation
2024 · 19 citations
Enhancing Generalization of Speech Large Language Models with Multi-Task Behavior Imitation and Speech-Text Interleaving
2025 · 2 citations
Genius: A Generalizable and Purely Unsupervised Self-Training Framework For Advanced Reasoning
2025 · 1 citations
Workflow-GYM: Towards Long-Horizon Evaluation of Computer-use Agentic tasks in Real-World Professional Fields
2026
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models
2026
E2E-VGuard: Adversarial Prevention for Production LLM-based End-To-End Speech Synthesis
2025
LSZone: A Lightweight Spatial Information Modeling Architecture for Real-time In-car Multi-zone Speech Separation
2025
OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows
2025
Towards Hallucination-Free Music: A Reinforcement Learning Preference Optimization Framework for Reliable Song Generation
2025
LightGrad: Lightweight Diffusion Probabilistic Model for Text-to-Speech
2023
Improving Mandarin Prosodic Structure Prediction with Multi-level Contextual Information
2023
RFWave: Multi-band Rectified Flow for Audio Waveform Reconstruction
2024
VoxInstruct: Expressive Human Instruction-to-Speech Generation with Unified Multilingual Codec Language Modelling
2024
Top co-authors
Ben Kao
· 1
Bowen Yang
· 1
Fangzhi Xu
· 1
Hang Yan
· 1
Jiahui Gao
· 1
Jialin Cao
· 1
Jianbing Zhang
· 1
Jingyang Gong
· 1
Kaiming Jin
· 1
Kanzhi Cheng
· 1
Liheng Chen
· 1
Lingpeng Kong
· 1
Topics
Text-to-Speech
Speech Recognition
Audio Generation
Multi-Agent
Evaluation
Benchmarks
Tool Use
Audio Understanding
Speech Translation
Music Generation