Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Mohd Mujtaba Akhtar — most-cited papers & profile · Speech Audio
← authors
·
overview
Mohd Mujtaba Akhtar
11
papers ·
0
citations ·
2
h-index
Veer Bahadur Singh Purvanchal University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Bridging the SEA Gap: An Initial Benchmark for Neural Audio Codec-Synthesized Speech Deepfakes in South-East Asian Languages
2026
Prosody as Supervision: Bridging the Non-Verbal--Verbal for Multilingual Speech Emotion Recognition
2026
Indic-CodecFake meets SATYAM: Towards Detecting Neural Audio Codec Synthesized Speech Deepfakes in Indic Languages
2026
Bridging Attribution and Open-Set Detection using Graph-Augmented Instance Learning in Synthetic Speech
2026
Rethinking Cross-Corpus Speech Emotion Recognition Benchmarking: Are Paralinguistic Pre-Trained Representations Sufficient?
2025
Investigating Polyglot Speech Foundation Models for Learning Collective Emotion from Crowds
2025
Enhancing In-Domain and Out-Domain EmoFake Detection via Cooperative Multilingual Speech Foundation Models
2025
PARROT: Synergizing Mamba and Attention-based SSL Pre-Trained Models via Parallel Branch Hadamard Optimal Transport for Speech Emotion Recognition
2025
Investigating the Reasonable Effectiveness of Speaker Pre-Trained Models and their Synergistic Power for SingMOS Prediction
2025
HYFuse: Aligning Heterogeneous Speech Pre-Trained Representations in Hyperbolic Space for Speech Emotion Recognition
2025
Top co-authors
Arun Balaji Buduru
· 8
Girish
· 6
Swarup Ranjan Behera
· 5
Orchid Chetia Phukan
· 4
Pailla Balakrishna Reddy
· 4
Orchid Chetia Phukan
· 2
Parabattina Bhagath
· 2
Rajesh Sharma
· 2
Farhan Sheth
· 1
Girish
· 1
Girish
· 1
Girish
· 1
Topics
Speech Recognition
Audio Understanding
Audio Generation
Speaker Analysis
Text-to-Speech
Multimodal Audio
Speech Translation
Voice Cloning
Music Generation