Awesome Multimodal
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shrikanth Narayanan — most-cited papers & profile · Multimodal
← authors
·
overview
Shrikanth Narayanan
54
papers ·
76
citations ·
76
h-index
University of Southern California · Viterbo University · California Southern University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
PEFT-SER: On the Use of Parameter Efficient Transfer Learning Approaches For Speech Emotion Recognition Using Pre-trained Speech Models
2023 · 25 citations
CapSpeech: Enabling Downstream Applications in Style-Captioned Text-to-Speech
2025 · 15 citations
Attribute Inference Attack of Speech Emotion Recognition in Federated Learning Settings
2021 · 11 citations
The FFSVC 2020 Evaluation Plan
2020 · 8 citations
Large Language Models based ASR Error Correction for Child Conversations
2025 · 5 citations
A long-form single-speaker real-time MRI speech dataset and benchmark
2025 · 3 citations
A study of semi-supervised speaker diarization system using gan mixture model
2019 · 2 citations
Speaker-invariant Affective Representation Learning via Adversarial Training
2019 · 2 citations
TrustSER: On the Trustworthiness of Fine-tuning Pre-trained Speech Embeddings For Speech Emotion Recognition
2023 · 2 citations
FedMultimodal: A Benchmark For Multimodal Federated Learning
2023 · 1 citations
Robust Self Supervised Speech Embeddings for Child-Adult Classification in Interactions involving Children with Autism
2023 · 1 citations
Examining Test-Time Adaptation for Personalized Child Speech Recognition
2024 · 1 citations
ChildVox: A Speech, Audio, and Large Audio-Language Model Benchmark in Understanding and Characterizing Sound across Childhood
2026
Accent Vector: Controllable Accent Manipulation for Multilingual TTS Without Accented Data
2026
Exploring Speech Foundation Models for Speaker Diarization Across Lifespan
2026
Topics
Speech Recognition
Audio Understanding
Speaker Analysis
Speech Translation
eess.AS
cs.SD
cs.CL
cs.AI
Text-to-Speech
Audio Generation