Awesome Speech Audio
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Nithin Rao Koluguri — most-cited papers & profile · Speech Audio
← authors
·
overview
Nithin Rao Koluguri
17
papers ·
60
citations ·
8
h-index
Nvidia (United States)
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
SpeakerNet: 1D Depth-wise Separable Convolutional Network for Text-Independent Speaker Recognition and Verification
2020 · 29 citations
BESTOW: Efficient and Streamable Speech Language Model with the Best of Two Worlds in GPT and T5
2024 · 6 citations
Property-Aware Multi-Speaker Data Simulation: A Probabilistic Modelling Technique for Synthetic Data Generation
2023 · 5 citations
The CHiME-7 Challenge: System Description and Performance of NeMo Team's DASR System
2023 · 5 citations
TitaNet: Neural Model for speaker representation with 1D Depth-wise separable convolutions and global context
2021 · 4 citations
Multi-scale Speaker Diarization with Dynamic Scale Weighting
2022 · 4 citations
A Compact End-to-End Model with Local and Global Context for Spoken Language Identification
2022 · 4 citations
Discrete Audio Representation as an Alternative to Mel-Spectrograms for Speaker and Speech Recognition
2023 · 2 citations
Less is More: Accurate Speech Recognition & Translation without Web-Scale Data
2024 · 1 citations
Speaker Targeting via Self-Speaker Adaptation for Multi-talker ASR
2025
Granary: Speech Recognition and Translation Dataset in 25 European Languages
2025
NEST: Self-supervised Fast Conformer as All-purpose Seasoning to Speech Processing Tasks
2024
Investigating End-to-End ASR Architectures for Long Form Audio Transcription
2023
Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs
2024
Codec-ASR: Training Performant Automatic Speech Recognition Systems with Discrete Speech Representations
2024
Top co-authors
Boris Ginsburg
· 16
Jagadeesh Balam
· 14
Kunal Dhawan
· 9
He Huang
· 7
Krishna C. Puvvada
· 6
Ivan Medennikov
· 3
Oleksii Hrinchuk
· 3
Taejin Park
· 3
Tae Jin Park
· 3
Vitaly Lavrukhin
· 3
Ante Juki\'c
· 2
Ante Jukić
· 2
Topics
Speech Recognition
Speaker Analysis
Audio Understanding
Speech Translation
Text-to-Speech
Audio Generation
Music Generation
Multimodal Audio
Speech Enhancement