Awesome Robotics
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β all topics
overview
Sound
loadingβ¦
π€
Ask AI
Awesome Sound β curated papers, datasets & benchmarks Β· Awesome Robotics
β all topics
overview
Sound
18 papers tagged Sound β re-sort below
Papers
π₯ Trending (default)
π Most cited
π Newest first
π€ A β Z by title
18 papers Β· trending (default)
numbers = π₯ heat
Phone Segmentation and Recognition through Phonological Activation Mapping
(2026)
Shikhar Bharadwaj et al.
5.88
Cover First, Disagree Softly: Rethinking Mismatch-First Active Learning for Frame-Level Audio Classification
(2026)
Shiqi Zhang et al.
4.39
Addressing Limited Data in Auditory Attention Decoding with Diffusion Generative Models
(2026)
David Rannaleet et al.
4.39
What the Waveform Knows: Transparent-first Speech and Audio Intelligence with Caption Studio
(2026)
Cheng Siong Chin et al.
4.39
A 3D-Printable Dataset for Fair Testing and Comparisons of Tactile Sensors
(2026)
Dexter R. Shepherd et al.
4.33
Local Multimodal Music Alignment from Global Supervision
(2026)
Irmak Bukey et al.
2.00
An Explainable FFT-Based Spatial-Frequency Fusion Framework for Deepfake Detection
(2026)
Pamela Kirui et al.
2.00
Physiological Signals as a Forensic Modality for Talking-Face Deepfake Detection
(2026)
Othmane Harraq et al.
2.00
Probing Speaker Identity Sensitivity in Audio Deepfake Detectors
(2026)
Daniyal Kabir Dar et al.
2.00
MemNMF: Memory-Augmented NMF on LPC Spectra for Anomalous Sound Detection
(2026)
Phurich Saengthong et al.
2.00
Synthetic Speech, Real Signal: Paralinguistic Preservation and Cross-Lingual Augmentation via Voice Cloning
(2026)
Roseline Polle et al.
2.00
IQ-JEPA: A Joint-Embedding Predictive Architecture with a Hermitian Vision Transformer for Sound Speed and Attenuation Estimation from Ultrasound IQ Data
(2026)
Masashi Sode et al.
2.00
Reflector: Arrangement-Aware Harmonic Retrieval for Sample-Based Composition
(2026)
Austin Rockman
2.00
Phylogenetic signal in marine mammal and bird vocalizations captured by audio foundation models: the limited benefit of domain-specific pretraining
(2026)
V\'ictor Rinc\'on Yepes
2.00
Interpretable EEG biomarkers with bag-of-waves: Spatial and temporal waveform dictionaries for low-data regimes
(2026)
Athanasios Papastathopoulos-Katsaros et al.
2.00
PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers
(2026)
Xiangyue Zhang et al.
1.89
HD3C: Efficient Medical Data Classification for Edge Devices
(2025)
Jianglan Wei et al.
1.44
Meta-Learning Approaches for Speaker-Dependent Voice Fatigue Models
(2025)
Roseline Polle et al.
1.22