Expresso
Emerging4papers using it
2023first seen
The Expresso Dataset [paper] [demo samples] [Original repository] Introduction The Expresso dataset is a high-quality (48kHz) expressive speech dataset that includes both expressively rendered read speech (8 styles, in mono wav format) and improvised dialogues (26 styles, in stereo wav format). The dataset includes 4 s
Papers using Expresso (4)
- TASLA: Text-Aligned Speech Tokens with Multiple Layer-AggregationNonverbalTTS: A Public English Corpus of Text-Aligned Nonverbal Vocalizations with Emotion Annotations for Text-to-SpeechAligning Paralinguistic Understanding and Generation in Speech LLMs via Multi-Task Reinforcement LearningEXPRESSO: A Benchmark and Analysis of Discrete Expressive Speech
Resynthesis