CVSS-C
Emerging8papers using it
2022first seen
The CVSS-C dataset/benchmark contains a collection of speech-to-speech translation data used to evaluate the performance of end-to-end speech-to-speech translation systems.
Papers using CVSS-C (8)
- Leveraging Audio-LLMs to Filter Speech-to-Speech Training DataFrom Flat Language Labels to Typological Priors: Structured Language Conditioning for Multilingual Speech-to-Speech TranslationRosettaSpeech: Zero-Shot Speech-to-Speech Translation without Parallel SpeechSLM-S2ST: A multimodal language model for direct speech-to-speech translationLeveraging unsupervised and weakly-supervised data to improve direct
speech-to-speech translationTextless Direct Speech-to-Speech Translation with Discrete Speech
RepresentationTextless Streaming Speech-to-Speech Translation using Semantic Speech
TokensPhonology-Guided Speech-to-Speech Translation for African Languages