F-5-TTS
Emerging5papers using it
2025first seen
The 'F-5-TTS' dataset/benchmark is used to evaluate text-to-speech systems, specifically in terms of their performance in cross-lingual voice cloning and audio quality.
Papers using F-5-TTS (5)
- Joint Residual Reweighting for Classifier Free Guidance in Flow-Matching Zero-Shot TTSFlowTTS-GRPO: Online Reinforcement Learning with Multi-Objective Reward Optimization for Flow-Matching Based Text-to-SpeechPFluxTTS: Hybrid Flow-Matching TTS with Robust Cross-Lingual Voice Cloning and Inference-Time Model FusionDiTReducio: A Training-Free Acceleration for DiT-Based TTS via Progressive CalibrationPseudo-Autoregressive Neural Codec Language Models for Efficient Zero-Shot Text-to-Speech Synthesis