KeSpeech
Emerging5papers using it
2024first seen
This dataset only contains test data, which is integrated into UltraEval-Audio(https://github.com/OpenBMB/UltraEval-Audio) framework. python audio_evals/main.py --dataset KeSpeech --model gpt4o_audio 🚀超凡体验,尽在UltraEval-Audio🚀 UltraEval-Audio——全球首个同时支持语音理解和语音生成评估的开源框架,专为语音大模型评估打造,集合了34项权威Benchmark,覆盖语音、声音、医疗及音乐四大领域,支持十
Papers using KeSpeech (5)
- Leveraging LLM and Self-Supervised Training Models for Speech Recognition in Chinese Dialects: A Comparative AnalysisDialect Identification Using Resource-Efficient Fine-Tuning ApproachesM2R-Whisper: Multi-stage and Multi-scale Retrieval Augmentation for
Enhancing WhisperQifusion-Net: Layer-adapted Stream/Non-stream Model for End-to-End
Multi-Accent Speech RecognitionSpeaker-Smoothed kNN Speaker Adaptation for End-to-End ASR