GigaSpeech
Canonical13papers using it
2022first seen
Dataset Card for Gigaspeech Dataset Description GigaSpeech is an evolving, multi-domain English speech recognition corpus with 10,000 hours of high quality labeled audio suitable for supervised training. The transcribed audio data is collected from audiobooks, podcasts and YouTube, covering both read and spontaneous sp
Papers using GigaSpeech (13)
- Hybrid Decoding: Rapid Pass and Selective Detailed Correction for Sequence ModelsExploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition SystemsCR-CTC: Consistency regularization on CTC for improved speech
recognitionCommunication-Efficient Personalized Federated Learning for
Speech-to-Text TasksGigaST: A 10,000-hour Pseudo Speech Translation CorpusEnd-to-end contextual asr based on posterior distribution adaptation for
hybrid ctc/attention systemDistillW2V2: A Small and Streaming Wav2vec 2.0 Based ASR ModelLongFNT: Long-form Speech Recognition with Factorized Neural TransducerConnecting Speech Encoder and Large Language Model for ASRUpdated Corpora and Benchmarks for Long-Form Speech RecognitionImproving Automatic Speech Recognition with Decoder-Centric
Regularisation in Encoder-Decoder ModelsLate fusion ensembles for speech recognition on diverse input audio
representationsAdvanced Long-Content Speech Recognition With Factorized Neural
Transducer