CORAAL
Emerging6papers using it
2021first seen
Dataset Card for CORAAL Dataset Summary This dataset comprises audio files, text files, and audio segments sourced from the Corpus of Regional African American Language (CORAAL). CORAAL is a subset of the Online Resources for African American Language (ORAAL) project, initiated by a team of linguistics researchers at t
Papers using CORAAL (6)
- Gumbel-BEARD: Automatic Layer Selection for Self-Supervised Adaptation of Whisper in Low-Resource DomainsWhisper-CD: Accurate Long-Form Speech Recognition using Multi-Negative Contrastive DecodingCORAA: a large corpus of spontaneous and prepared speech manually
validated for speech recognition in Brazilian PortugueseContinual Learning for End-to-End ASR by Averaging Domain ExpertsInvestigating End-to-End ASR Architectures for Long Form Audio
TranscriptionImproving Speech Recognition for African American English With Audio
Classification