How-2
Emerging8papers using it
2021first seen
The 'How-2' dataset is a benchmark used to evaluate the performance of models in generating human-like text summaries from spoken content.
Papers using How-2 (8)
- Leveraging Large Text Corpora for End-to-End Speech SummarizationAVFormer: Injecting Vision into Frozen Speech Models for Zero-Shot
AV-ASRXNOR-FORMER: Learning Accurate Approximations in Long Speech
TransformersAttention-based Multi-hypothesis Fusion for Speech SummarizationAVATAR: Unconstrained Audiovisual Speech RecognitionSpeech Summarization using Restricted Self-AttentionBASS: Block-wise Adaptation for Speech SummarizationAn End-to-End Speech Summarization Using Large Language Model