← all datasets

WavCaps

Emerging
1papers using it
2024first seen

WavCaps is a dataset that contains audio clips paired with corresponding video content, used to evaluate the effectiveness of text-to-audio technology in the context of foley audio dubbing.

Papers using WavCaps (1)

WavCaps β€” datasets β€” reinforcement-learning