Visqol V3: An Open Source Production Ready Objective Speech And Audio Metric
2020 Β· Michael Chinen, Felicia S. C. Lim, Jan Skoglund, et al.
Abstract
Estimation of perceptual quality in audio and speech is possible using a variety of methods. The combined v3 release of ViSQOL and ViSQOLAudio (for speech and audio, respectively,) provides improvements upon previous versions, in terms of both design and usage. As an open source C++ library or binary with permissive licensing, ViSQOL can now be deployed beyond the research context into production usage. The feedback from internal production teams at Google has helped to improve this new release, and serves to show cases where it is most applicable, as well as to highlight limitations. The new model is benchmarked against real-world data for evaluation purposes. The trends and direction of future work is discussed.
Authors
(none)
Tags
Stats
Related papers
- VERSA: A Versatile Evaluation Toolkit For Speech, Audio, And Music (2024)15.28
- Torchaudio-squim: Reference-less Speech Quality And Intelligibility Measures In Torchaudio (2023)0.00
- MMMOS: Multi-domain Multi-axis Audio Quality Assessment (2025)0.00
- SCOREQ: Speech Quality Assessment With Contrastive Regression (2024)8.09
- Openace: An Open Benchmark For Evaluating Audio Coding Performance (2024)2.16
- Squid: Measuring Speech Naturalness In Many Languages (2022)9.41
- Non-intrusive Speech Quality Assessment Using Neural Networks (2019)13.74
- Speechllm-as-judges: Towards General And Interpretable Speech Quality Evaluation (2025)2.60