Awesome Speech Audio
Papers
Topics
Trending
Leaderboards
Authors
Datasets
Learn
Ask AI
Map
Tools
Packs
News
Videos
All sections
Browse
Papers
The full index, filterable and sortable.
Topics
The same papers grouped by subject tag.
Trending
What moved this week, and by how much.
Authors
Who publishes here, and who they publish with.
Map
The collection laid out by embedding similarity.
Compare
Leaderboards
Benchmark tables, with the paper behind each number.
Datasets
The datasets these papers train and evaluate on.
Tools
Code and libraries released alongside the papers.
Read
Learn
Ordered routes from background reading to current work.
Ask AI
Ask a question and get answers cited to these papers.
Videos
The most-watched talks and lectures in this field.
Packs
Short curated sets built around one question.
Follow
News
Press and coverage tied back to the papers.
Blogs
Author and lab write-ups of their own work.
Newsletter
Email digest of what changed, on a schedule.
Research Radar
Paste an abstract, get matches across every collection.
Yours
Saved
Papers you bookmarked in this browser.
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
Ask AI
Anna Korhonen — most-cited papers & profile · Speech Audio
← authors
·
overview
Anna Korhonen
33
papers ·
0
citations ·
0
h-index
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Dial HEALTHDIAL for Advice: A Multilingual and Multi-Parallel Spoken Dialogue Dataset for Knowledge-Grounded Information Seeking
2026
Learning Distributed Representations of Sentences from Unlabelled Data
2016
Specializing Unsupervised Pretraining Models for Word-Level Semantic Similarity
2019
Prototype-to-Style: Dialogue Generation with Style-Aware Editing on Retrieval Memory
2020
Fast, Effective, and Self-Supervised: Transforming Masked Language Models into Universal Lexical and Sentence Encoders
2021
MirrorWiC: On Eliciting Word-in-Context Representations from Pretrained Language Models
2021
Cross-Lingual Dialogue Dataset Creation via Outline-Based Generation
2022
Improving Word Translation via Two-Stage Contrastive Learning
2022
Probing Cross-Lingual Lexical Knowledge from Multilingual Sentence Encoders
2022
MULTI3NLU++: A Multilingual, Multi-Intent, Multi-Domain Dataset for Natural Language Understanding in Task-Oriented Dialogue
2022
Language-Agnostic Bias Detection in Language Models with Bias Probing
2023
Multi3WOZ: A Multilingual, Multi-Domain, Multi-Parallel Dataset for Training and Evaluating Culturally Adapted Task-Oriented Dialog Systems
2023
DIALIGHT: Lightweight Multilingual Development and Evaluation of Task-Oriented Dialogue Systems with Large Language Models
2024
Top co-authors
Ivan Vuli\'c
· 10
Nigel Collier
· 5
Fangyu Liu
· 4
Edoardo Maria Ponti
· 3
Evgeniia Razumovskaia
· 3
Songbo Hu
· 3
Xiaobin Wang
· 2
Abdullatif K\"oksal
· 1
Ahmet Akbiyik
· 1
Alexander Fraser
· 1
Alexandra Birch
· 1
and Goran Glava\v{s
· 1
Topics
cs.CL
Uncategorized
cs.AI
cs.LG
Music Generation
Audio Generation
cs.IR