Awesome AI for Code
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Hilde Kuehne — most-cited papers & profile · AI for Code
← authors
·
overview
Hilde Kuehne
15
papers ·
32
citations ·
0
h-index
University of Tübingen
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
NeuralNetwork-Viterbi: A Framework for Weakly Supervised Video Learning
2018 · 22 citations
Action Sets: Weakly Supervised Action Segmentation without Ordering Constraints
2017 · 3 citations
Weakly Supervised Grounding for VQA in Vision-Language Transformers
2022 · 3 citations
TAEC: Unsupervised Action Segmentation with Temporal-Aware Embedding and Clustering
2023 · 2 citations
mWhisper-Flamingo for Multilingual Audio-Visual Noise-Robust Speech Recognition
2025 · 1 citations
ConMe: Rethinking Evaluation of Compositional Reasoning for Modern VLMs
2024 · 1 citations
Towards Audio Token Compression in Large Audio Language Models
2025
VOLD: Reasoning Transfer from LLMs to Vision-Language Models via On-Policy Distillation
2025
Omni-R1: Do You Really Need Audio to Fine-Tune Your Audio LLM?
2025
M2SVid: End-to-End Inpainting and Refinement for Monocular-to-Stereo Video Conversion
2025
Style Agnostic 3D Reconstruction via Adversarial Style Transfer
2021
Comparison of Multilingual Self-Supervised and Weakly-Supervised Speech Pre-Training for Adaptation to Unseen Languages
2023
Whisper-Flamingo: Integrating Visual Features into Whisper for Audio-Visual Speech Recognition and Translation
2024
State-Space Large Audio Language Models
2024
Top co-authors
Adhiraj Ghosh
· 1
Alessio Tonioni
· 1
Ameya Prabhu
· 1
Ana Klimovic
· 1
Andreas Hochlehnert
· 1
Bernt Schiele
· 1
Dhruba Ghosh
· 1
Elaine Sui
· 1
Elisa Ricci
· 1
Federico Tombari
· 1
Hasan Hammoud
· 1
Jehanzeb Mirza
· 1
Topics
Multimodal Audio
Video Understanding
Speech Recognition
Audio Understanding
Segmentation
Vision-Language Models
Visual QA & Reasoning
Benchmarks
Image Generation
3D Vision