Awesome Large Language Models
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β authors
Β·
overview
Loading authorβ¦
π€
Ask AI
Yi-Cheng Lin β most-cited papers & profile Β· Large Language Models
β authors
Β·
overview
Yi-Cheng Lin
22
papers Β·
58
citations Β·
0
h-index
Google Scholar β
Semantic Scholar β
OpenAlex β
Most-cited papers
DeSTA2.5-Audio: Toward General-Purpose Large Audio Language Model with Self-Generated Cross-Modal Alignment
2025 Β· 46 citations
Fake-Mamba: Real-Time Speech Deepfake Detection Using Bidirectional Mamba as Self-Attention's Alternative
2025 Β· 6 citations
Towards audio language modeling -- an overview
2024 Β· 4 citations
How Does Instrumental Music Help SingFake Detection?
2025 Β· 1 citations
Improving Speech Emotion Recognition in Under-Resourced Languages via Speech-to-Speech Translation with Bootstrapping Data Selection
2024 Β· 1 citations
WaveSP-Net: Learnable Wavelet-Domain Sparse Prompt Tuning for Speech Deepfake Detection
2025
Pseudo2Real: Task Arithmetic for Pseudo-Label Correction in Automatic Speech Recognition
2025
MI-Fuse: Label Fusion for Unsupervised Domain Adaptation with Closed-Source Large-Audio Language Model
2025
DeSTA2.5-Audio: Toward General-Purpose Large Audio Language Model with Self-Generated Cross-Modal Alignment
2025
Multi-Distillation from Speech and Music Representation Models
2025
BreezyVoice: Adapting TTS for Taiwanese Mandarin with Enhanced Polyphone Disambiguation -- Challenges and Insights
2025
Leveraging Joint Spectral and Spatial Learning with MAMBA for Multichannel Speech Enhancement
2024
Efficient Training of Self-Supervised Speech Foundation Models on a Compute Budget
2024
Codec-SUPERB @ SLT 2024: A lightweight benchmark for neural audio codec models
2024
Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks
2024
Topics
Audio Understanding
Speech Recognition
Multimodal Audio
Speech Translation
Speech Enhancement
Speaker Analysis
Text-to-Speech
cs.SD
eess.AS
Vision-Language Models