Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Daya Guo — most-cited papers & profile · Large Language Models
← authors
·
overview
Daya Guo
17
papers ·
3882
citations ·
24
h-index
Sun Yat-sen University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
CodeBERT: A Pre-Trained Model for Programming and Natural Languages
2020 · 2434 citations
CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation
2021 · 415 citations
DeepSeek-V3 Technical Report
2024 · 248 citations
CodeBLEU: a Method for Automatic Evaluation of Code Synthesis
2020 · 188 citations
GraphCodeBERT: Pre-training Code Representations with Data Flow
2020 · 158 citations
Automating Code Review Activities by Large-Scale Pre-training
2022 · 155 citations
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
2024 · 116 citations
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
2024 · 92 citations
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
2024 · 49 citations
SparseCoder: Identifier-Aware Sparse Transformer for File-Level Code Summarization
2024 · 10 citations
LongCoder: A Long-Range Pre-trained Language Model for Code Completion
2023 · 7 citations
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
2025 · 5 citations
RLCoder: Reinforcement Learning for Repository-Level Code Completion
2024 · 3 citations
RepoTransBench: A Real-World Multilingual Benchmark for Repository-Level Code Translation
2024 · 2 citations
PhoenixRepair: Rethinking Repair Strategy Exploration in Software Agents
2026
Topics
Code Generation
Code Models
Software Engineering
Code Understanding
Code Translation
Code Agents
Multi-Agent
Evaluation
Tool Use
Planning