Awesome Large Language Models
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Shiwei Liu — most-cited papers & profile · Large Language Models
← authors
·
overview
Shiwei Liu
28
papers ·
64
citations ·
12
h-index
Beihang University
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
Consmax: Hardware-friendly Alternative Softmax With Learnable Parameters
2024 · 17 citations
Junk DNA Hypothesis: Pruning Small Pre-trained Weights Irreversibly And Monotonically Impairs "difficult" Downstream Tasks In Llms
2023 · 13 citations
Demystifying the Roles of LLM Layers in Retrieval, Knowledge, and Reasoning
2025 · 10 citations
Learning from the Self-future: On-policy Self-distillation for dLLMs
2026
Acttail: Global Activation Sparsity In Large Language Models
2026
When Fewer Layers Break More Chains: Layer Pruning Harms Test-Time Scaling in LLMs
2025
O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning
2025
Mask-Enhanced Autoregressive Prediction: Pay Less Attention to Learn More
2025
Stable-SPAM: How to Train in 4-Bit More Stably than 16-Bit Adam
2025
SoS1: O1 and R1-Like Reasoning LLMs are Sum-of-Square Solvers
2025
LIFT the Veil for the Truth: Principal Weights Emerge after Rank Reduction for Reasoning-Focused Supervised Fine-Tuning
2025
Chain-of-Experts: Unlocking the Communication Power of Mixture-of-Experts Models
2025
GPTailor: Large Language Model Pruning Through Layer Cutting and Stitching
2025
GPAS: Accelerating Convergence of LLM Pretraining via Gradient-Preserving Activation Scaling
2025
Diffusion Language Models Know the Answer Before Decoding
2025
Top co-authors
Lu Yin
· 10
Li Shen
· 5
Pengxiang Li
· 4
Zhangyang Wang
· 4
Zhenyu Zhang
· 4
Ajay Jaiswal
· 3
Tianjin Huang
· 3
Xinyuan Song
· 3
Gaojie Jin
· 2
Guinan Su
· 2
Jiawei Zhao
· 2
Jonas Geiping
· 2
Topics
Efficiency
Fine-Tuning
Model Architecture
Training Techniques
fine-tuning
Code
reasoning
Evaluation
LLMs
RAG