Awesome Computer Vision
π
Papers
π§
Topics
π₯
Trending
πΊοΈ
Map
π
Leaderboards
π
Learn
π€
Ask AI
β―
More
π₯
Authors
π
Reading Packs
π
Datasets
π οΈ
Tools
π°
News
π
Blogs
βοΈ
Newsletter
π―
Research Radar
π
Saved
+ Add Paper
βΎ
β
β authors
Β·
overview
Loading authorβ¦
π€
Ask AI
Haozhe Zhao β most-cited papers & profile Β· Computer Vision
β authors
Β·
overview
Haozhe Zhao
10
papers Β·
4
citations
Google Scholar β
Semantic Scholar β
OpenAlex β
Most-cited papers
Towards End-to-End Embodied Decision Making via Multi-modal Large Language Model: Explorations with GPT4-Vision and Beyond
2023 Β· 4 citations
From Context to Skills: Can Language Models Learn from Context Skillfully?
2026
Step-wise Rubric Rewards for LLM Reasoning
2026
A Goal Without a Plan Is Just a Wish: Efficient and Effective Global Planner Training for Long-Horizon Agent Tasks
2025
Teaching Large Language Models to Maintain Contextual Faithfulness via Synthetic Tasks and Reinforcement Learning
2025
ML-Bench: Large Language Models Leverage Open-source Libraries for Machine Learning Tasks
2023
An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models
2024
Selecting Influential Samples for Long Context Alignment via Homologous Models' Guidance and Contextual Awareness Measurement
2024
Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey
2024
Topics
Training Techniques
Vision-Language
Evaluation
In-Context Learning
Reinforcement Learning
Fine-Tuning
Safety & Alignment
attention
Model Architecture
Efficiency