VL-RewardBench
Emerging7papers using it
2024first seen
Dataset Card for VLRewardBench Project Page: https://vl-rewardbench.github.io Dataset Summary VLRewardBench is a comprehensive benchmark designed to evaluate vision-language generative reward models (VL-GenRMs) across visual perception, hallucination detection, and reasoning tasks. The benchmark contains 1,250 high-qua
Papers using VL-RewardBench (7)
- DeltaRubric: Generative Multimodal Reward Modeling via Joint Planning and VerificationLearning What Matters: Dynamic Dimension Selection and Aggregation for Interpretable Vision-Language Reward ModelingMSRL: Scaling Generative Multimodal Reward Modeling via Multi-Stage Reinforcement LearningSelf-Improving VLM Judges Without Human AnnotationsSkywork-VL Reward: An Effective Reward Model for Multimodal Understanding and ReasoningFrom Captions to Rewards (CAREVL): Leveraging Large Language Model
Experts for Enhanced Reward Modeling in Large Vision-Language ModelsVL-RewardBench: A Challenging Benchmark for Vision-Language Generative Reward Models