LiveCodeBench-v-6
Emerging12papers using it
2025first seen
LiveCodeBench-v-6 is a benchmark dataset used to evaluate code generation and reasoning capabilities, specifically focusing on the effectiveness of different strategies in achieving successful code outputs.
Papers using LiveCodeBench-v-6 (12)
- SCOPE: Leveraging Subgoal Critiques for Code GenerationCast a Wider Net: Coordinated Pass@K Policy Optimization for Code ReasoningPrimal Generation, Dual Judgment: Self-Training from Test-Time ScalingEmbarrassingly Simple Self-Distillation Improves Code GenerationBreaking Training Bottlenecks: Effective and Stable Reinforcement Learning for Coding ModelsBACE: LLM-based Code Generation through Bayesian Anchored Co-Evolution of Code and Test PopulationsStep 3.5 Flash: Open Frontier-Level Intelligence with 11B Active ParametersX-Coder: Advancing Competitive Programming with Fully Synthetic Tasks, Solutions, and TestsCoreThink: A Symbolic Reasoning Layer to reason over Long Horizon Tasks with LLMsPromptCoT 2.0: Scaling Prompt Synthesis for Large Language Model ReasoningRethinking Verification for LLM Code Generation: From Generation to TestingPromptCoT 2.0: Scaling Prompt Synthesis for Large Language Model
Reasoning