Search-QA
Emerging5papers using it
491HF downloads
23HF likes
2026first seen
We publicly release a new large-scale dataset, called SearchQA, for machine comprehension, or question-answering. Unlike recently released datasets, such as DeepMind CNN/DailyMail and SQuAD, the proposed SearchQA was constructed to reflect a full pipeline of general question-answering. That is, we start not from an exi
π€ Hugging Faceβ unknown
Papers using Search-QA (5)
- StepOPSD: Step-Aware Online Preference Distillation for Agent Reinforcement LearningT$^2$PO: Uncertainty-Guided Exploration Control for Stable Multi-Turn Agentic Reinforcement LearningDynamic Skill Lifecycle Management for Agentic Reinforcement LearningSelf-Distilled Agentic Reinforcement LearningSKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization