BrowseComp-Plus
Emerging10papers using it
2025first seen
BrowseComp-Plus BrowseComp-Plus is a new benchmark for Deep-Research system, isolating the effect of the retriever and the LLM agent to enable fair, transparent comparisons of Deep-Research agents. The benchmark sources challenging, reasoning-intensive queries from OpenAI's BrowseComp. However, instead of searching the
Papers using BrowseComp-Plus (9)
- FineVerify: Scaling Test-Time Compute with Fine-Grained Self-Verification for Agentic SearchLACUNA: Safe Agents as Recursive Program HolesWebSwarm: Recursive Multi-Agent Orchestration for Deep-and-Wide Web SearchActiveMem: Distributed Active Memory for Long-Horizon LLM ReasoningLLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via State ProprioceptionECHO: Prune To Act, Trace To Learn With Selective Turn Memory In Agentic RLDr-DCI: Scaling Direct Corpus Interaction via Dynamic Workspace ExpansionNatural Language Query to Configuration for Retrieval AgentsSkyRL-Agent: Efficient RL Training for Multi-turn LLM Agent