Artifact Benchmark
BrowseComp-Plus
Canonical ID https://id.searchplex.net/browsecomp-plus/
A controlled benchmark for evaluating deep-research agents and retrievers, derived from BrowseComp and built around a fixed curated corpus with supporting documents and hard negatives.
Scope
BrowseComp-Plus is designed to make evaluation of deep-research agents and retrieval methods more fair and transparent by controlling the document corpus and evidence. It is related to BrowseComp but should not be treated as identical to live-web browsing evaluation.
Observed terms
- BCP
- BrowseComp Plus
Evaluates
Not equivalent to
- BrowseComp-Plus is not BEIR. BrowseComp-Plus evaluates deep-research agents and retrieval methods over BrowseComp-derived tasks; BEIR is an information-retrieval benchmark suite.
- BrowseComp-Plus is not BrowseComp. BrowseComp-Plus derives from BrowseComp but uses a fixed curated corpus and supporting evidence; BrowseComp evaluates browsing agents under live web browsing conditions.
- BrowseComp-Plus is not MTEB. BrowseComp-Plus evaluates deep-research agents and retrievers; MTEB evaluates text embedding models across multiple tasks.
Learn more
ID
- Canonical ID
https://id.searchplex.net/browsecomp-plus/ - Scheme
Searchplex ID - Machine
JSONJSON-LD