Retrieval, embeddings and long context / Benchmark leaderboard
BEIR
Zero-shot retrieval across heterogeneous domains and corpora.
No published company result yet.
This benchmark is in the research directory. Its task package, adapter and grading protocol need qualification before a hosted run can be offered.
Arrange an evaluation for your company →What this benchmark measures
Metrics
nDCG@10, recall and other declared retrieval metrics.
Execution requirements
Search/index pipeline with fixed corpus, chunking and embeddings.
Scope and limitations
For database comparisons, hold embeddings, retrieval configuration and recall targets fixed; report latency separately.
Evaluation availability
Catalog entry. Request a managed evaluation to qualify your agent interface and the benchmark’s native grading requirements.
Sources and company fit
- Original benchmark ↗Established
- Dataset / catalog ↗
Companies whose products may fit
Research recommendations based on product capabilities. These companies have not necessarily run this benchmark or integrated with Blobfish.
Compatibility notes for each company
Cohere: Direct component API + selected system adapters
Glean: Capability-aligned; adapter/access to qualify
Jina AI: Direct component API + selected system adapters
Pinecone: Composite system
Qdrant: Composite system
Vectara: Capability-aligned; adapter/access to qualify
Voyage AI: Direct component API + selected system adapters
Weaviate: Composite system
Zilliz: Composite system