explainx.ai0k
TrendingAI News TodayPathwaysSkills
Pricing
explainx.ai

Upskill in AI — 16 free pathways, live workshops & bootcamps, and 50+ courses from practitioners. Plus the skills, tools, and MCP servers to practice on.

follow us

follow on google

Add explainx.ai as a preferred source

corporate training

support@explainx.ai

get started

Find your pathTake Free Evaluation

community

Join the community

learn

mind: share how you thinkpathways — start freeworkshopsbootcampscoursescompare Explainxcertificationsmock testsexplainx universitycorporate traininglearn skills & mcp

discover

skillsmcp serversexplainx mcptoolsmdx readeragentsllmsdesignsdictionarypeopleagi trackerfelony benchranks

company

aboutvisionmissionteaminstructorsteach on explainxpartnershipscommunityhackathonscareers

content

daily AI newsstate of AI — live resultsblogreleasespromptsgeneratorsresource libraryfor LLMsexplainx.ai kids

solutions

all solutionsdeveloper upskillingmarketing upskillingproduct manager upskillingleadership upskilling

newsletter · weekly

Get AI news, tools, and insights in your inbox.

supportcontactprivacytermsdata rightshow we create contentsubmission guidelines

© 2026 AISOLO Technologies Pvt Ltd

explainx.ai

  1. Home
  2. /
  3. Dictionary
  4. /
  5. context-bench
Evaluation & Benchmarksaka turbopuffer context-benchaka Context Bench

context-bench

A private turbopuffer benchmark for contextual embeddings that scores document disambiguation, answer-chunk recall, and supporting-evidence recall on long documents.

Ask Melo about this← all terms

Released alongside Perplexity's pplx-embed-v2-context-9b-preview on September 30, 2026, context-bench holds 2,099 queries over 38,894 long documents (about 2.46 million sentence chunks) across 21 domains. Labels stay private to limit contamination; labs request evals at contextbench@turbopuffer.com. Metrics include Document@K, Answer@K, and Evidence Recall@K. Figures quoted from Perplexity's post are vendor-submitted.

Related terms

AI BenchmarkRecallQ2D-WebRetrieval-Augmented GenerationGSM8KHuman Evaluation