About this benchmark
This benchmark, created by OpenAI, evaluates the ability of language models to answer short, fact-seeking questions.
.png)
Empower your AI agents, copilots, search applications and automated workflows
indexed documents
for web retrieval at scale
requests per day
in sustained workloads
P99 latency
for production workloads
Get LLM-ready excerpts from web pages: more context than classic search snippets, fewer tokens than full-page scraping.

Pay-as-you-go pricing with transparent, consumption-based rates. Volume discounts available upon request.
Least-privilege IAM, short-lived tokens, and encryption with customer-managed keys — backed by ISO/IEC certifications.
Data minimization, zero-data retention for enterprise clients, transparent GDPR-tailored data processing terms.
.png)
Quick start
Run queries via Yandex Cloud ML SDK / REST / gRPC with quickstart flow (service account + payment method + API key).
.png)
Al-native architecture
MCP-compatible and designed to integrate seamlessly with modern agent frameworks.

Ranked links and short snippets from Yandex search engine for grounding and citations, available in XML and HTML formats.

Original, query-relevant excerpts from each page — more context than a classic snippet, fewer tokens than the full page. Built for RAG and AI agents.

Search images by text query or URL and retrieve results in JSON.

Get a single AI-generated answer synthesized from the retrieved results.

This benchmark, created by OpenAI, evaluates the ability of language models to answer short, fact-seeking questions.
You pay via Yandex Cloud billing. Individuals typically pay by credit/debit card. Businesses can pay by card or bank transfer, depending on the billing account type and country.
Connect Yandex search capabilities to your products — no server infrastructure needed, with automatic scaling and generative answer support.
