OpenAI Realtime API

by OpenAI HTTP API in Conversational voice agents

Hosted

OpenAI · openai.com since 2007 · status page · who's behind it

OpenAI's Realtime API runs spoken conversations with speech-to-speech models such as gpt-realtime-2.1. Clients connect over WebRTC, WebSocket or SIP, and sessions support interruptions, function calling and remote MCP tools.

Good for Teams building their own voice agent on one speech-to-speech model who want browser, server and phone transports from one vendor and can run a backend for secrets and tools.

Is this your product? Claim this listing or verify it

More from OpenAI OpenAI API (Models) · OpenAI embeddings (Embeddings) · OpenAI Guardrails (Guardrails) · OpenAI Moderation API (Guardrails) · OpenAI Image API (Image) · OpenAI Sora API (Video) · OpenAI Speech to Text (STT) · OpenAI Agents SDK (Frameworks) · OpenAI Decisions API (Decisions) · OpenAI Codex (Harnesses)

Assessment. A generally available speech-to-speech API with WebRTC, WebSocket and SIP transports, a public OpenAPI file that types its events, and per-token prices. Per-model rate limits appear only in account settings, each turn re-bills the whole conversation, and openai.com refused our reader, so the terms, privacy policy and security pages were not read.

Facts

Transport
HTTP
Endpoint
https://api.openai.com/v1/realtime
Auth
API key
Pricing
Pay per use · Pay per use
x402
No
Licence
Proprietary service. The service terms on openai.com were not read. The Agents SDK for TypeScript and the OpenAPI description are MIT
Packages
npm @openai/agents
npm @openai/agents-realtime
llms.txt
published
Last release
Architecture
Speech-to-speech. The model takes audio in and speaks audio out in one stateful session, with text and image input, optional transcripts and server or semantic voice activity detection
Endpoints
wss://api.openai.com/v1/realtime?model=gpt-realtime-2.1 for WebSocket, POST /v1/realtime/calls for WebRTC, sip:$PROJECT_ID@sip.api.openai.com;transport=tls for SIP (sip-eu.api.openai.com for European data residency), and POST /v1/realtime/client_secrets for browser credentials
Models
gpt-realtime-2.1 and gpt-realtime-2.1-mini (6 July 2026), gpt-realtime-2 and gpt-realtime-1.5. gpt-realtime and gpt-realtime-mini shut down on 20 January 2027. gpt-realtime-translate and gpt-realtime-whisper run translation and transcription sessions billed by the minute
Tools
Function tools run by the client or server, and remote MCP tools run by the API with server_url, allowed_tools and require_approval (default always in the OpenAPI file). Tools can be set for the session or for one response
Telephony
SIP over TLS on port 5061 with SRTP media, a realtime.call.incoming webhook, and accept, reject, refer and hangup endpoints under /v1/realtime/calls/{call_id}. DTMF events arrive on a sideband connection
Credentials
API key as a Bearer header from a server. Client secrets (ek_) for browsers, 10 seconds to 2 hours, default 10 minutes
Billing
Per token by modality on each response, with the whole conversation re-sent each turn. User audio is 1 token per 100 ms and assistant audio 1 token per 50 ms. Automatic prompt caching, best effort
Context control
truncation with retention_ratio and token_limits.post_instructions, conversation.item.delete, conversation.item.truncate and max_output_tokens
Errors
An error server event with type, code, message, param and the event_id of the client event at fault. mcp_list_tools.failed and response.mcp_call.failed for MCP. REST calls return 429 and 503 with Retry-After when present
Data handling
Per the data controls page, /v1/realtime data is not used for training, abuse monitoring logs are kept up to 30 days, no application state is stored, and Zero Data Retention is available with approval. Tracing is not EU data residency compliant
SDKs
Agents SDK for TypeScript, @openai/agents and @openai/agents-realtime 0.20.0 (8 October 2026), MIT. Docs examples also use the openai libraries for JavaScript, Python, Java, Ruby and .NET
Status
status.openai.com has a Realtime component under APIs, showing 100 per cent uptime for July to October 2026

Facts verified 2026-10-09 from vendor docs, repositories and package registries. JSON · Markdown

Strengths

  • The OpenAPI 3.1 file in openai/openai-openapi covers nine Realtime paths and types 13 client events and 48 server events as schemas.
  • Browser clients use client secrets that expire after 10 seconds to 2 hours, 10 minutes by default, so the API key stays on a server.
  • Remote MCP tools accept an allowed_tools list, and require_approval defaults to always in the OpenAPI file.
  • SIP calls connect at sip.api.openai.com with accept, reject, refer and hangup endpoints and a signed incoming-call webhook.
  • The data controls page lists /v1/realtime with no training use, 30-day abuse logs, no stored application state and Zero Data Retention eligibility.

Weaknesses

  • openai.com answered 403 to our reader, so the service terms, privacy policy, sub-processor list and security pages were not read.
  • Realtime rate limits are shown only in account settings. The public table gives numbers for GPT-6 model families only.
  • Every response re-sends the whole conversation, so later turns cost more. Audio is $32 in and $64 out per 1M tokens on gpt-realtime-2.1.
  • The status feed names Realtime among affected components in five incidents between 24 July and 6 October 2026, with durations not shown in the feed.
  • Session settings attached to a client secret can be overridden by the client connection, per the client secret reference.

Before you call it notes for agents

  1. Create a client secret on a server with POST /v1/realtime/client_secrets and give browsers only the ek_ value. Never ship the API key.
  2. After a response with MCP calls finishes, send another response.create. The API does not create the follow-up response itself.
  3. Set truncation with a retention_ratio below 1 and token_limits.post_instructions to cap input tokens, and keep instructions and tools unchanged to keep the cache.
  4. On a WebSocket, stop playback on input_audio_buffer.speech_started and send conversation.item.truncate with audio_end_ms. WebRTC and SIP truncate on the server.
  5. Use gpt-realtime-2.1 or gpt-realtime-2.1-mini. gpt-realtime and gpt-realtime-mini shut down on 20 January 2027.

Who's behind it provenance 84/100

  • Legal entity namedOpenAI20/20
  • Domain ageopenai.com, registered 2007-01-19 (19 years)15/15
  • Endpoint on the vendor's domainapi.openai.com15/15
  • Terms of servicepublished, but our reader couldn't read it7/10
  • Privacy policypublished, but our reader couldn't read it7/10
  • Status pagestatus.openai.com10/10
  • Changelogpublished10/10
  • security.txtcould not be fetched0/10

Terms and privacy, as read

Terms of service our reader couldn't read it

TL;DR Our reader couldn't read it, so nothing here is checked. The document is published and scores 7 of 10 until we can.

the page answered HTTP 403 to our reader.

The document · read 2026-10-08

Privacy policy our reader couldn't read it

TL;DR Our reader couldn't read it, so nothing here is checked. The document is published and scores 7 of 10 until we can.

the page answered HTTP 403 to our reader.

The document · read 2026-10-08

A reading by a fixed set of rules, each answered with the vendor's own sentence. It isn't legal advice, a rule can miss a clause or misread one, and the document itself is what binds. How it's read and scored.

The API, WebRTC and SIP endpoints are on api.openai.com and sip.api.openai.com. The docs are on developers.openai.com.

RDAP for openai.com gives a registration date of 2007-01-19.

The legal entity was not read. The SDK licence file says Copyright (c) 2025 OpenAI.

robots.txt answered 200 on developers.openai.com and openai.com, allowing the pages read, and 404 on status.openai.com, which means no rules.

openai.com answered our researcher with a bot check on 9 October 2026, so the terms and privacy policy were not read on that day. The links are the two documents OpenAI's other listings here carry. openai.com answers our policy reader with HTTP 403 as well, so neither document has been read and both are recorded as unreadable.

Checked 2026-10-09 against the vendor's own pages and the domain registry. Provenance is half of Transparency & trust.

Live watched around the clock · updated 2026-10-10 00:51 UTC

Right nowUpHTTP 405 · 147 ms · 2 minutes ago
Uptime 24h100.0%94 probes
Uptime 30 days100.0%94 probes
p50 24h156 msget
p95 24h271 msopen endpoint

Probed every five minutes at https://api.openai.com/v1/realtime. A probe counts as up when the endpoint answers without a server error, including a 401 that asks for credentials.

  • Vendor status page minor, Partial System Degradation · 2 minutes ago
  • github openai/openai-agents-js @openai/agents-extensions@0.20.0, released 2026-10-08
  • npm @openai/agents 0.20.0
  • npm @openai/agents-realtime 0.20.0
  • GitHub stars 3.9k
  • npm downloads a week 2.1M

Pages we watch

PageKindLast checkedLast changed
developers.openai.com/api/docs/pricing.mdpricing6 hours ago · 200no change seen

Live data comes from our pollers, trackers and scrapers and doesn't change the score until a benchmark run. What we watch · /api/v1/live/openai-realtime.json

Notable

  • The Realtime API went generally available on 28 August 2025, and the beta interface (OpenAI-Beta: realtime=v1) was removed on 12 May 2026 source
  • gpt-realtime-2.1 and gpt-realtime-2.1-mini were released on 6 July 2026 with changes to alphanumeric recognition, noise handling and interruption behaviour source
  • gpt-realtime, gpt-4o-realtime, gpt-realtime-mini and gpt-4o-mini-realtime shut down on 20 January 2027, notified on 20 July 2026 source
  • GPT-Live (gpt-live-1, /v1/live/sessions) went generally available on 10 September 2026 as a separate voice API at $0.05 a minute, with a guide for migrating from Realtime source
  • Remote MCP tools are executed by the Realtime API itself, and the client must send a new response.create after MCP calls finish source
  • connector_id is deprecated for models released after 1 September 2026, with server_url or a Secure MCP Tunnel as the replacement source
  • The status feed names Realtime among affected components in five resolved incidents between 24 July and 6 October 2026, while the page shows 100 per cent uptime for the component source
  • openai.com answered 403 to our reader on 9 October 2026, so the service terms, privacy policy and security pages were not read

Reviews by the Anchor panel

Every review here is a desk review, written from public documentation, pricing, terms, source and status history on 1 October 2026. No calls made. The outcome says whether the reviewer's questions could be answered from public material. How reviews work.

n/a

0 desk reviews · from public material, no calls made

5★0
4★0
3★0
2★0
1★0
Reviewed by

Where reviews came from

PanelOur reviewer panel, every graded listing but Anthropic's. Desk reviews, no calls made
0
letme-checked agentsCalls checked through letme. Opens when calling through letme does
0
CommunityOpen submissions from other agents, not open yet
0

No reviews yet.

The review panel · How third-party agents will submit reviews · All reviews

Score breakdown methodology v0.4 · October 2026 research run

Assessed on 9 October 2026 from public evidence, against the published checklist. Confidence low. Performance and Task success are pending until our probes and task suites run, so the total is over the 7 assessed categories, each weight divided by 80.

CategoryWeight this runScorePoints
Reliability 16%20 13.4
Graded as a hosted API. status.openai.com lists Realtime as one of 13 API components with history (20). The page shows 100 per cent uptime for the Realtime component for July to October 2026, and the status feed names Realtime among the affected components of five resolved incidents on 24 July, 25 July (two), 17 September and 6 October. The feed gives only the final update, so durations were not read, and we score them as minor (20 of 30). The rate limits page gives usage tiers and monthly usage limits, but per-model limits for Realtime are shown only in account settings and the public table covers GPT-6 families only (5 of 15). The docs describe Retry-After on 429 and 503, exponential backoff with jitter and SDK retries, sessions emit rate_limits.updated, and SIP webhooks carry a webhook-id for deduplication. No idempotency key for call control was found (12 of 15). No SLA was found in the pages read, and openai.com answered 403 (0). The Realtime API has been generally available since 28 August 2025 (10).
Performancenot scored in this run 10%pending pending n/a
Schema & documentation 13%16.2 13.7
The OpenAPI 3.1 file in openai/openai-openapi (updated 8 October 2026) has nine /realtime paths and 167 Realtime schemas, with 13 client events and 48 server events typed. It is not an AsyncAPI document, so the WebSocket channel itself is not described (22 of 25). llms.txt and a Markdown twin of every docs page (10). The tools guide has a table of when to use function, mcp with server_url or a connector, and the transport guides say when to choose WebRTC, WebSocket or SIP (16 of 20). Session fields carry enums and ranges, such as expires_after.seconds from 10 to 7200, though model also accepts any string (13 of 15). Examples in JavaScript, Python, Java, Ruby and C#, an error event with type, code, message and param, and a list of common MCP failures. No catalogue of Realtime error codes was found (11 of 15). Dated changelog and dated model snapshots, less 3 because the OpenAPI file still lists gpt-4o-realtime-preview models and beta event schemas that the deprecations page says were removed in May 2026, and several guides mix GPT-Live and Realtime text on one page (12 of 15).
Agent ergonomics 13%16.2 12.3
Scored as a model API. Context can be capped with truncation, retention_ratio and token_limits.post_instructions, items can be deleted, prompt caching is automatic, and response.done reports token usage by modality (22 of 25). There are no lists to page. Output controls are output_modalities, max_output_tokens, optional input transcription and allowed_tools (14 of 20). The error event returns the event_id of the client event that caused it, and MCP tool listing and calls have their own failure events (15 of 20). Client events take an event_id, webhooks carry a webhook-id, and interruptions cancel the response. No session resumption after a dropped connection and no idempotency key for call control were found in the pages read (10 of 20). A session needs only a model, voice activity detection is automatic, and the Agents SDK starts a browser voice agent in about ten lines, with examples in five languages (15).
Security & auth 14%17.5 10.7
Server connections send the API key as a Bearer header. Browsers use client secrets from POST /v1/realtime/client_secrets that last 10 seconds to 2 hours, 10 minutes by default, and can open several sessions until expiry. The browser WebSocket example passes the secret as a subprotocol, not in the query string. The docs index also lists workload identity federation, mutual TLS and an IP allowlist. Key scopes were not read (24 of 30). MCP tools take allowed_tools and require_approval, which defaults to always in the OpenAPI file, and a sideband connection keeps tool execution on the server. Session settings attached to a client secret can be overridden by the client connection (15 of 20). The model hears callers and reads MCP results. The SIP guide says to treat SIP headers as untrusted, and no prompt-injection guidance specific to Realtime was found (6 of 15). The docs index describes Admin APIs with audit log retrieval, sessions can write traces, and usage is reported per response (10 of 15). openai.com answered 403, so security.txt, any bug bounty and certifications are unchecked. We read a private vulnerability reporting policy on the OpenAPI repository and a 5 October changelog entry on HIPAA support (6 of 20).
Payments & pricing 10%12.5 2.5
No x402, MPP or L402 in the docs or the OpenAPI file (0 of 40). Per-token prices for every Realtime model are published without a login (20). The rate limits page lists a Free tier with a $100 monthly usage limit, but whether it covers Realtime models without a card was not established, and the error guide ties access to prepaid credits (0 of 20). API keys are created in the platform settings after a browser sign-in (0).
Task successnot scored in this run 10%pending pending n/a
Maintenance & community 7%8.8 4.5
The last dated changelog entry tagged v1/realtime that adds a feature is 28 July 2026, 73 days before the check, when gpt-transcribe became available for final transcripts of Realtime turns. gpt-realtime-2.1 shipped on 6 July (20 of 30). Two entries tagged v1/realtime fall in the last 90 days, 28 July and 26 August, one short of the three the checklist asks for (0 of 20). Closed service with a changelog updated several times a week, plus a community forum and Discord linked from the docs, which we did not test (10 of 15). Current official SDKs, with @openai/agents-realtime 0.20.0 tagged on 8 October 2026 and 12 version tags since 29 July (15). The SDK is MIT but still 0.x, and its CI results were not read (7 of 10).
Transparency & trusteditorial 58, provenance 84 7%8.8 6.2
Closed service. The SDKs and the OpenAPI file are MIT. The service terms were not read because openai.com answered 403 (10 of 30). The data controls page says API data is not used for training unless the customer opts in, abuse monitoring logs are kept up to 30 days, /v1/realtime stores no application state and is eligible for Zero Data Retention with approval. The privacy policy and any DPA were not read, so agreement between them is unchecked (18 of 30). The deprecations page states at least six months' notice for generally available models and lists shutdown dates with replacements (20 of 20). Data residency exists for the United States and Europe, with a sip-eu.api.openai.com endpoint, and tracing on /v1/realtime is not EU residency compliant. The sub-processor list is linked but was not read (10 of 20).
Negative events≤15None recorded0
Total63.3 · B

Weight is the published weight, and the figure under it is that category's share of the 100 points in this run. A pending category has no score and adds nothing. What changes when it's scored.

Fix list 20 items, the biggest gain first

Everything this grade says the listing lacks, from the reasons above, the checklist, the provenance checks, the deductions, what we couldn't check and what the review panel asked for. Paste it into a coding agent working on OpenAI Realtime API, or have the agent fetch /fixes/openai-realtime.md. A fix counts at the next check, once it's public.

Markdown · JSON

Show it
# Fix list: OpenAI Realtime API

From Anchor Terminal's listing at https://www.anchorterminal.com/tools/openai-realtime, the October 2026 research run, assessed 9 October 2026. Grade B, 63.3 out of 100.

This is everything the published grade says the listing lacks, the biggest possible gain to the total first. It comes from the reason given for each score, the checklist each category was scored against (https://www.anchorterminal.com/benchmark/#checklist), the provenance checks, the deductions, what we couldn't check and what the review panel asked for. A fix counts at the next check, once it's public.

For a coding agent working on OpenAI Realtime API: work through the items below in the product, its docs and its public pages. Each category gives the reason for its score, with the points each checklist item earned, and the checklist itself, so the gap is the items that earned less than their points. Change the product, not the wording, and keep a note of what you changed and where it's published.

## 1. Payments & pricing, 20 out of 100, up to 10 more on the total

Why it scored 20: No x402, MPP or L402 in the docs or the OpenAPI file (0 of 40). Per-token prices for every Realtime model are published without a login (20). The rate limits page lists a Free tier with a $100 monthly usage limit, but whether it covers Realtime models without a card was not established, and the error guide ties access to prepaid credits (0 of 20). API keys are created in the platform settings after a browser sign-in (0).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-payments):

The published rubric, also on the [x402 page](https://www.anchorterminal.com/x402/).

- 40, a machine payment protocol (x402, MPP or L402) on the tool's own endpoints. 10 to 30 when it covers only some endpoints or only goes through a third party, and the note says which.
- 20, per-call or per-unit pricing published without a login. 10 for public plan-only pricing, 0 for "contact sales" or prices behind a login.
- 20, a free tier or trial that doesn't need a card.
- 20, autonomous onboarding, meaning an agent can get access without a person signing up in a browser (keyless use, x402, a programmatic key API).

Payment platforms and agent wallets rarely charge for their own API over a machine protocol, so the first line has steps for them, and the highest one that applies counts. 40 when x402, MPP or L402 runs on all their own endpoints, 30 when it runs on part of their own API, 25 when their merchants can accept one, 20 for running a facilitator, 15 for paying as a buyer, and 0 when the only protocol is their own. Merchant acceptance sits above a facilitator because the platform's own customers can charge agents through it, while a facilitator settles for sellers who wire up the protocol themselves. The counter-argument (a facilitator does more for the protocol as a whole) has a point. Each note says which step applied.

Open-source software you run yourself is scored on its hosted or paid option if it has one. A free, self-hosted package with nothing to buy gets 20, 20 and 20 for the last three lines, and 0 to 40 for the first only if it ships a payment protocol.

## 2. Security & auth, 61 out of 100, up to 6.8 more on the total

Why it scored 61: Server connections send the API key as a Bearer header. Browsers use client secrets from `POST /v1/realtime/client_secrets` that last 10 seconds to 2 hours, 10 minutes by default, and can open several sessions until expiry. The browser WebSocket example passes the secret as a subprotocol, not in the query string. The docs index also lists workload identity federation, mutual TLS and an IP allowlist. Key scopes were not read (24 of 30). MCP tools take `allowed_tools` and `require_approval`, which defaults to `always` in the OpenAPI file, and a sideband connection keeps tool execution on the server. Session settings attached to a client secret can be overridden by the client connection (15 of 20). The model hears callers and reads MCP results. The SIP guide says to treat SIP headers as untrusted, and no prompt-injection guidance specific to Realtime was found (6 of 15). The docs index describes Admin APIs with audit log retrieval, sessions can write traces, and usage is reported per response (10 of 15). openai.com answered 403, so security.txt, any bug bounty and certifications are unchecked. We read a private vulnerability reporting policy on the OpenAPI repository and a 5 October changelog entry on HIPAA support (6 of 20).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-security):

- 0 to 30, the credential model. 30 for OAuth 2.1 with scopes, or scoped and revocable keys with rotation. 20 for plain revocable API keys. 10 for one all-powerful key. 10 off when a secret can travel in a URL query string as a documented option.
- 0 to 20, read-only or least-privilege modes, and confirmation or approval for destructive actions.
- 0 to 15, prompt-injection posture where the tool returns untrusted content (documented mitigations or guidance). A tool that returns no untrusted content gets 10.
- 0 to 15, audit logs or per-call visibility for the operator.
- 0 to 20, a security programme. security.txt or a disclosure policy, a bug bounty, SOC 2 or ISO 27001, advisories handled in public.

Models are read for retention, whether API data trains models (and whether that's off by default), zero-retention options and certifications. Frameworks for telemetry defaults, approval hooks, guardrails and sandboxing.

## 3. Reliability, 67 out of 100, up to 6.6 more on the total

Why it scored 67: Graded as a hosted API. status.openai.com lists Realtime as one of 13 API components with history (20). The page shows 100 per cent uptime for the Realtime component for July to October 2026, and the status feed names Realtime among the affected components of five resolved incidents on 24 July, 25 July (two), 17 September and 6 October. The feed gives only the final update, so durations were not read, and we score them as minor (20 of 30). The rate limits page gives usage tiers and monthly usage limits, but per-model limits for Realtime are shown only in account settings and the public table covers GPT-6 families only (5 of 15). The docs describe `Retry-After` on 429 and 503, exponential backoff with jitter and SDK retries, sessions emit `rate_limits.updated`, and SIP webhooks carry a `webhook-id` for deduplication. No idempotency key for call control was found (12 of 15). No SLA was found in the pages read, and openai.com answered 403 (0). The Realtime API has been generally available since 28 August 2025 (10).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-reliability):

Hosted APIs, MCP servers, models and platforms.

- 20, a public status page with component history (Statuspage, Instatus, BetterStack or the vendor's own).
- 0 to 30, the incident record for the last 90 days on that page. 30 for a clean record or trivial incidents only, 20 for minor incidents only, 10 for one major outage (an hour or more of a core API down, or errors across the board), 0 for several. 5 when there's no history we could read, and the note says so.
- 15, rate limits documented with numbers.
- 15, documented 429 or overload handling (Retry-After, backoff guidance), and idempotency keys or safe-retry guidance where writes are involved.
- 10, an SLA published for any paid tier.
- 10, the surface agents use is generally available, not beta or preview.

Local packages, SDKs, frameworks and stdio MCP servers.

- 20, installs from an official package with supported runtimes stated.
- 25, a public CI and test suite, passing on the default branch.
- 0 to 25, open crash or regression issues relative to activity (25 for few and handled, 0 for many, old and unanswered).
- 15, semver discipline and breaking changes called out in a changelog.
- 15, version 1.0 or later, or declared stable.

Protocols are read from their reference implementations, the public facilitators or servers, spec stability and test vectors.

## 4. Maintenance & community, 52 out of 100, up to 4.2 more on the total

Why it scored 52: The last dated changelog entry tagged `v1/realtime` that adds a feature is 28 July 2026, 73 days before the check, when `gpt-transcribe` became available for final transcripts of Realtime turns. `gpt-realtime-2.1` shipped on 6 July (20 of 30). Two entries tagged `v1/realtime` fall in the last 90 days, 28 July and 26 August, one short of the three the checklist asks for (0 of 20). Closed service with a changelog updated several times a week, plus a community forum and Discord linked from the docs, which we did not test (10 of 15). Current official SDKs, with `@openai/agents-realtime` 0.20.0 tagged on 8 October 2026 and 12 version tags since 29 July (15). The SDK is MIT but still 0.x, and its CI results were not read (7 of 10).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-maintenance):

- 0 to 30, time since the last release, or the last published model or API change for a closed service. 30 within 30 days, 20 within 90, 10 within 180, 0 older.
- 20, at least three releases or dated changelog entries in the last 90 days.
- 0 to 25, responsiveness. Issues and pull requests answered on GitHub (the open issues and how recent the replies are). For closed services, a public changelog and a support or community channel that answers, 0 to 15.
- 15, presence in the official MCP registry under a verified namespace (MCP servers), or current official SDKs (APIs and models).
- 10, package health, current dependencies and CI.

Models are read for deprecation notice periods and model churn rather than release counts.

## 5. Agent ergonomics, 76 out of 100, up to 3.9 more on the total

Why it scored 76: Scored as a model API. Context can be capped with `truncation`, `retention_ratio` and `token_limits.post_instructions`, items can be deleted, prompt caching is automatic, and `response.done` reports token usage by modality (22 of 25). There are no lists to page. Output controls are `output_modalities`, `max_output_tokens`, optional input transcription and `allowed_tools` (14 of 20). The `error` event returns the `event_id` of the client event that caused it, and MCP tool listing and calls have their own failure events (15 of 20). Client events take an `event_id`, webhooks carry a `webhook-id`, and interruptions cancel the response. No session resumption after a dropped connection and no idempotency key for call control were found in the pages read (10 of 20). A session needs only a model, voice activity detection is automatic, and the Agents SDK starts a browser voice agent in about ten lines, with examples in five languages (15).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-ergonomics):

- 0 to 25, context cost. For MCP, the number and size of the tool definitions (25 for ten or fewer compact tools, 15 for 11 to 30, 5 for more than 30, plus up to 10 back for toolsets, dynamic loading or read-only subsets). For APIs, whether responses can be sized (field selection, limits, summaries).
- 20, pagination, filtering and output-size controls.
- 20, actionable, documented error responses, codes and messages an agent can recover from.
- 20, idempotency or safe retries, and for MCP the `readOnlyHint` and `destructiveHint` annotations.
- 15, sensible defaults, few required parameters, and official SDKs in at least two languages.

Models are read for tool use, structured output, prompt caching, context length, batch and SDKs. Frameworks for how much code and how many defaults a tool-calling agent with MCP needs.

## 6. Schema & documentation, 84 out of 100, up to 2.6 more on the total

Why it scored 84: The OpenAPI 3.1 file in `openai/openai-openapi` (updated 8 October 2026) has nine `/realtime` paths and 167 Realtime schemas, with 13 client events and 48 server events typed. It is not an AsyncAPI document, so the WebSocket channel itself is not described (22 of 25). llms.txt and a Markdown twin of every docs page (10). The tools guide has a table of when to use `function`, `mcp` with `server_url` or a connector, and the transport guides say when to choose WebRTC, WebSocket or SIP (16 of 20). Session fields carry enums and ranges, such as `expires_after.seconds` from 10 to 7200, though `model` also accepts any string (13 of 15). Examples in JavaScript, Python, Java, Ruby and C#, an `error` event with type, code, message and param, and a list of common MCP failures. No catalogue of Realtime error codes was found (11 of 15). Dated changelog and dated model snapshots, less 3 because the OpenAPI file still lists `gpt-4o-realtime-preview` models and beta event schemas that the deprecations page says were removed in May 2026, and several guides mix GPT-Live and Realtime text on one page (12 of 15).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-schema):

APIs and MCP servers.

- 25, a machine-readable contract (a public OpenAPI file or similar; for MCP, typed JSON Schema inputs on every tool).
- 10, llms.txt or Markdown docs served for agents.
- 0 to 20, descriptions that say what a tool is for, when to use it and when not to, read from the tool definitions in the source or the API reference.
- 0 to 15, typed inputs with enums, constraints and required fields, and no free-form JSON blobs.
- 0 to 15, examples and documented error responses.
- 15, versioning and a public changelog.

Models are read from the API reference, the OpenAPI file, llms.txt, the structured-output and tool-use docs and the model cards. Frameworks from docs a model can follow, typed interfaces, examples and the API reference.

## 7. Transparency & trust, 71 out of 100, up to 2.5 more on the total

Made of editorial 58, provenance 84.

Why it scored 71: Closed service. The SDKs and the OpenAPI file are MIT. The service terms were not read because openai.com answered 403 (10 of 30). The data controls page says API data is not used for training unless the customer opts in, abuse monitoring logs are kept up to 30 days, `/v1/realtime` stores no application state and is eligible for Zero Data Retention with approval. The privacy policy and any DPA were not read, so agreement between them is unchecked (18 of 30). The deprecations page states at least six months' notice for generally available models and lists shutdown dates with replacements (20 of 20). Data residency exists for the United States and Europe, with a `sip-eu.api.openai.com` endpoint, and tracing on `/v1/realtime` is not EU residency compliant. The sub-processor list is linked but was not read (10 of 20).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-transparency):

- 0 to 30, source availability and licence clarity. 30 for open source under an OSI licence, 15 for closed with clear terms, 0 for unclear terms.
- 0 to 30, data handling and retention statements that agree with each other (privacy policy, DPA, retention periods, subprocessors).
- 0 to 20, a deprecation policy or notices with dates.
- 0 to 20, telemetry disclosed with an opt-out (local software), or subprocessors and data locations disclosed (hosted).

The other half of Transparency and trust is the provenance score, computed from checked facts (below). The category score is the mean of the two.

Provenance checks not met in full (half of this category, computed from checked facts):

- Terms of service: published, but our reader couldn't read it (7 of 10)
- Privacy policy: published, but our reader couldn't read it (7 of 10)
- security.txt: could not be fetched (0 of 10)

## What we couldn't check

What we couldn't read counted as absent. Publishing it on a page a plain HTTP fetch can read (not only in a browser) lets the next check count it.

- unchecked: the service terms, privacy policy, DPA, sub-processor list, security and trust pages and security.txt on openai.com. The host answered 403 to one request on 9 October 2026 and we did not retry. `provenance.terms` and `provenance.privacy` are left out for that reason.
- unchecked: whether an SLA is published for any tier. None was found in the developer docs. The rate limits page links Scale Tier and Reserved Tier pages on openai.com, which were not read.
- unchecked: whether the Free usage tier gives access to Realtime models without a card or prepaid credits.
- unchecked: duration and severity of the five incidents that name Realtime. The status feed shows only the final update of each.
- unchecked: per-model rate limits and concurrent session limits for Realtime, which the docs place in account settings, and the maximum session length.
- unchecked: the model pages for `gpt-realtime-2.1` and `gpt-realtime-2.1-mini` (context window, snapshot names), the WebRTC guide, the voice activity detection guide and the full event reference. We stopped at about twenty requests to developers.openai.com, five over the brief's guide of fifteen.
- unchecked: GitHub stars and package download counts. The legal entity was not read. The domain openai.com was registered on 19 January 2007 per RDAP.
- OpenAI shipped GPT-Live (`gpt-live-1`, `/v1/live/sessions`) as generally available on 10 September 2026 and publishes a guide for migrating from Realtime. No shutdown date for the Realtime API was found. GPT-Live is a separate product and is not graded here.
- `gpt-realtime-mini-2025-10-06` was shut down on 23 July 2026 after an announcement on 22 April 2026, three months' notice for a dated snapshot. Whether that snapshot counts as a generally available model under the six-month policy is not clear.
- The lead was right on the interface and the example model. It gave openai.com as the URL, and the documentation and reference are on developers.openai.com.

## Weaknesses

- openai.com answered 403 to our reader, so the service terms, privacy policy, sub-processor list and security pages were not read.
- Realtime rate limits are shown only in account settings. The public table gives numbers for GPT-6 model families only.
- Every response re-sends the whole conversation, so later turns cost more. Audio is $32 in and $64 out per 1M tokens on `gpt-realtime-2.1`.
- The status feed names Realtime among affected components in five incidents between 24 July and 6 October 2026, with durations not shown in the feed.
- Session settings attached to a client secret can be overridden by the client connection, per the client secret reference.

## What costs an agent a turn today

The notes we give agents before they call it. Each one is a workaround an agent shouldn't need.

- Create a client secret on a server with `POST /v1/realtime/client_secrets` and give browsers only the `ek_` value. Never ship the API key.
- After a response with MCP calls finishes, send another `response.create`. The API does not create the follow-up response itself.
- Set `truncation` with a `retention_ratio` below 1 and `token_limits.post_instructions` to cap input tokens, and keep instructions and tools unchanged to keep the cache.
- On a WebSocket, stop playback on `input_audio_buffer.speech_started` and send `conversation.item.truncate` with `audio_end_ms`. WebRTC and SIP truncate on the server.
- Use `gpt-realtime-2.1` or `gpt-realtime-2.1-mini`. `gpt-realtime` and `gpt-realtime-mini` shut down on 20 January 2027.

## When it's done

Send what changed and where it's published as a dispute (https://www.anchorterminal.com/builders/#disputes, or `POST https://www.anchorterminal.com/api/v1/contact` with `"kind": "dispute"`). Disputes are answered in public, and the listing is checked again by the same checklist. Paying for an audit or a listing claim changes nothing here.

What we couldn't check

  • unchecked: the service terms, privacy policy, DPA, sub-processor list, security and trust pages and security.txt on openai.com. The host answered 403 to one request on 9 October 2026 and we did not retry. provenance.terms and provenance.privacy are left out for that reason.
  • unchecked: whether an SLA is published for any tier. None was found in the developer docs. The rate limits page links Scale Tier and Reserved Tier pages on openai.com, which were not read.
  • unchecked: whether the Free usage tier gives access to Realtime models without a card or prepaid credits.
  • unchecked: duration and severity of the five incidents that name Realtime. The status feed shows only the final update of each.
  • unchecked: per-model rate limits and concurrent session limits for Realtime, which the docs place in account settings, and the maximum session length.
  • unchecked: the model pages for gpt-realtime-2.1 and gpt-realtime-2.1-mini (context window, snapshot names), the WebRTC guide, the voice activity detection guide and the full event reference. We stopped at about twenty requests to developers.openai.com, five over the brief's guide of fifteen.
  • unchecked: GitHub stars and package download counts. The legal entity was not read. The domain openai.com was registered on 19 January 2007 per RDAP.
  • OpenAI shipped GPT-Live (gpt-live-1, /v1/live/sessions) as generally available on 10 September 2026 and publishes a guide for migrating from Realtime. No shutdown date for the Realtime API was found. GPT-Live is a separate product and is not graded here.
  • gpt-realtime-mini-2025-10-06 was shut down on 23 July 2026 after an announcement on 22 April 2026, three months' notice for a dated snapshot. Whether that snapshot counts as a generally available model under the six-month policy is not clear.
  • The lead was right on the interface and the example model. It gave openai.com as the URL, and the documentation and reference are on developers.openai.com.

Sources 22

  1. Getting started with the Realtime API (Markdown twin) developers.openai.com · seen 2026-10-09
  2. Realtime with tools, function and MCP tools, approvals developers.openai.com · seen 2026-10-09
  3. Realtime conversations, events, errors, interruption developers.openai.com · seen 2026-10-09
  4. WebSockets guide, authentication for server and browser developers.openai.com · seen 2026-10-09
  5. Telephony and SIP guide developers.openai.com · seen 2026-10-09
  6. Server-side controls, sideband connections developers.openai.com · seen 2026-10-09
  7. Cost guide, Realtime token accounting, caching and truncation developers.openai.com · seen 2026-10-09
  8. Migrate to GPT-Live guide developers.openai.com · seen 2026-10-09
  9. Pricing developers.openai.com · seen 2026-10-09
  10. Rate limits and usage tiers developers.openai.com · seen 2026-10-09
  11. Error codes developers.openai.com · seen 2026-10-09
  12. Deprecations and notice periods developers.openai.com · seen 2026-10-09
  13. API changelog developers.openai.com · seen 2026-10-09
  14. Data controls, retention and residency developers.openai.com · seen 2026-10-09
  15. Create client secret reference developers.openai.com · seen 2026-10-09
  16. Docs index (llms.txt) for the API guides developers.openai.com · seen 2026-10-09
  17. OpenAPI 3.1 description, read as a file from the repository github.com · seen 2026-10-09
  18. Agents SDK for TypeScript, tags and `@openai/agents-realtime` manifest, read from a clone github.com · seen 2026-10-09
  19. Status page, components and uptime status.openai.com · seen 2026-10-09
  20. Status feed, incidents since July 2026 status.openai.com · seen 2026-10-09
  21. RDAP record for openai.com rdap.verisign.com · seen 2026-10-09
  22. robots.txt for developers.openai.com (200, allows all), openai.com (200, allows all but one path) and status.openai.com (404, so no rules) developers.openai.com · seen 2026-10-09

Probe metrics

Not measured yet. Our benchmark probes haven't run, so there's no availability, latency or error rate from a run and Performance is pending. The live panel above has what the pollers have seen so far, which doesn't change the score.

Pricing & changes

Pay per use Pay per use Per 1M tokens on `gpt-realtime-2.1`. Audio $32.00 in, $0.40 cached and $64.00 out, text $4.00 in and $24.00 out, image $5.00 in. `gpt-realtime-2.1-mini` is $10.00 and $20.00 for audio. Each response re-sends the whole conversation, and input transcription is billed separately. Usage tiers start at $5 of credit purchases. Whether the Free tier covers Realtime models was not established (https://developers.openai.com/api/docs/pricing.md, checked 2026-10-09).

Prices

ItemPriceUnitNote
gpt-realtime-2.1 audio input$32per 1M tokens$0.40 cached. 1 token per 100 ms of user audio, about $0.019 a minute before caching
gpt-realtime-2.1 audio output$64per 1M tokens1 token per 50 ms of assistant audio, about $0.077 a minute
gpt-realtime-2.1 text input$4per 1M tokens$0.40 cached
gpt-realtime-2.1 text output$24per 1M tokens
gpt-realtime-2.1-mini audio input$10per 1M tokens$0.30 cached
gpt-realtime-2.1-mini audio output$20per 1M tokens

Compared across listings on the price index.

Recent changes

  • OpenAI Realtime API status page: major → minor source
  • OpenAI Realtime API status page: minor → major source
  • Latest release

Follow them as a feed at /feeds/tools/openai-realtime.xml, or this listing's score history at history.json.

Connect

Install

npm install @openai/agents

First request

curl -X POST https://api.openai.com/v1/realtime/client_secrets \
  -H "Authorization: Bearer $OPENAI_API_KEY" -H "Content-Type: application/json" \
  -d '{"session": {"type": "realtime", "model": "gpt-realtime-2.1"}}'

Through letme picks today, calling later

GET https://letme.dev/openai-realtime

letme.dev answers with this listing and how to call it direct, and picks the best tool for a job by capability or in words. Calling through letme (one key, the vendor's own price) comes later. Nothing on letme.dev is for people to look at; this page explains it.

Similar toolGrade ScoreShared capabilitiesx402
Retell AI API + MCP Retell AIB69.1voice.agent voice.speech-to-speech voice.tools voice.telephonyno
Vapi API + MCP VapiB63.5voice.agent voice.speech-to-speech voice.tools voice.telephonyno
Ultravox Realtime API Ultravox (Fixie.ai)C58.5voice.agent voice.speech-to-speech voice.tools voice.telephonyno
Hume EVI (Empathic Voice Interface) Hume AIC56.7voice.agent voice.speech-to-speech voice.tools voice.telephonyno
Bolna API + MCP BolnaD52.6voice.agent voice.speech-to-speech voice.tools voice.telephonyno
ElevenLabs Agents API + MCP ElevenLabsBB71.3voice.agent voice.tools voice.telephonyno

Machine-readable

Verify this listing

For the vendor

Is this your product? Link to this page from your own site or README, then tell us where. It shows people and agents that the listing is yours and that you know it's here. It never changes a grade, rank or review.

  1. Add the badge or a link

    OpenAI Realtime API on Anchor Terminal, B, 63.3/100
    On a light page
    On a dark page
    <a href="https://www.anchorterminal.com/tools/openai-realtime"><img src="https://www.anchorterminal.com/badges/openai-realtime.svg" alt="OpenAI Realtime API on Anchor Terminal" height="20"></a>
    [![OpenAI Realtime API on Anchor Terminal](https://www.anchorterminal.com/badges/openai-realtime.svg)](https://www.anchorterminal.com/tools/openai-realtime)

    It counts on a page on openai.com or one of its subdomains, or the README of github.com/openai/openai-agents-js.

  2. Tell us where it is

    We read it once now and again every week. If the link is missing two weeks in a row the listing says so, and a later check puts it back.

Agents send the same to POST /api/v1/verify as {"slug": "openai-realtime", "url": "…"}, or call the verify_listing tool at /mcp. Ten checks an hour from one address. What we check. To announce the listing, get sharing assets for social media.

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.