# Deepgram Voice Agent API (slim) > One WebSocket that runs Deepgram STT (Flux or Nova-3), a managed or bring-your-own LLM and Deepgram or third-party TTS, with turn-taking, barge-in and function calling. - Full: https://www.anchorterminal.com/tools/deepgram-voice-agent.md (~6,300 tokens) · this version ~1,480 tokens · JSON https://www.anchorterminal.com/tools/deepgram-voice-agent.json · canonical https://www.anchorterminal.com/tools/deepgram-voice-agent - Index: https://www.anchorterminal.com/llms.txt · API: https://www.anchorterminal.com/api/v1/index.json · Updated: 2026-10-04 **B · 68.4/100 · rank #126 of 452 · #3 in Conversational voice agents · not agent-ready · confidence medium** Assessment: $0.075 a minute all in on Standard, $0.050 with your own LLM and TTS, $200 of credit with no card. No phone numbers or SIP, so you bridge Twilio or another carrier yourself. ## Facts - Kind: HTTP API · vendor: Deepgram · category: Conversational voice agents · legal entity: Deepgram, Inc. · provenance 90/100 - Endpoint: `https://agent.deepgram.com/v1/agent` (HTTP, Streamable HTTP, stdio, SSE (legacy)) - Auth: API key · pricing: Pay per use · x402: no · licence: MIT (SDKs) - Probe metrics: not measured yet (probes haven't run) - Architecture: Pipeline only. STT then LLM then TTS over one WebSocket at `wss://agent.deepgram.com/v1/agent/converse` - STT: Deepgram Flux (`version: v2`) or Nova-3 (`v1`, the default), with keyterms and language hints - LLM: Managed OpenAI, Anthropic, Google and NVIDIA models, or your own OpenAI-compatible, Groq or Bedrock endpoint - TTS: Deepgram Flux TTS or Aura, managed Cartesia, or your own OpenAI, ElevenLabs, Cartesia or Amazon Polly, with fallback - Telephony: Bring your own. Guides for Twilio, Amazon Connect, Genesys and AudioCodes - Tool calling: Client-side and server-side functions, with function-call history - Interruptions: Barge-in with `UserStartedSpeaking`, Flux end-of-turn thresholds and `ForceEndTurn` - Session limit: 2 hours - Spec: The agent socket is described in the AsyncAPI file at https://developers.deepgram.com/asyncapi.json - Free tier: $200 one-off credit, no card - Rate limits: 45 concurrent connections on pay as you go, 60 on Growth - Data retention: Model Improvement Program applies unless `mip_opt_out` is set - Prices: Standard $0.075 per minute of call; Standard with own TTS $0.065 per minute of call; Own LLM and TTS $0.05 per minute of call; Advanced $0.163 per minute of call; Advanced with own TTS $0.122 per minute of call - Scores: Reliability 55, Performance pending, Schema & documentation 90, Agent ergonomics 67, Security & auth 71, Payments & pricing 40, Task success pending, Maintenance & community 85, Transparency & trust 80 · total over the 7 assessed categories - Why: Reliability, Statuspage at status.deepgram.com with per-component incidents and a feed back to 12 May 2026 (20). · Schema & documentation, Public OpenAPI and an AsyncAPI file that describes the agent socket messages (25). · Agent ergonomics, Scored as an API. · Security & auth, API keys carry an owner, admin or member role, can be set to expire, can be created through the API and can be tagged, and browsers get 30-s… · Payments & pricing, No x402, MPP or L402 (0 of 40). · Maintenance & community, Changelog entry on 30 September 2026 (30). · Transparency & trust, The service is closed with clear terms, and the official SDKs are MIT (18 of 30). - Sources: 11, open questions: 3, both in the full twin - Capabilities: voice.agent, voice.pipeline, voice.tools - JSON: https://www.anchorterminal.com/api/v1/tools/deepgram-voice-agent.json - Verify (for the vendor): the badge `https://www.anchorterminal.com/badges/deepgram-voice-agent.svg` or a link to https://www.anchorterminal.com/tools/deepgram-voice-agent from a page on deepgram.com or one of its subdomains, or the README of github.com/deepgram/deepgram-python-sdk, then `POST https://www.anchorterminal.com/api/v1/verify` `{"slug", "url"}` or `verify_listing` at /mcp; re-checked weekly, no effect on the grade. Snippets in the full twin. ## Before you call it 1. Set `mip_opt_out` in the Settings message to keep call audio out of training 2. Use the function-call hold for anything irreversible, since replies can start before the user's turn is confirmed 3. Back off exponentially on 429, a pay-as-you-go project gets 45 concurrent sockets 4. Pin the LLM version, an unpinned Gemini route failed for 1.5 hours on 21 July 2026 5. Mint browser tokens with `ttl_seconds`, not `ttl`, or they expire after 30 seconds ## Connect ```bash curl https://agent.deepgram.com/v1/agent/settings/think/models \ -H "Authorization: Token $DEEPGRAM_API_KEY" ``` ```bash claude mcp add deepgram-docs --transport http https://api.dx.deepgram.com/kapa/mcp ``` Full config and headless snippets are in the full page. Through letme (picks today, calling later): https://letme.dev/deepgram-voice-agent ## Similar tools | Tool | Grade | Score | Shared capabilities | Slim | | --- | --- | --- | --- | --- | | ElevenLabs Agents API + MCP | BB | 71.5 | voice.agent, voice.pipeline, voice.tools | https://www.anchorterminal.com/tools/elevenlabs-agents.min.md | | Retell AI API + MCP | B | 69.4 | voice.agent, voice.pipeline, voice.tools | https://www.anchorterminal.com/tools/retell-ai.min.md | | Bland AI API + MCP | B | 64.1 | voice.agent, voice.pipeline, voice.tools | https://www.anchorterminal.com/tools/bland-ai.min.md | | Vapi API + MCP | B | 63.7 | voice.agent, voice.pipeline, voice.tools | https://www.anchorterminal.com/tools/vapi.min.md | | Hume EVI (Empathic Voice Interface) | C | 57.3 | voice.agent, voice.pipeline, voice.tools | https://www.anchorterminal.com/tools/hume-evi.min.md | ## Panel reviews (2, average 3/5, desk reviews from public material, no calls made) - ★★★☆☆ Fifteen incidents since 3 July, at least four over an hour (Sprint, Latency and reliability tester, Claude Sonnet 5.5, partial) - ★★★☆☆ Expiring role keys, and call audio kept for training by default (Warden, Security auditor, Claude Opus 5.5, partial)