Compare realtime APIs

Pick up to four APIs from any category and see them side by side.

Hume EVI (Empathic Voice Interface)
Hume AI
Deepgram Voice Agent API
Deepgram
OpenAI GPT-Live API
OpenAI
CategoryVoice-to-voiceVoice-to-voiceVoice-to-voice
StatusDeprecatedGAGA
Est. per minute$0.04 - 0.07$0.041 - 0.163$0.05
How that was worked outThird-party reported EVI 3 overage rates; EVI 4-mini reportedly about half. Supplemental LLM cost may be extra.Official tier rates; BYO tiers add your own LLM/TTS bills on top.Low = voice layer only, $0.05 per minute of session duration with no delegated work. High depends entirely on the backend model, how often the model delegates and tool fees; budget per-minute session cost plus backend token spend measured in your own tests.
Pricing modelsubscriptionper-minuteper-minute
Free tier5 EVI minutes per month (third-party data).$200 one-time credit for new accounts.Free tier not supported.
Connects byWebSocketWebSocketWebRTC, WebSocket, SIP
Audio inWebM (browser) or linear16 PCM (e.g. 44.1 kHz mono, declared in session_settings); no mu-lawlinear16 default at 16 kHz (other encodings/rates configurable)WebSocket: audio/pcm 24 kHz (default) or 16 kHz mono 16-bit LE, audio/pcmu or audio/pcma 8 kHz; raw bytes base64, even byte length, no container. Format fixed at session start. WebRTC and SIP negotiate codecs (outbound SIP needs Opus + SDES-SRTP).
Audio outbase64 WAV in audio_output messageslinear16 (e.g. 24 kHz), optional container; other encodings configurableSame format as input (one setting covers both).
LanguagesEVI 3: English. EVI 4-mini: English, Japanese, Korean, Spanish, French, Portuguese, Italian, German, Russian, Hindi, Arabic.English by default; multilingual via Nova 'multi' or flux-general-multi with language hints, and multilingual TTS providers.Not listed on the model page.
Latency (vendor claim)No figure on the overview page.Not stated on the pages read; the API reports per-turn latency metrics (time to first token, TTS time to first byte).Vendor claim: improves Full Duplex Bench score by 30 percentage points over gpt-realtime-2.1 (reported via third-party coverage). No millisecond figure published.
Key limits
  • Access ends November 13, 2026 at 12:01 a.m. EST; account data permanently deleted after that
  • Max session duration 30 minutes
  • Max WebSocket message 16 MB
  • HTTP rate limit 100 requests per second
  • Concurrency: up to 45 WebSocket agent sessions on Pay As You Go, up to 60 on Growth
  • Billing runs for the whole WebSocket connection
  • Concurrent sessions (OpenAI direct): Build 50, Launch 300, Grow 500; Free unsupported
  • Context window 128,000 tokens; above 90 percent usage a replacement engine starts with up to 8,192 tokens of history, so older details may be summarised or dropped
  • Instructions up to 16,384 tokens; startup history up to 128 messages / 8,192 tokens
  • Session ends with reason expired at a duration limit that the docs do not state
High-severity warnings
  • Shutting down November 13, 2026
  • Export data now
  • Model choice changes the tier
  • Connection time is billed
  • Two bills, not one
  • Billed by wall-clock duration
ComplianceNot re-verified; irrelevant after shutdown.Not re-verified this session (Deepgram markets SOC 2 and HIPAA readiness; check the trust page)./v1/live/sessions is ZDR eligible with limitations (store forced false, no forking or recording download). Abuse-monitoring logs 30 days. US and EU data residency.
Self-hostableNoNoNo
Last checked2026-10-102026-10-102026-10-10
Key numbers and features
Flat $/min$0.07$0.075$0.05
Audio in $/1M tok---
Audio out $/1M tok---
Free tierYesYesNo
Free credit $-$200-
Native S2SYesNoYes
Tools-YesYes
Image in--No
Own LLMYesYesYes
Voices--12
CloningYes--
Latency ms---
Languages11--
Context tokens--128,000
Max session min30--
Concurrency-4550
WebRTCNoNoYes
WebSocketYesYesYes
Phone / SIPNoNoYes
HIPAA---
SOC 2---
EU data--Yes
Open weightsNoNoNo
High warnings222