OpenAI's low-latency streaming speech-to-text model. Served only through the gateway's /v1/realtime WebSocket endpoint as a transcription session (intent=transcription) or as the input-audio transcription model of a speech-to-speech session, billed per minute of audio.
LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.
OpenAI's low-latency streaming speech-to-text model. Served only through the gateway's /v1/realtime WebSocket endpoint as a transcription session (intent=transcription) or as the input-audio transcription model of a speech-to-speech session, billed per minute of audio. You can access it through LLM Gateway's OpenAI-compatible API with automatic provider routing, fallback, and cost analytics.
Pricing for GPT Live Transcribe on LLM Gateway starts at $0.00 per million input tokens and $0.00 per million output tokens, depending on the provider. The pricing table above always reflects the current per-provider rates.
GPT Live Transcribe is served by OpenAI through LLM Gateway. Requests are automatically routed to the best available provider, with fallback when a provider has issues.
No. GPT Live Transcribe does not currently support tool calling or structured JSON outputs through LLM Gateway.
GPT Live Transcribe was released on July 28, 2026.
GPT Image 2.5 Flare
$5.00 in / $0.00 out per 1M
GPT Image 2.5 Sunburst
$5.00 in / $0.00 out per 1M
GPT-6 Astra
1,050,000 context$10.00 in / $50.00 out per 1M
GPT Transcribe
$0.00 in / $0.00 out per 1M
GPT-5.6 Luna
1,050,000 context$0.20 in / $1.20 out per 1M
GPT-5.6 Terra
1,050,000 context$2.00 in / $12.00 out per 1M