All models / gpt-realtime-2 / OpenAI
gpt-realtime-2
model id: openai/gpt-realtime-2/openai
Audio in → Audio out
Price
$32 / 1M audio input tokens
per 1M tokens, by modality — billed by the session's real usage across every response.done event, not a flat per-minute or per-call rate
| Token type | Input | Cached input | Output |
|---|---|---|---|
| Text | $4 | $0.4 | $24 |
| Audio | $32 | $0.4 | $64 |
| Image | $5 | $0.5 | $0 |
+ VideoRouter platform fee: 2% of provider cost — billed as its own line, never blended into the rates above.
import asyncio, json, websockets
async def main():
async with websockets.connect(
"wss://videorouter.sh/v1/realtime?model=gpt-realtime-2",
additional_headers={"Authorization": "Bearer llmr_sk_..."},
) as ws:
print(json.loads(await ws.recv())) # -> {"type": "session.created", ...}
await ws.send(json.dumps({"type": "session.update",
"session": {"type": "realtime", "instructions": "You are a helpful assistant."}}))
# ... input_audio_buffer.append chunks, then response.create ...
await ws.send(json.dumps({"type": "response.create"}))
print(json.loads(await ws.recv())) # -> response.* stream, ending in response.done + usage
asyncio.run(main())
Customer reviews
Specific to the OpenAI hosting of gpt-realtime-2 — other providers serving the same model have their own reviews.
Log in to write a review of gpt-realtime-2.
★★★★★
No reviews yet — be the first to share how this model performed for you.