All models / gpt-realtime-2.1-mini / Microsoft
gpt-realtime-2.1-mini
model id: openai/gpt-realtime-2.1-mini/microsoft
Audio in → Audio out
Price
$10 / 1M audio input tokens
per 1M tokens, by modality — billed by the session's real usage across every response.done event, not a flat per-minute or per-call rate
| Token type | Input | Cached input | Output |
|---|---|---|---|
| Text | $0.6 | $0.06 | $2.4 |
| Audio | $10 | $0.3 | $20 |
| Image | $0.8 | $0.3 | $0 |
+ VideoRouter platform fee: 2% of provider cost — billed as its own line, never blended into the rates above.
import asyncio, json, websockets
async def main():
async with websockets.connect(
"wss://videorouter.sh/v1/realtime?model=azure/gpt-realtime-2.1-mini",
additional_headers={"Authorization": "Bearer llmr_sk_..."},
) as ws:
print(json.loads(await ws.recv())) # -> {"type": "session.created", ...}
await ws.send(json.dumps({"type": "session.update",
"session": {"type": "realtime", "instructions": "You are a helpful assistant."}}))
# ... input_audio_buffer.append chunks, then response.create ...
await ws.send(json.dumps({"type": "response.create"}))
print(json.loads(await ws.recv())) # -> response.* stream, ending in response.done + usage
asyncio.run(main())
Customer reviews
Specific to the Microsoft hosting of gpt-realtime-2.1-mini — other providers serving the same model have their own reviews.
Log in to write a review of gpt-realtime-2.1-mini.
★★★★★
No reviews yet — be the first to share how this model performed for you.