Cost optimization
Every render is priced against the cheapest host serving that model at your requested resolution and duration — automatically, not by hand.
Access Sora, Kling, Veo, Minimax, and 15+ video models through a single interface. VideoRouter provides a unified video generation API for developers.
One flat 2% platform fee — vs. 5% on OpenRouter.
The OpenRouter alternative for video & image models
The same video & image model, priced by every host we route to — see how much cheaper our best price is than Fal.ai's own rate for each one.
| Model | Cheapest | Fal price | Discount (compared to Fal.ai price) |
|---|---|---|---|
| Loading live rates… | |||
Install the VideoRouter node pack and bill Kling, MiniMax, Veo, and 20+ more Partner Nodes through your credits — no comfy.org account needed.
comfyui-videorouter node pack
20+ Partner Nodes for image & video generation, billed through your VideoRouter credits.
The architecture
One endpoint in. Every model, max uptime. We shop Fal, WaveSpeedAI, Atlas Cloud, Replicate, Novita, and more for the best price on every model, every request — and health-check every host continuously so your app fails over automatically and stays online.
Your application
VideoRouter
Inference Optimization Layer
The full layer
Everything below runs automatically, on every request, under a single flat 2% fee.
Price, latency, and quality for every provider — in one place, live — so you never have to guess who's cheapest or fastest for the model you want.
Every render is priced against the cheapest host serving that model at your requested resolution and duration — automatically, not by hand.
Set budgets and limits per key, project, or team, and get alerted before spend runs away — not after the invoice.
One API for every video & image model — Fal, WaveSpeedAI, Atlas Cloud, Replicate, Novita, and closed hosts like Sora, Veo, and Kling. Swap models with a single string change.
If a host degrades or rate-limits mid-render, we fail over transparently — your app never sees the outage.
Every provider hosting a given model, side by side — live price, measured latency, and benchmark quality — so you can see who's actually cheapest and fastest, not guess.
We track latency and throughput per host and route toward whoever is fastest for your render right now.
Quickstart
Point your base URL at VideoRouter and call any open-source or closed video/image model directly — same request format, one API key, a flat 2% fee.
# pip install openai
from openai import OpenAI
client = OpenAI(
base_url="https://videorouter.sh/api/v1",
api_key="llmr_sk_...",
)
# Any video model, from any host, one API key
video = client.videos.create(
model="wan-2.2",
prompt="A drone shot over a neon-lit city at night",
seconds=5,
)
print(video.id)
# → routed to whichever host has Wan 2.2 cheapest right now
From the blog
Real cross-provider price data on the models you're actually calling — Seedance, Wan, Veo, and more.
Create a developer account, grab an API key, and point your OpenAI SDK at us — every video & image host, price, latency, and failover, optimized automatically.
OpenAI-compatible · one API key · 2% fee.