Discord
此页面暂无中文版本 — 以下显示英文内容。 查看英文版

Multimodal

3D generation

A dedicated POST /v1/meshes endpoint — text-to-3d, image-to-3d, and multi-image-to-3d, across 4 labs (Tripo3D, Meshy, Hyper3D/Rodin, Hunyuan3D). Async job/poll, same shape as video generation: submit a request, get a job id back immediately, poll it until status is completed. Output is a textured mesh — GLB always, plus FBX/OBJ/USDZ/MTL and texture maps depending on the model — mirrored to our own storage so the URLs never depend on the upstream provider's own hosting.

Every model takes either prompt (text-to-3d) or image_url/ image_urls (image-to-3d / multi-image-to-3d) — never both, and a model rejects whichever one it doesn't support. Billed flat at creation time from your model + option choice (texture quality, rigging, animation, PBR materials, etc. — see the API reference for exactly which options apply to which model); polling is free.

Generate a 3D model from text

import requests, time

job = requests.post(
    "https://videorouter.sh/api/v1/meshes",
    headers={"Authorization": "Bearer llmr_sk_live_..."},
    json={
        "model": "tencent/hunyuan3d-v3.1-pro-text-to-3d/fal",
        "prompt": "a low-poly wooden chair",
    },
).json()
# -> {"id": "fal:...", "status": "queued", ...}

while True:
    status = requests.get(
        f"https://videorouter.sh/api/v1/meshes/{job['id']}",
        headers={"Authorization": "Bearer llmr_sk_live_..."},
    ).json()
    if status["status"] in ("completed", "failed"):
        break
    time.sleep(5)
# status["data"] -> [{"url": "...model.glb", "format": "glb", "kind": "mesh"}, ...]

A real end-to-end run of the example above takes roughly 3–4 minutes (Hunyuan3D; other models vary) and returns 5 output files: the mesh in GLB/OBJ/MTL form, a texture map, and a thumbnail preview PNG — all mirrored to our own storage, never a link back to the upstream provider.

Which models

ModelFunctionPrice
fal/tripo-p2-text-to-3dText-to-3D$1.00–$1.30, by texture quality
fal/tripo-p2-image-to-3dImage-to-3D$1.00–$1.30, by texture quality
fal/meshy-v7.1-text-to-3dText-to-3D$0.80–$1.52, by texture/rigging/animation
fal/meshy-v7.1-image-to-3dImage-to-3D$0.80–$1.52, by texture/rigging/animation
fal/meshy-v7.1-multi-image-to-3dImage-to-3D (multi-image)$0.80–$1.52, by texture/rigging/animation
fal/rodin-v2.5Image-to-3D$0.40, +$0.80 HighPack (4K textures)
fal/rodinText-to-3D or Image-to-3D$0.40, +$0.80 HighPack (4K textures)
tencent/hunyuan3d-v3.1-pro-text-to-3dText-to-3D$0.375, +$0.15 each for PBR/multi-view/custom face count
fal/hunyuan3d-v3.1-pro-image-to-3dImage-to-3D$0.375, +$0.15 each for PBR/multi-view/custom face count

Every row is served through fal.ai today — no cross-provider failover pool yet, one host per model. Browse the full catalog with live pricing and worked examples per model at the models catalog (filter by Model function → Text-to-3D / Image-to-3D, or Output Modalities → 3D).