Lowest-cost media generation API
Transparent, pay-as-you-go pricing across image, video, audio and text. Search the full catalog below and see the exact cost for every configuration — no monthly fees, no surprises.
Showing 169 models
| Model | Configuration | Price | Actions |
|---|---|---|---|
Qwen3.8-Max alibaba | Text → Text · Input tokens | $2 per 1M tokens | |
Fast Image Resizer grey-hound432 | Image → Image · Keep aspect ratio: true | $0.00045 per second | |
LatentSync 1.0 bytedance | Video → Video | $0.014 per second | |
BiRefNet Background Removal zhengpeng7 | Image → Image · Variant: general | $0.002 per second | |
Kokoro-82M hexgrad | Text → Audio | $0.000225 per second | |
Rodin Gen-2 by Hyper3D hyper3d | Text/Image | $0.399 per output | |
Wan 3.0 wan-video | Text/Image → Video · 480P | $0.05 per second | |
Qwen-Image 3.0 alibaba | Text → Image | $0.03 per output | |
Happy Horse 1.1 alibaba | Text/Image → Video · 720P | $0.14 per second | |
Gemini 3.1 Flash TTS google | Text → Audio | $0.0006 per second | |
| Text → Text · Input tokens | $2.8 per 1M tokens | ||
DeepSeek V4 Pro deepseek | Text → Text · Input tokens | $2.4 per 1M tokens | |
GLM-5.2 z-ai | Text → Text · Input tokens | $1.4 per 1M tokens | |
Wan 2.7 Video Extend wan-video | Video → Video · 720P | $0.1 per second | |
| Image → Video · 720P | $0.1 per second | ||
Wan 2.7 wan-video | Text/Image → Video · 720P | $0.1 per second | |
Gemini 3.5 Flash-Lite google | Text → Text · Input tokens | $0.3 per 1M tokens | |
Gemini 3.7 Flash google | Text → Text · Input tokens | $0.75 per 1M tokens | |
Gemini 3.5 Flash google | Text → Text · Input tokens | $1.5 per 1M tokens | |
Seedance 2.5 bytedance | Text/Image → Video · 480p | $0.154 per second | |
Seedance 2.0 Mini bytedance | Text/Image → Video · 480p | $0.053 per second | |
Seedance 2.0 Fast bytedance | Text/Image → Video · 480p | $0.084 per second | |
Seedance 2.0 bytedance | Text/Image → Video · 480p | $0.105 per second | |
Lyria 3 Clip google | Text → Audio | $0.04 per output | |
| Text → Image | $0.034 per output |
Qwen3.8-Max
alibaba
- Text → Text · Input tokens
- $2per 1M tokens
- Text → Text · Cached input
- $0.25per 1M tokens
- Text → Text · Output tokens
- $6per 1M tokens
Fast Image Resizer
grey-hound432
- Image → Image · Keep aspect ratio: true
- $0.00045per second
LatentSync 1.0
bytedance
- Video → Video
- $0.014per second
BiRefNet Background Removal
zhengpeng7
- Image → Image · Variant: general
- $0.002per second
Kokoro-82M
hexgrad
- Text → Audio
- $0.000225per second
Rodin Gen-2 by Hyper3D
hyper3d
- Text/Image
- $0.399per output
- Text/Image → Video · 480P
- $0.05per second
- Text/Image → Video · 720P
- $0.1per second
- Text/Image → Video · 1080P
- $0.2per second
Qwen-Image 3.0
alibaba
- Text → Image
- $0.03per output
Happy Horse 1.1
alibaba
- Text/Image → Video · 720P
- $0.14per second
- Text/Image → Video · 1080P
- $0.18per second
- Text → Audio
- $0.0006per second
- Text → Text · Input tokens
- $2.8per 1M tokens
- Text → Text · Cached input
- $0.7per 1M tokens
- Text → Text · Output tokens
- $8.8per 1M tokens
DeepSeek V4 Pro
deepseek
- Text → Text · Input tokens
- $2.4per 1M tokens
- Text → Text · Cached input
- $0.2per 1M tokens
- Text → Text · Output tokens
- $4.8per 1M tokens
- Text → Text · Input tokens
- $1.4per 1M tokens
- Text → Text · Cached input
- $0.35per 1M tokens
- Text → Text · Output tokens
- $4.4per 1M tokens
Wan 2.7 Video Extend
wan-video
- Video → Video · 720P
- $0.1per second
- Video → Video · 1080P
- $0.15per second
- Image → Video · 720P
- $0.1per second
- Image → Video · 1080P
- $0.15per second
- Text/Image → Video · 720P
- $0.1per second
- Text/Image → Video · 1080P
- $0.15per second
- Text → Text · Input tokens
- $0.3per 1M tokens
- Text → Text · Cached input
- $0.03per 1M tokens
- Text → Text · Output tokens
- $2.5per 1M tokens
Gemini 3.7 Flash
google
- Text → Text · Input tokens
- $0.75per 1M tokens
- Text → Text · Cached input
- $0.075per 1M tokens
- Text → Text · Output tokens
- $3.75per 1M tokens
Gemini 3.5 Flash
google
- Text → Text · Input tokens
- $1.5per 1M tokens
- Text → Text · Cached input
- $0.15per 1M tokens
- Text → Text · Output tokens
- $9per 1M tokens
Seedance 2.5
bytedance
- Text/Image → Video · 480p
- $0.154per second
- Text/Image → Video · 720p
- $0.347per second
- Video → Video
- ≈$2.07for a 5s 720p clipInput + output duration are billed · smallest run $0.56
Seedance 2.0 Mini
bytedance
- Text/Image → Video · 480p
- $0.053per second
- Text/Image → Video · 720p
- $0.113per second
Seedance 2.0 Fast
bytedance
- Text/Image → Video · 480p
- $0.084per second
- Text/Image → Video · 720p
- $0.181per second
Seedance 2.0
bytedance
- Text/Image → Video · 480p
- $0.105per second
- Text/Image → Video · 720p
- $0.227per second
- Text/Image → Video · 1080p
- $0.561per second
- Video → Video
- ≈$1.52for a 5s 720p clipInput + output duration are billed · smallest run $0.41
Lyria 3 Clip
google
- Text → Audio
- $0.04per output
- Text → Image
- $0.034per output
Page 1 of 5
Frequently Asked Questions
Everything you need to know about ModelRunner
How is ModelRunner different from other AI providers?
ModelRunner provides a unified API for generative AI models across image and video. Instead of managing multiple provider integrations, you access Google, ByteDance, and other providers through a single endpoint. This simplifies development, reduces integration overhead, and gives you flexibility to switch between providers or models without changing your code.
What models does ModelRunner support?
ModelRunner supports models across four categories: text-to-image, image-to-image, text-to-video, and image-to-video. We integrate with leading providers including Google (Veo 3.1, Imagen) and ByteDance (Seedance, Seedream), as well as open-source models like FLUX running on serverless compute. New models are added regularly, and you can test any model instantly in our Playground before integrating.
How does pricing work?
ModelRunner offers transparent, pay-as-you-go pricing with no hidden fees. Pricing varies by model type: per-second GPU time for serverless models, per-output for simple generation, or per-output-second for video (based on duration). All pricing is clearly displayed for each model, including GPU costs for serverless inference. You purchase credits via Stripe and only pay for what you use.
Can I try models before integrating?
Yes. Every model on ModelRunner has an interactive Playground where you can test generation with full parameter control. You can explore example runs to see real inputs and outputs, then copy ready-to-use code snippets for your integration. The Playground mirrors the exact API behavior, so what you see is what you get in production.
How do I integrate ModelRunner into my application?
Integration is straightforward: sign up, purchase credits, and create an API key. You can then make requests via our REST API or use our JavaScript SDK for a more streamlined experience. All models share a consistent interface—same authentication, same request/response patterns—so switching between models requires minimal code changes.
Is my data private and secure?
Yes. ModelRunner does not use your inputs or outputs for training. We support secure authentication including passkeys (WebAuthn), OAuth (GitHub, Google), and two-factor authentication, and your API keys are managed securely. Request inputs and outputs are stored so they appear in your dashboard history, and you control for how long: shorten retention on an individual request, set an account-wide default, opt out of storing payloads entirely with a single header, or delete a request's data — including the files it generated — whenever you want.
Can I use ModelRunner for commercial projects?
Yes. You can use ModelRunner-generated content for commercial purposes. However, licensing terms depend on the underlying model provider. We clearly document licensing for each model in our catalog. For open-source models, standard open-source licenses apply. For provider APIs like Google or ByteDance, their respective terms of service govern commercial use.
Does ModelRunner support enterprise workloads?
Yes. ModelRunner supports production applications requiring high throughput and reliability. Our async queue system handles requests at scale with real-time status updates via SSE. For enterprise needs like dedicated capacity, custom SLAs, or volume pricing, contact our sales team at modelrunner.ai/enterprise to discuss tailored solutions.
Do pro models ever use faster or cheaper models under the hood?
No. When you select a pro model, you always get that exact model—we never substitute it with a faster or cheaper alternative. Every request is processed by the model you chose, ensuring consistent quality and predictable results. This transparency is core to how ModelRunner operates: what you select is exactly what runs your generation.
Do you offer free credits for open-source projects?
Yes. Maintainers of active open-source projects can apply to the ModelRunner Open Source Program for free monthly API credits, from $25 for a small project up to $400+ a month for ecosystem-critical infrastructure. Credits land on your normal account balance and are spent at the same published per-model rates as any other account. Apply at modelrunner.ai/oss-program.

