Discover curated collections of AI models for every use case
Transfer motion and expression from a driving video onto a character image. Compare AI motion control and motion transfer models by price, and run them via one API.
Kling 3.0 Motion Control
kuaishou
Transfer motion and facial expression from a driving video onto a character image, animating your character to perform the reference movements.
Wan 2.2 Animate Move
wan-video
Motion control for a character image: transfer the motion and facial expressions of a driving video onto your character — no text prompt needed, just a photo and a reference clip.
Wan VACE Video Edit
Edit an existing video from a text prompt — restyle scenes, swap subjects or backgrounds, and apply reference-image-guided changes while preserving the source motion.
OmniHuman 1.5
bytedance
Animate a still photo of a person to speak and move in sync with an audio track, producing a natural talking-head video.
Find the best AI image generator in 2026. Compare top AI image generation and editing models side by side — see real outputs, pricing per image, speed benchmarks, and API integration options.
Seedream V5 Text to Image
Generate high-quality, intelligent images from text prompts using the fast Lite version of Seedream 5.0.
Seedream V4.5 Text to Image
A next-generation text-to-image model by ByteDance, capable of high-fidelity generation, precise text rendering, and complex stylistic control for highly detailed visual compositions.
Nano Banana 2 Text to Image
google
Nano Banana 2 is a fast and versatile text-to-image model. It excels at creating high-quality images, from photorealistic scenes to complex infographics with accurate text, and can optionally use Google Search to generate content based on real-time information.
BitDance Text to Image
shallowdream204
Generate fast, high-resolution, photorealistic images from text prompts using an advanced autoregressive model for efficient, high-quality results.
Z-Image Turbo
tongyi-mai
High-speed 6B parameter text-to-image generation optimized for cost efficiency and volume. Produces up to 4MP images in an 8-step pipeline suitable for rapid prototyping.
Z-Image Base
Generate high-quality, stylistically diverse images with precise prompt adherence using the Z-Image foundation model.
SDXL Lightning 4-step
SDXL-Lightning is a lightning-fast text-to-image generation model that produces high-quality 1024px images in just a few steps, distilled from Stable Diffusion XL.
Seedream V5 Image Editing
Edit and seamlessly compose images using text prompts and multiple reference images with the fast, high-quality Seedream 5.0 Lite model.
Seedream V4.5 Image Editing
Advanced image editing model by ByteDance that uses text prompts and up to 10 reference images to stylize, transform, and seamlessly composite visuals.
Nano Banana 2 Image Editing
Edit images with text prompts. Make targeted changes like adding or removing objects, changing styles, or modifying specific elements while preserving the rest of the image.
Firered Image Edit Text to Image
fireredteam
An advanced image editing model that modifies images based on text prompts. It supports single-image edits and multi-image compositions for tasks like style transfer or virtual try-on.
Nano Banana
State of the art image editing model from Google Gemini 2.5.
Seedream v4
Seedream 4.0 is a next-generation image creation model that unifies generation and editing in a single architecture, enabling advanced multimodal reasoning and reference consistency while delivering stunning 4K images with significantly faster inference.
GPT Image 2
openai
Generate images from a text prompt, with precise instruction-following (counts, layout, multi-part requests) and accurate, legible text rendered inside the image.
Ideogram V4
ideogram
Generate images from a text prompt with industry-leading accurate, legible in-image text for logos, posters, and signage.
Luma Uni-1
luma
Generate a single high-quality image from a text prompt — a well-rounded default with optional manga styling, web-grounded references, and up to nine guidance images.
Luma Uni-1 Max
Generate a high-fidelity image from a text prompt, with optional manga styling, web-grounded references, and up to nine reference images to steer composition.
Krea 2 Large
krea
Generate one high-fidelity, photorealistic image from a text prompt, tuned for sharp detail and strong prompt adherence.
Recraft V3
recraft
Generate images from a text prompt across a large library of professional design styles, with true vector (SVG) output when a vector style is selected — ideal for logos, icons, and brand graphics.
FLUX.2 [dev]
black-forest-labs
Generate high-quality images from a text prompt with strong prompt adherence and efficient, fast inference.
Qwen-Image
qwen
Generate high-quality images from a text prompt, with standout accurate, legible in-image text in English and Chinese.
Qwen-Image-Edit
Edit an existing image from a text instruction — recolor, restyle, swap backgrounds, add or remove elements — with accurate, legible in-image text in English and Chinese.
Imagen 4
Generate high-quality, photorealistic images from a text prompt, with strong prompt adherence, improved in-image text rendering, and up to 2K resolution.
Imagen 4 Fast
Generate high-quality, photorealistic images from a text prompt fast and at the best price, with strong prompt adherence and improved in-image text rendering.
FLUX.1 Kontext [dev]
Edit an existing image from a text instruction — change objects, style, background, or text — while keeping the rest of the photo consistent.
GPT Image 2 Edit
Edit an existing image from a text instruction, with precise instruction-following and accurate, legible in-image text, using up to 16 reference images and an optional mask.
Ideogram Character
Generate new images of the same character from one reference photo and a text prompt, keeping facial features and distinctive traits consistent across scenes.
Edit images with AI — inpainting, outpainting, style transfer, background removal, and upscaling. Run top image-to-image AI models via API with full parameter control and per-request pricing.
Topaz Upscale Image
topazlabs
Upscale and enhance your images using a variety of powerful AI models. Increase resolution up to 4x, restore details, and use specialized modes for photos, CGI, and text, with optional face enhancement for professional-quality results.
Real Esrgan Image Upscaler
nightmareai
High-quality image upscaler with optional face enhancement
Inspyrenet Image Mask
swook
Helps find and highlight important objects in high-resolution images. It works without needing special high-quality training data and gives sharp, accurate results.
Ideogram V4 Image-to-Image
Transform a reference image with a text prompt — restyle, edit, or reinterpret it while rendering accurate, legible in-image text for logos, posters, and signage.
Bria GenFill V2
bria
Fill a masked region of your image with new content from a text instruction, with commercially safe outputs cleared for risk-free business use.
Z-Image Turbo Image to Image
Generate images from text and an initial image using Tongyi-MAI's super-fast Z-Image Turbo model.
FASHN Virtual Try-On v1.6
fashn
Dress a person in any garment from a product photo, generating realistic on-model fashion images at high resolution from a model photo and a garment photo.
Bria Background Remove
Remove the background from an image and return a transparent-PNG cutout of the subject, trained on fully licensed commercial data.
Bria Background Replace
Keep the foreground subject of a photo and generate a brand-new background from a text prompt (or match a reference image), with commercially safe outputs.
Bria Expand
Expand (outpaint) an image onto a larger canvas, generating new surroundings that match the original — with commercially safe outputs.
Bria Fibo Edit Colorize
Add realistic color to black-and-white photos, or apply a curated color treatment (vivid, sepia vintage), with commercially safe outputs.
Recraft Vectorize
Convert a raster image (PNG/JPG/WEBP logo, icon, or sketch) into a clean, infinitely scalable SVG vector file.
CodeFormer
sczhou
Restore and enhance blurry, low-resolution, compressed, or old face photos into a sharp, detailed image.
NAFNet Denoise
megvii-research
Remove noise and grain from a photo and restore its quality, returning a clean image at the input's native resolution.
IC-Light V2
lllyasviel
Relight an existing photo to a described lighting setup — change light direction, mood, and background while keeping the subject intact.
Segment Anything 2 (Auto-Segment)
meta
Automatically segment a photo into a combined object/region mask — no prompts, points, or clicks needed.
Animate still images into high-quality videos. Compare top image-to-video AI models by quality, resolution, audio support, and API pricing.
Seedance V1.5 Image to Video
Transform static images into dynamic videos with synchronized audio. Supports text-guided animations and start/end frame keying.
Veo 3.1 Image to Video
Turn static images into high-fidelity 720p or 1080p videos with synchronized native audio using text prompts to guide the animation.
Veo 3.1 First/Last Frame to Video
Generate seamless 8-second video transitions by interpolating between a first and last frame with high-fidelity visuals and native audio.
Veo 3.1 Reference to Video
Generate high-fidelity, cinematic videos with synchronized audio by using text prompts and up to three reference images to guide visual style and content.
LongCat-Video i2v
meituan-longcat
LongCat-Video turns a single still image into minutes-long, smooth 480p, 30fps video with stable style, lighting, and identity — fast, consistent, production-ready animation from one frame.
Grok Imagine 1.5 Image to Video
xai
Turn a still image into a short video clip guided by a motion prompt, with synchronized audio generated automatically and included free.
Seedance 1.0 Pro
Seedance 1.0 generates 1080P videos with smooth motion, rich detail, and diverse styles, while the pro version adds multi-shot narrative and advanced instruction following for cinematic results.
Kling 2.5 Turbo Pro Image-to-Video
Animate a still image into a short, cinematic video clip with fluid, natural motion guided by a text prompt.
Generate music, speech, sound effects, and transcriptions with AI. Compare top audio models for text-to-music, text-to-speech, and speech-to-text.
ACE-Step
ace-studio
Generate full songs or instrumental music from genre tags and optional lyrics, with duration you control up to 4 minutes.
meta / musicgen
A fast, controllable auto-regressive Transformer for high-fidelity music generation.
MiniMax Speech-02 HD
minimax
Turn text into natural, high-fidelity speech in 30+ languages with 300+ voices plus emotion, speed, pitch, and volume control.
LTX-2.3 Text-to-Audio
lightricks
Generate sound effects, ambience, and spoken-style audio from a text prompt, with duration you control down to the frame.
ElevenLabs Scribe v1
elevenlabs
Transcribe speech audio into accurate text with word-level timestamps, speaker labels, and audio-event tags across 99 languages.
Lyria 2
Generate ~30 seconds of high-fidelity instrumental music from a text prompt, as a 48kHz WAV file.
ElevenLabs Sound Effects V2
Generate sound effects, Foley, and ambience from a text prompt, returning a hosted MP3.
Whisper Large v3
Transcribe or translate speech audio into text across 99 languages, with segment/word timestamps and optional speaker diarization.
Chatterbox TTS
resemble-ai
Turn text into expressive speech and clone any voice from a short reference recording, with fine control over emotional intensity.
ElevenLabs Multilingual v2
Turn text into natural, expressive speech in 29 languages with ElevenLabs Multilingual v2 voices, with controls for stability, similarity, and style.
DeepFilterNet 3
rikorose
Clean up a noisy speech recording by removing background noise and upsampling it to studio-quality 48 kHz audio.
SAM Audio — Separate
Isolate any sound from an audio mixture by describing it in plain language.
ElevenLabs Audio Isolation
Strip background noise and music from a recording to isolate clean, studio-quality speech, returning the isolated voice as an MP3.
Enhance, upscale, lipsync, and extend existing videos with AI. Covers video upscaling, audio sync, and video extension.
Topaz Upscale Video
Enhance and enlarge your videos with professional-grade upscaling, detail recovery, and smooth frame interpolation.
Sync Lipsync 2
sync
Re-sync a talking-head video's mouth movements to a new audio track for dubbing, translation, and re-voicing.
Veo 3.1 Extend Video
Extend existing Veo-generated videos by seamlessly adding 7 seconds of high-fidelity footage and synchronized audio using text prompts.
RIFE Video Interpolation
Interpolate new in-between frames to boost a video's frame rate and produce smooth slow-motion.
MMAudio V2
mmaudio
Add realistic, synchronized sound effects, Foley, ambience, or music to a silent video from a text prompt, returning the video with a new audio track.
Generate textured 3D meshes (GLB) from text prompts or images, ready for games, AR, and product viewers.
Tripo v2.5 Text-to-3D
tripo3d
Generate a downloadable, textured 3D mesh (GLB) from a text prompt — no input photo, ready for games, AR, and product viewers.
Tripo v2.5 Image-to-3D
Turn a single object or product photo into a downloadable, textured 3D mesh (GLB) ready for games, AR, and product viewers.
Compare top text to video AI models in 2026. Generate AI videos from text prompts in 480p, 720p, and 1080p.
Seedance V1.5 Text to Video
Generate high-quality videos with synchronized audio directly from text prompts using Seedance 1.5.
Veo 3.1 Text to Video
Create cinematic 8-second videos with Veo 3.1, Google’s latest text-to-video model in the Gemini API — now with native audio, frame control, and reference image support.
LongCat-Video t2v
Turn plain text into cinematic, on-brand video. Describe the scene and camera feel; LongCat generates smooth, consistent shots with adjustable length, FPS, and quality.
Kling 2.6 Pro Text-to-Video
Generate a short, cinematic video from a text prompt with smooth, fluid motion and strong prompt adherence.
Kling 2.5 Turbo Pro Text-to-Video
Wan 2.7 Text to Video
Generate cinematic, high-fidelity video from a text prompt with smooth, coherent motion and strong prompt adherence, at 720p or 1080p.
MiniMax Hailuo-02 Standard Text-to-Video
Generate a short, cinematic 768p video from a text prompt with smooth, lifelike motion and strong prompt adherence.
Run leading open-weight AI models via API — open-source image, video, and audio generation you can inspect and self-host.
Find the best AI video generator for your project. Compare AI video generation models by output quality, resolution, speed, and cost. Run models like Seedance and Veo via API or playground.
Compare top text to image AI models in 2026. Generate high-quality images from text prompts. Try free with real-time previews and API access.
Boogu-Image
boogu
Generate high-quality images from a text prompt with fine-grained control over the CFG guidance schedule.