Models

Image, video, 3D, speech and audio models from leading labs. Compare prices and try any of them with one balance.

636 of 636 models
Example made with openai/gpt-image-2.5-flare/text-to-imagePhoto
OpenAI

openai/gpt-image-2.5-flare/text-to-image

Images from text4K

OpenAI's GPT Image 2.5 Flare Text-to-Image generates high-quality images from natural-language prompts, with five quality tiers up to 4K. Flare is the fast, balanced GPT Image 2.5 tier for everyday generation at low latency.

from€0.05per renderDetails →
Example made with openai/gpt-image-2.5-sunburst/editPhoto
OpenAI

openai/gpt-image-2.5-sunburst/edit

Image editing4K

OpenAI's GPT Image 2.5 Sunburst Edit edits one or more reference images from natural-language instructions, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail.

from€0.08per renderDetails →
Example made with openai/gpt-image-2.5-sunburst/text-to-imagePhoto
OpenAI

openai/gpt-image-2.5-sunburst/text-to-image

Images from text4K

OpenAI's GPT Image 2.5 Sunburst Text-to-Image generates high-quality images from natural-language prompts, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail.

from€0.05per renderDetails →
Example made with openai/gpt-image-2.5-flare/editPhoto
OpenAI

openai/gpt-image-2.5-flare/edit

Image editing4K

OpenAI's GPT Image 2.5 Flare Edit edits one or more reference images from natural-language instructions, with five quality tiers up to 4K. Flare is the fast, balanced GPT Image 2.5 tier for everyday generation at low latency.

from€0.08per renderDetails →
Example made with bytedance/seedance-2.5/talking-avatarVideo
Bytedance

bytedance/seedance-2.5/talking-avatar

Talking avatarsPortraits

Animate a portrait image from an audio recording with Seedance 2.5. Generate talking-avatar videos in 480p or 720p, processing up to the first 120 seconds of audio.

from€2.20per renderDetails →
Example made with bytedance/seedance-2.5/image-to-videoVideo
Bytedance

bytedance/seedance-2.5/image-to-video

Animating imagesRealistic video

Seedance 2.5 (Image-to-Video) generates Hollywood-grade cinematic videos from reference images and text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability.

from€1.80per renderDetails →
Example made with bytedance/seedance-2.5/video-editVideo
Bytedance

bytedance/seedance-2.5/video-edit

Video editingRealistic video

Seedance 2.5 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed.

from€2.20per renderDetails →
Example made with bytedance/seedance-2.5/video-edit-turboVideo
Bytedance

bytedance/seedance-2.5/video-edit-turbo

Video editing

Seedance 2.5 (Video-Edit Turbo) is the turbo tier for editing an input video from a natural-language prompt — faster, more affordable high-resolution output while preserving subject identity, composition, and motion.

from€2.60per renderDetails →
Example made with bytedance/seedance-2.5/video-extendVideo
Bytedance

bytedance/seedance-2.5/video-extend

Extending videoRealistic video

Seedance 2.5 (Video-Extend) extends an input video with a new cinematic continuation generated from its last frame and a natural-language prompt.

from€2.20per renderDetails →
Example made with bytedance/seedance-2.5/text-to-videoVideo
Bytedance

bytedance/seedance-2.5/text-to-video

Video from textRealistic video

Seedance 2.5 (Text-to-Video) generates Hollywood-grade cinematic videos from text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability.

from€1.80per renderDetails →
Example made with alibaba/wan-3.0-prime/video-extendVideo
Alibaba

alibaba/wan-3.0-prime/video-extend

Extending video

Wan 3.0 Prime Video Extend continues a video from its final frame and appends a newly generated 2-30 second segment. Inputs longer than 120 seconds retain their last 120 seconds. Supports 480p, 720p, and 1080p output.

from€1.50per renderDetails →
Video
Alibaba

alibaba/wan-3.0-prime/image-to-video-spicy

Animating images

Alibaba WAN 3.0 Prime Spicy Image-to-Video converts a first-frame image into a video with optional last-frame guidance, flexible 2-30 second duration, resolution, aspect ratio, audio, and deep-thinking controls.

from€1.50per renderDetails →
Example made with alibaba/wan-3.0-prime/image-to-videoVideo
Alibaba

alibaba/wan-3.0-prime/image-to-video

Animating images

Alibaba WAN 3.0 Prime Image-to-Video converts a first-frame image into a video with optional last-frame guidance, flexible 2-30 second duration, resolution, aspect ratio, audio, and deep-thinking controls.

from€1.50per renderDetails →
Example made with alibaba/wan-3.0-prime/reference-to-videoVideo
Alibaba

alibaba/wan-3.0-prime/reference-to-video

Animating images

Alibaba WAN 3.0 Prime Reference-to-Video combines reference images, videos, and audio with prompts to create coherent videos with flexible 2-30 second duration, resolution, aspect ratio, audio, and deep-thinking controls.

from€1.50per renderDetails →
Example made with alibaba/wan-3.0-prime/video-editVideo
Alibaba

alibaba/wan-3.0-prime/video-edit

Video editing

Wan 3.0 Prime Video-Edit edits an input video using a text prompt and optional reference images and audio. Inputs longer than 15 seconds are trimmed to the first 15 seconds. Output aspect ratio is explicitly selected from the input display dimensions automatically.

from€1.50per renderDetails →
Video
Alibaba

alibaba/wan-3.0-prime/text-to-video

Video from text

Alibaba WAN 3.0 Prime Text-to-Video generates videos from text prompts with flexible 2-30 second duration, resolution, aspect ratio, audio, and deep-thinking controls.

from€1.50per renderDetails →
Video
Bytedance

bytedance/seedance-2.5/image-to-video-spicy

Animating imagesRealistic video

Seedance 2.5 Spicy (Image-to-Video) generates unlimited high-quality cinematic clips from a reference image and prompt, optimized for scalable content generation with smooth animations and stable aesthetics.

from€1.80per renderDetails →
Example made with alibaba/wan-3.0/reference-to-videoVideo
Alibaba

alibaba/wan-3.0/reference-to-video

Animating images

Alibaba WAN 3.0 Reference-to-Video combines reference images, videos, and audio with prompts to create coherent videos with flexible 2-30 second duration, resolution, aspect ratio, audio, and deep-thinking controls.

from€1.00per renderDetails →
Example made with alibaba/wan-3.0/image-to-videoVideo
Alibaba

alibaba/wan-3.0/image-to-video

Animating images

Alibaba WAN 3.0 Image-to-Video converts a first-frame image into a video with optional last-frame guidance, flexible 2-30 second duration, resolution, aspect ratio, audio, and deep-thinking controls.

from€1.00per renderDetails →
Example made with alibaba/wan-3.0/text-to-videoVideo
Alibaba

alibaba/wan-3.0/text-to-video

Video from text

Alibaba WAN 3.0 Text-to-Video generates videos from text prompts with flexible 2-30 second duration, resolution, aspect ratio, audio, and deep-thinking controls.

from€1.00per renderDetails →
Example made with alibaba/wan-3.0/video-editVideo
Alibaba

alibaba/wan-3.0/video-edit

Video editing

Wan 3.0 Video-Edit edits an input video using a text prompt and optional reference images and audio. Inputs longer than 15 seconds are trimmed to the first 15 seconds. Output aspect ratio is explicitly selected from the input display dimensions automatically.

from€1.00per renderDetails →
Video
Alibaba

alibaba/wan-3.0/image-to-video-spicy

Animating images

Alibaba WAN 3.0 Spicy Image-to-Video converts a first-frame image into a video with optional last-frame guidance, flexible 2-30 second duration, resolution, aspect ratio, audio, and deep-thinking controls.

from€1.00per renderDetails →
Example made with alibaba/wan-3.0/video-extendVideo
Alibaba

alibaba/wan-3.0/video-extend

Extending video

Wan 3.0 Video Extend continues a video from its final frame and appends a newly generated 2-30 second segment. Inputs longer than 120 seconds retain their last 120 seconds. Supports 480p, 720p, and 1080p output.

from€1.00per renderDetails →
Example made with bytedance/seedance-2.5/image-to-video-turboVideo
Bytedance

bytedance/seedance-2.5/image-to-video-turbo

Animating imagesRealistic video

Seedance 2.5 (Image-to-Video Turbo) generates cinematic 720p/1080p videos from reference images and text prompts —a faster, more affordable high-resolution tier with native audio-visual synchronization, director-level control, and exceptional motion stability.

from€2.00per renderDetails →
Example made with bytedance/seedance-2.5/text-to-video-turboVideo
Bytedance

bytedance/seedance-2.5/text-to-video-turbo

Video from textRealistic video

Seedance 2.5 (Text-to-Video Turbo) generates cinematic videos from text prompts at 720p and 1080p with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability — optimized for turbo output.

from€2.00per renderDetails →
Example made with bytedance/seedream-v5.0-pro/editPhoto
Bytedance

bytedance/seedream-v5.0-pro/edit

Image editing

Seedream 5.0 Pro API Preview Edit is ByteDance's advanced image editing model for single-image and multi-reference image generation.

from€0.09per renderDetails →
Example made with bytedance/seedream-v5.0-proPhoto
Bytedance

bytedance/seedream-v5.0-pro

Images from text

Seedream 5.0 Pro API Preview is ByteDance's advanced image generation model for text-to-image and reference-image generation.

from€0.09per renderDetails →
Example made with bytedance/seedream-v5.0-flash/editPhoto
Bytedance

bytedance/seedream-v5.0-flash/edit

Image editing

Seedream 5.0 Flash Edit is ByteDance's fast, lower-cost image editing model for single-image and multi-reference image generation.

from€0.06per renderDetails →
Example made with bytedance/seedream-v5.0-flashPhoto
Bytedance

bytedance/seedream-v5.0-flash

Images from text

Seedream 5.0 Flash is ByteDance's fast, lower-cost image generation model for text-to-image generation.

from€0.06per renderDetails →
Example made with alibaba/qwen-image-3.0/editPhoto
Alibaba

alibaba/qwen-image-3.0/edit

Image editing

Qwen Image 3.0 edit model with high-quality image editing and advanced instruction understanding.

from€0.06per renderDetails →
Example made with alibaba/qwen-image-3.0-pro/editPhoto
Alibaba

alibaba/qwen-image-3.0-pro/edit

Image editing

Qwen Image 3.0 Pro edit model with professional-grade editing quality and advanced instruction understanding.

from€0.08per renderDetails →
Example made with alibaba/qwen-image-3.0/text-to-imagePhoto
Alibaba

alibaba/qwen-image-3.0/text-to-image

Images from text

Qwen Image 3.0 text-to-image model with high-quality image generation and advanced prompt understanding.

from€0.06per renderDetails →
Example made with alibaba/qwen-image-3.0-pro/text-to-imagePhoto
Alibaba

alibaba/qwen-image-3.0-pro/text-to-image

Images from text

Qwen Image 3.0 Pro text-to-image model with professional-grade quality and advanced prompt understanding.

from€0.08per renderDetails →
Example made with x-ai/grok-imagine-image-v2.0/editPhoto
xAI

x-ai/grok-imagine-image-v2.0/edit

Image editing

Grok Imagine Image V2.0 Edit transforms an input image according to a text prompt with configurable resolution and quality.

from€0.12per renderDetails →
Example made with x-ai/grok-imagine-image-v2.0/text-to-imagePhoto
xAI

x-ai/grok-imagine-image-v2.0/text-to-image

Images from text

Grok Imagine Image V2.0 Text to Image generates images from text prompts with configurable aspect ratio, resolution, and quality.

from€0.10per renderDetails →
Example made with black-forest-labs/flux-3/image-to-videoVideo
Black Forest Labs

black-forest-labs/flux-3/image-to-video

Animating images

Text-to-image generation with FLUX.3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene. .

from€1.70per renderDetails →
Example made with black-forest-labs/flux-3/start-end-to-videoVideo
Black Forest Labs

black-forest-labs/flux-3/start-end-to-video

Animating images

Text-to-image generation with FLUX.3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene. .

from€1.70per renderDetails →
Example made with black-forest-labs/flux-3/video-extendVideo
Black Forest Labs

black-forest-labs/flux-3/video-extend

Extending video

Text-to-image generation with FLUX.3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene. .

from€4.10per renderDetails →
Example made with black-forest-labs/flux-3/text-to-videoVideo
Black Forest Labs

black-forest-labs/flux-3/text-to-video

Video from text

Text-to-image generation with FLUX.3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene. .

from€1.70per renderDetails →
Example made with minimax/h3/reference-to-videoVideo
Minimax

minimax/h3/reference-to-video

Animating images

MiniMax H3 is a 2K reference-to-video model that combines reference images, videos, and audio with natural-language instructions to create coherent video scenes.

from€1.40per renderDetails →
Example made with minimax/h3/text-to-videoVideo
Minimax

minimax/h3/text-to-video

Video from text

MiniMax H3 is a 2K text-to-video model that generates coherent videos from natural-language prompts with flexible duration and aspect-ratio controls.

from€1.40per renderDetails →
Example made with minimax/h3/image-to-videoVideo
Minimax

minimax/h3/image-to-video

Animating images

MiniMax H3 is a 2K image-to-video model that animates a first-frame image with natural-language motion and scene instructions, with optional last-frame control.

from€1.40per renderDetails →
Example made with pruna-ai/p-video-2-pro/text-to-videoVideo
Pruna Ai

pruna-ai/p-video-2-pro/text-to-video

Video from text

Pruna P-Video-2-Pro text-to-video generation at 480p or 768p, with generated audio.

from€0.04per renderDetails →
Example made with pruna-ai/p-video-2/text-to-videoVideo
Pruna Ai

pruna-ai/p-video-2/text-to-video

Video from text

Pruna P-Video-2 text-to-video generation with explicit duration, resolution, draft mode, and audio output.

from€0.05per renderDetails →
Example made with pruna-ai/p-video-2/image-to-videoVideo
Pruna Ai

pruna-ai/p-video-2/image-to-video

Animating images

Pruna P-Video-2 image-to-video generation with explicit duration, resolution, draft mode, and audio output.

from€0.05per renderDetails →
Example made with pruna-ai/p-video-2-pro/image-to-videoVideo
Pruna Ai

pruna-ai/p-video-2-pro/image-to-video

Animating images

Pruna P-Video-2-Pro image-to-video with optional last-frame guidance generation at 480p or 768p, with generated audio.

from€0.04per renderDetails →
Example made with bytedance/seedance-2.0-mini/video-edit-turboVideo
Bytedance

bytedance/seedance-2.0-mini/video-edit-turbo

Video editingRealistic video

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 720p-1080p, 5-12s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16.

from€1.90per renderDetails →
Example made with bytedance/seedance-2.0-mini/image-to-video-turboVideo
Bytedance

bytedance/seedance-2.0-mini/image-to-video-turbo

Animating imagesRealistic video

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 720p-1080p, 5-12s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16.

from€1.40per renderDetails →
Example made with bytedance/seedance-2.0-mini/video-editVideo
Bytedance

bytedance/seedance-2.0-mini/video-edit

Video editingRealistic video4K

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16.

from€1.50per renderDetails →
Example made with bytedance/seedance-2.0-mini/image-to-videoVideo
Bytedance

bytedance/seedance-2.0-mini/image-to-video

Animating imagesRealistic video4K

Seedance 2.0 Mini is ByteDance's faster, lower-cost tier of Seedance 2.0 for cinematic multi-shot video — narrative sequences, AI camera control (zoom/pan/tracking), and consistent characters across scenes, from text or image prompts. 480p-4k, 4-15s, aspect ratios 16:9 / 4:3 / 1:1 / 3:4 / 9:16.

from€1.20per renderDetails →
Example made with alibaba/happyhorse-1.0/video-editVideo
Alibaba

alibaba/happyhorse-1.0/video-edit

Video editing

Alibaba Happy Horse 1.0 (Video Edit) performs prompt-driven video editing with multi-image reference support, supporting 720p/1080p output. Smooth, context-aware edits while preserving motion and temporal consistency.

from€1.40per renderDetails →
Example made with alibaba/happyhorse-1.1/reference-to-videoVideo
Alibaba

alibaba/happyhorse-1.1/reference-to-video

Animating images

Alibaba Happy Horse 1.1 (Reference-to-Video) generates new video scenes guided by reference images, maintaining consistent characters, styles, and visual identity. Smooth camera movement and expressive, stable motion.

from€1.40per renderDetails →
Example made with alibaba/happyhorse-1.1/video-extendVideo
Alibaba

alibaba/happyhorse-1.1/video-extend

Extending video

Alibaba Happy Horse 1.1 (Video Extend) extends existing videos with seamless AI-generated continuation, supporting 720p/1080p output. Natural, motion-consistent extension.

from€1.40per renderDetails →
Example made with alibaba/happyhorse-1.0/video-extendVideo
Alibaba

alibaba/happyhorse-1.0/video-extend

Extending video

Alibaba Happy Horse 1.0 (Video Extend) extends existing videos with seamless AI-generated continuation, supporting 720p/1080p output. Natural, motion-consistent extension.

from€1.40per renderDetails →
Example made with alibaba/happyhorse-1.0/reference-to-videoVideo
Alibaba

alibaba/happyhorse-1.0/reference-to-video

Animating images

Alibaba Happy Horse 1.0 (Reference-to-Video) generates new video scenes guided by reference images, maintaining consistent characters, styles, and visual identity. Smooth camera movement and expressive, stable motion.

from€1.40per renderDetails →
Example made with alibaba/wan-2.7/image-to-video-proVideo
Alibaba

alibaba/wan-2.7/image-to-video-pro

Animating imagesRealistic video4K

Alibaba WAN 2.7 Pro converts images into ultra-high-resolution videos (1080p/2K/4K) with cinematic detail and smooth motion.

from€1.20per renderDetails →
Example made with alibaba/wan-2.7/image-edit-proPhoto
Alibaba

alibaba/wan-2.7/image-edit-pro

Image editing4K

Alibaba WAN 2.7 Image Edit Pro performs prompt-driven image editing with multi-image reference support and up to 4K output.

from€0.15per renderDetails →

Prices are per render and charged from your balance only when a render succeeds. Failed renders are refunded automatically. The exact price shows before you generate.