Flux 3 Text-to-Video-Draft
Text-to-image generation with FLUX.3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene. .
Examples
A young drummer waits alone beneath a concrete overpass with a small drum kit while evening traffic passes overhead. A guitarist arrives, then a saxophone player, then a woman carrying an upright bass. Without speaking, they begin playing together. The camera starts with a static wide shot, moves into rhythmic handheld close-ups of each musician, then circles around the group as pedestrians gradually stop to listen. By the end, the empty space has turned into a spontaneous street concert. Cinematic urban music film, natural performances, golden-hour light, energetic but realistic camera work.
Try this prompt →Specs
Similar models
Prices are per render and charged from your balance only when a render succeeds. Failed renders are refunded automatically. Examples show what the model can do; your result depends on your prompt and settings.