Skip to main content
Black Forest Labs FLUX 3 software interface for image and audio-video editing, showcasing advanced features.

Editorial illustration for Black Forest Labs Launches FLUX 3 for Images and Audio-Video

Black Forest Labs Launches FLUX 3 for Images and Audio-Video

3 min read

Black Forest Labs is done treating images, video and audio as separate products stitched together. The Freiburg-based lab announced FLUX 3 today, a multimodal model that generates images or 20-second audio-video clips from a single prompt, and extends that same architecture into robotic vision and action. Unlike earlier approaches that bolt together specialized models behind a shared interface, BFL says FLUX 3 is trained jointly across all these modalities at once.

This marks the company's first public video generation model, arriving through four separate product lines: FLUX 3 Video, FLUX 3 Image, FLUX 3 Action, and an open source version called FLUX 3 Dev still to come. Video and Action are opening now through a gated Early Access program that anyone can apply for, though BFL controls who gets in. There's no API access yet, public or through partners, with Image expected to roll out over the coming weeks ahead of a wider release.

The staggered rollout puts BFL in step with how other frontier labs in the U.S. have been handling major model launches lately.

Black Forest Labs (BFL) is expanding its FLUX family beyond image generation with today's launch of FLUX 3, a multimodal frontier model trained to understand and generate images, or combined audio/video clips up to 20 seconds from a single prompt — and to extend the same underlying architecture to robotic vision and actions.

Why this matters

BFL is betting that a single backbone trained across images, video, audio and robotic vision beats stitching together specialist models, which is a real architectural claim worth watching, not just a marketing line. But the launch details matter more than the pitch. No downloadable weights, no open source license, just a promise that FLUX 3 Dev arrives "later this year" as open-weight access.

That's the same playbook BFL used with earlier FLUX releases: ship the closed flagship first, dangle the open version for developers who actually want to build on it. For founders evaluating whether to integrate FLUX 3 now, that gap matters, you're building on an API you don't control the roadmap for. Researchers curious about the joint-training approach will have to take BFL's technical blog at its word until weights are inspectable.

The robotics angle, extending the architecture to robotic vision and actions, is the detail to track over the next few months. If BFL follows through on open weights and shows real robotics use cases, this stops being another video generator and starts looking like infrastructure. Until then, it's a capable but locked box.

LIVE21:49Free ChatGPT Users Get Worse Health Advice From Older AI Model