← All stories
● Covered by 1 source · 1 reportMedium impact1 neutral

Black Forest Labs Launches FLUX 3 Multimodal AI for Image and 20-Second Video Generation

🔄 Updated 1h ago
New to BrevFeed? We gather this story from every outlet covering it into one summary — ranked by real-world impact, not just the latest headline — so you never miss what matters. What is BrevFeed? →

Key points

  • FLUX 3 generates images and 20-second video with audio from a single prompt.
  • It is a multimodal model, jointly trained across image, video, and audio modalities.
  • FLUX 3 Video and FLUX 3 Action are in gated Early Access.
  • Open-source and open-weight versions are planned for later this year.

FLUX 3 Multimodal AI Release

Black Forest Labs (BFL) has launched FLUX 3, expanding its FLUX model family beyond image generation. This new multimodal frontier model is trained to understand and generate images, or combined audio/video clips up to 20 seconds, from a single prompt. It also extends its underlying architecture to robotic vision and actions.

Integrated Multimodal Training

BFL states that FLUX 3 is jointly trained across image, video, and audio modalities, rather than integrating separate models. This approach is central to the company's vision of visual intelligence, aiming to connect creative generation, simulation, computer use, and robotics as applications of a single capability.

Product Lines and Availability

FLUX 3 will be offered through four product lines: FLUX 3 Video, FLUX 3 Image, FLUX 3 Action, and the upcoming open-source FLUX 3 Dev. Currently, FLUX 3 Video (with optional native audio generation) and FLUX 3 Action are available through a gated "Early Access" program. Public access via API or partners is not yet available, but FLUX 3 Image is expected to roll out in the coming weeks, followed by general availability.

Future Open-Source and Open-Weight Plans

BFL has not yet announced pricing, service-level commitments, or evaluation methodologies. While FLUX 3 is not launching with downloadable weights or an open-source license, the company plans to release faster and open-weight versions later this year. FLUX 3 Dev is described as providing open-weight access to a multimodal backbone for content creation and action prediction, a broader commitment than previous image-only FLUX Dev releases.

✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →

The daily brief

One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.

One email a day. Unsubscribe in one click, any time.

Today's brief

Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.

~39 min · 35 stories · Jul 22

▶ Play today's brief Listen on Spotify

New every morning, and the back catalogue is archived by date.

Reporting from

Black Forest Labs (BFL) released FLUX 3, a multimodal AI model capable of generating images and video clips up to 20 seconds with audio from a single prompt. This release marks BFL's first public video generation model and aims to connect creative generation, simulation, computer use, and robotics through a single visual intelligence capability.