← All stories
● Covered by 2 sources · 3 reportsMedium impact3 neutral

Black Forest Labs Launches FLUX 3 Multimodal AI for Image and 20-Second Video Generation

🔄 Updated 23d ago — new reporting from Hacker News Front Page
New to BrevFeed? We gather this story from every outlet covering it into one summary — ranked by real-world impact, not just the latest headline — so you never miss what matters. What is BrevFeed? →

Key points

  • FLUX 3 generates images and 20-second video with audio from a single prompt.
  • It is a multimodal model, jointly trained across image, video, and audio modalities.
  • FLUX 3 Video and FLUX 3 Action are in gated Early Access.
  • Open-source and open-weight versions are planned for later this year.
  • Black Forest Labs (BFL) released FLUX 3.
  • FLUX 3 is BFL's first public video generation model.
  • FLUX 3 is offered through FLUX 3 Video, FLUX 3 Image, FLUX 3 Action, and FLUX 3 Dev.
  • FLUX 3 Dev will be open source.
  • FLUX 3 Video has optional native audio generation.
  • Black Forest Labs (BFL) is based in Freiburg, Germany.
  • FLUX 3 is running on robots through a collaboration with Mimic Robotics.
  • FLUX-mimic is a video-action model developed with Mimic Robotics.
  • FLUX-mimic has been tested and deployed at Audi.
  • FLUX 1 and FLUX 2 generate images.

FLUX 3 Multimodal AI Release

Black Forest Labs (BFL) has launched FLUX 3, expanding its FLUX model family beyond image generation. This new multimodal frontier model is trained to understand and generate images, or combined audio/video clips up to 20 seconds, from a single prompt. It also extends its underlying architecture to robotic vision and actions.

Integrated Multimodal Training

BFL states that FLUX 3 is jointly trained across image, video, and audio modalities, rather than integrating separate models. This approach is central to the company's vision of visual intelligence, aiming to connect creative generation, simulation, computer use, and robotics as applications of a single capability.

Product Lines and Availability

FLUX 3 will be offered through four product lines: FLUX 3 Video, FLUX 3 Image, FLUX 3 Action, and the upcoming open-source FLUX 3 Dev. Currently, FLUX 3 Video (with optional native audio generation) and FLUX 3 Action are available through a gated "Early Access" program. Public access via API or partners is not yet available, but FLUX 3 Image is expected to roll out in the coming weeks, followed by general availability.

Future Open-Source and Open-Weight Plans

BFL has not yet announced pricing, service-level commitments, or evaluation methodologies. While FLUX 3 is not launching with downloadable weights or an open-source license, the company plans to release faster and open-weight versions later this year. FLUX 3 Dev is described as providing open-weight access to a multimodal backbone for content creation and action prediction, a broader commitment than previous image-only FLUX Dev releases.

Updates

🕒 2026-07-24 · new reporting from Hacker News Front Page
  • Black Forest Labs (BFL) is based in Freiburg, Germany.
  • FLUX 3 is running on robots through a collaboration with Mimic Robotics.
  • FLUX-mimic is a video-action model developed with Mimic Robotics.
  • FLUX-mimic has been tested and deployed at Audi.
  • FLUX 1 and FLUX 2 generate images.
🕒 2026-07-24 · new reporting from Hacker News Front Page
  • Black Forest Labs (BFL) released FLUX 3.
  • FLUX 3 is BFL's first public video generation model.
  • FLUX 3 is offered through FLUX 3 Video, FLUX 3 Image, FLUX 3 Action, and FLUX 3 Dev.
  • FLUX 3 Dev will be open source.
  • FLUX 3 Video has optional native audio generation.

✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →

The daily brief

One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.

One email a day. Unsubscribe in one click, any time.

Today's brief

Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.

~11 min · 9 stories · Aug 16

▶ Play today's brief Listen on Spotify

New every morning, and the back catalogue is archived by date.

How outlets covered it

FLUX 3, a new multimodal foundation model, is now running on robots through a collaboration with Mimic Robotics, resulting in FLUX-mimic, a video-action model. This development extends FLUX's capabilities from generating images and audio-visual content to controlling robots by modeling physical world behavior.

FLUX 3, a new multimodal foundation model, is now available in Early Access. This model learns jointly from images, videos, and audio within a unified architecture to better represent real-world dynamics, which is significant for advancing AI perception and content creation.

Black Forest Labs (BFL) released FLUX 3, a multimodal AI model capable of generating images and video clips up to 20 seconds with audio from a single prompt. This release marks BFL's first public video generation model and aims to connect creative generation, simulation, computer use, and robotics through a single visual intelligence capability.