
Black Forest Labs Unveils FLUX 3, A New Multimodal Frontier Model For Visual Intelligence
23.7.2026 17:00:00 CEST | GlobeNewswire by notified | Press release
Jointly trained across image, video, audio, and action prediction modalities within a unified architecture, FLUX 3 brings a coherent understanding of the real world to every kind of visual creation—from generative media to robotics and beyond
FREIBURG, Germany, July 23, 2026 (GLOBE NEWSWIRE) -- Black Forest Labs, the global frontier AI research lab building the foundation layer for visual intelligence, today introduces FLUX 3, its new multimodal frontier model. FLUX 3 jointly learns from images, video, and audio within a unified architecture, and can also be extended to predict actions. The model is Black Forest Labs’ newest addition to its FLUX family of visual AI models, which are known for pairing state-of-the-art capabilities with open access, and for powering generative features inside leading creative, developer, and consumer platforms like Adobe Photoshop, Picsart, Nous Research’s Hermes Agent, and more.
The unveiling of FLUX 3 comes at an inflection point for the industry, as the focus moves beyond language models toward intelligent systems that can understand both the visual and physical world. While this shift has produced a cluster of adjacent fields—image and video generation, world models, simulation, robotics and physical AI—Black Forest Labs sees these categories as interconnected expressions of the same underlying capability: visual intelligence, or models that can perceive, predict, and act across physical and digital environments.
Multimodality for Perception and Simulation
The world is multimodal. Intelligent species and systems have evolved with the capability to process many kinds of information, including vision, sound, and other sensory signals. Vision in particular is the human brain's highest-bandwidth input system. But perception is only one component of intelligence; equally as important is the ability to form a coherent representation of the world, and to act within it.
“We place vision at the center of our approach because it is the most signal-rich medium of the physical world. Images convey structure, images and video teach spatial relationships, video teaches dynamics, and actions reveal causal relationships. But vision alone is not the complete picture,” said Robin Rombach, Co-Founder and CEO of Black Forest Labs. “True intelligence means perceiving the world: predicting how it will change, taking action, and learning from the results. Joint training within one unified architecture is what will get us there, because each training modality strengthens the others. Audio conveys timing, prosody, and physical events that elude vision. Language conveys goals, abstractions, and instructions that pixels cannot easily express.”
One Model, Multiple Capabilities
FLUX 3 builds on Self-Flow, the company’s breakthrough approach for efficiently aligning multimodal generation and understanding within the same underlying architecture. By scaling up compute and data resources, FLUX 3 was trained across multiple modalities at the same time. Testing ultimately showed that generative video generation and action prediction do not actually require separate foundations, as the same underlying architecture could be extended to action prediction without sacrificing the capabilities learned from videos.
“You can’t cheat reality. A model that only learns images can only generate images. But the world is not made of still frames. It moves, sounds, changes, and responds,” said Rombach. “That’s why FLUX 3 is trained across those signals together, so the model can build a deeper understanding of how the world works. That is what visual intelligence requires, whether the application is video generation, simulation, or robotics.”
FLUX 3 is built to support applications including in creative tooling, media, design, e-commerce, and physical AI—including video generation with synchronized audio, precisely edited images, the maintenance of product and material consistency across motion, and action prediction for robotics.
The model will be available through FLUX 3 Video, FLUX 3 Image, FLUX 3 Action, and FLUX 3 Dev. FLUX 3 is already being tested by Canva, Burda, Magnific (formerly Freepik), Krea, and Picsart. While still in development, FLUX 3 Video already leads in early evaluations against frontier video models, and is particularly strong in capturing human facial expressions, associating sounds with physical events, and multilingual capabilities.
Visual Intelligence in the Real World
Black Forest Labs and mimic robotics today also introduced FLUX-mimic, a video-action model built on FLUX 3. FLUX-mimic is currently undergoing testing and deployment with leading manufacturing companies, like Audi.
FLUX-mimic is designed for general-purpose robotic manipulation: helping robots understand a visual scene, predict the consequences of an action, and adapt to new tasks with far less task-specific data. Depending on task difficulty, the model can be fine-tuned for a specific manipulation task with as little as 30 minutes of robot data, where prior approaches have required 30 or more hours.
“The hardest part of robotics is data,” said Elvis Nava, CTO of mimic robotics. “Every new task normally means hours of a robot repeating itself. Because FLUX-mimic is built on top of frontier video models that already understand how the physical world behaves, it picks up a new task in minutes, not days. This way, we can leapfrog the current state of the art in robot learning."
“In partnership with mimic, Audi has been testing and deploying FLUX-mimic. We have seen these robots solve complex soft body manipulation work that would have been simply impossible with conventional robotics. This can have a major impact in assisting our employees, increasing efficiency, and expanding flexible automation across production and logistics operations,” said Christoph Schneider, Audi Production Lab. “For us, partnering with pioneering companies such as mimic and Black Forest Labs is essential in pushing the frontier of Physical AI and validating these innovations in real-world production environments.”
FLUX 3 Availability
FLUX 3 Video with optional native audio generation and FLUX 3 Action are now available for early access, while FLUX 3 Image will roll out in the coming weeks. Developers and teams can apply for access here, with general availability to follow. Black Forest Labs will publish full FLUX 3 benchmark results and methodology alongside broader availability.
Black Forest Labs will also release faster and open-weight versions of FLUX 3 later this year, consistent with the company’s long-held belief in building and sharing its technology openly so researchers and developers around the world can build, test, and deploy new applications on top of it. Open weights make secure, low-latency local deployment possible for applications like robotic control systems, and let teams adapt FLUX 3 to their own data, products, and workflows—turning a general model into a secure foundation for any creative, simulation, or physical AI stack.
“We are only beginning to scratch the surface of versatile, capable, unified visual models,” said Rombach. “From interactive image and video editing to simulation, physical AI, and computer use, the frontier is wide open.”
Black Forest Labs has quickly become a standard visual AI engine across the builder and creative ecosystems since its 2024 launch from stealth. Its founders have been pioneering modern visual AI for far longer, from VQGAN, latent diffusion, and Stable Diffusion to Black Forest Labs’ family of FLUX models, which power AI capabilities for leading enterprise platforms and creative industry professionals like legendary film director Martin Scorsese. To date, models from the Black Forest Labs founders have been downloaded over half a billion times.
About Black Forest Labs
Black Forest Labs is a global frontier AI research lab building the foundation layer for visual intelligence. Founded by the researchers who pioneered latent diffusion, Stable Diffusion and the FLUX model family, the 100-person team advances research and ships open, production-ready models from Freiburg, Germany and San Francisco, USA. Valued at $3.25 billion, the company has raised more than $450 million from leading venture and strategic investors including a16z, AMP, Salesforce Ventures, NVIDIA, General Catalyst, Adobe Ventures, Figma Ventures, Canva and Deutsche Telekom’s T.Capital. Learn more at bfl.ai.
Media Contact
press@blackforestlabs.ai
Subscribe to releases from GlobeNewswire by notified
Subscribe to all the latest releases from GlobeNewswire by notified by registering your e-mail address below. You can unsubscribe at any time.
Latest releases from GlobeNewswire by notified
Iveco Group signs a 150 million euro term loan facility with Cassa Depositi e Prestiti to support investments in research, development and innovation11.6.2024 12:00:00 CEST | Press release
Turin, 11th June 2024. Iveco Group N.V. (EXM: IVG), a global automotive leader active in the Commercial & Specialty Vehicles, Powertrain and related Financial Services arenas, has successfully signed a term loan facility of 150 million euros with Cassa Depositi e Prestiti (CDP), for the creation of new projects in Italy dedicated to research, development and innovation. In detail, through the resources made available by CDP, Iveco Group will develop innovative technologies and architectures in the field of electric propulsion and further develop solutions for autonomous driving, digitalisation and vehicle connectivity aimed at increasing efficiency, safety, driving comfort and productivity. The financed investments, which will have a 5-year amortising profile, will be made by Iveco Group in Italy by the end of 2025. Iveco Group N.V. (EXM: IVG) is the home of unique people and brands that power your business and mission to advance a more sustainable society. The eight brands are each a
DSV, 1115 - SHARE BUYBACK IN DSV A/S11.6.2024 11:22:17 CEST | Press release
Company Announcement No. 1115 On 24 April 2024, we initiated a share buyback programme, as described in Company Announcement No. 1104. According to the programme, the company will in the period from 24 April 2024 until 23 July 2024 purchase own shares up to a maximum value of DKK 1,000 million, and no more than 1,700,000 shares, corresponding to 0.79% of the share capital at commencement of the programme. The programme has been implemented in accordance with Regulation No. 596/2014 of the European Parliament and Council of 16 April 2014 (“MAR”) (save for the rules on share buyback programmes set out in MAR article 5) and the Commission Delegated Regulation (EU) 2016/1052, also referred to as the Safe Harbour rules. Trading dayNumber of shares bought backAverage transaction priceAmount DKKAccumulated trading for days 1-25478,1001,023.01489,100,86026:3 June 20247,0001,050.597,354,13027:4 June 20245,0001,055.705,278,50028:6 June20243,0001,096.273,288,81029:7 June 20244,0001,106.174,424,68
Landsbankinn hf.: Offering of covered bonds11.6.2024 11:16:36 CEST | Press release
Landsbankinn will offer covered bonds for sale via auction held on Thursday 13 June at 15:00. An inflation-linked series, LBANK CBI 30, will be offered for sale. In connection with the auction, a covered bond exchange offering will take place, where holders of the inflation-linked series LBANK CBI 24 can sell the covered bonds in the series against covered bonds bought in the above-mentioned auction. The clean price of the bonds is predefined at 99,594. Expected settlement date is 20 June 2024. Covered bonds issued by Landsbankinn are rated A+ with stable outlook by S&P Global Ratings. Landsbankinn Capital Markets will manage the auction. For further information, please call +354 410 7330 or email verdbrefamidlun@landsbankinn.is.
Relay42 unlocks customer intelligence with a new insights and reporting module, powered by Amazon QuickSight11.6.2024 11:00:00 CEST | Press release
AMSTERDAM, June 11, 2024 (GLOBE NEWSWIRE) -- Relay42, a leading European Customer Data Platform (CDP), is leveraging Amazon QuickSight to power its new real-time customer intelligence, reporting, and dashboard module. Harnessing the breadth and quality of customer data, the new Insights module empowers marketing teams to dive deep into customer behaviors and gain invaluable insights into the performance of their marketing programs across all online, offline, paid, and owned marketing channels. Preview of the Relay42 Insights module, in pre-beta version Key capabilities of the Relay42 Insights module include: Deep insights into customer behaviors: With the Relay42 Insights module, marketers can ask unlimited questions about their data and gain a deeper understanding of how to serve their customers more effectively. Simplicity with AI-powered querying: Marketers can use artificial intelligence to query their data using natural language search, reducing the reliance on data scientists. Us
Metasphere Labs Announces X Spaces Event on the Topic of Green Bitcoin Mining and Sound Money for Sustainability11.6.2024 10:30:00 CEST | Press release
VANCOUVER, British Columbia, June 11, 2024 (GLOBE NEWSWIRE) -- Metasphere Labs Inc. (formerly Looking Glass Labs Ltd., "Metasphere Labs" or the "Company") (Cboe Canada: LABZ) (OTC: LABZF) (FRA: H1N) is thrilled to announce an engaging Twitter Spaces event on Green Bitcoin mining, energy markets, and sustainability on July 3, 2024 at 2 p.m. ET. Follow us on X at MetasphereLabs for updates and to join the event. What We'll Discuss Bitcoin Mining Basics: Understand the fundamentals of Bitcoin mining.Energy Market Dynamics: Explore how Bitcoin mining interacts with energy markets.Sustainable Innovations: Learn about our efforts to promote sustainability in Bitcoin mining.Sound Money: Discover how tamper-proof currency can enhance stability.Efficient Payment Rails: See how fast, neutral payment systems support humanitarian projects.Carbon Footprint: Compare Bitcoin's environmental impact with traditional banking. "We're excited to host this event and dive into the critical topics of Bitcoin