Home
Black Forest Labs releases Flux3, first multimodal model with native 20-second audio-visual sync
On July 23, German AI startup Black Forest Labs officially launched the multimodal foundation model Flux3, marking a significant leap forward in generative video technology. Built on the Self-Flow architecture, the model integrates dedicated codecs for image, video, audio, and motion, enabling unified understanding and generation across both physical and digital environments.

As the first model to support native audio generation, Flux3 can produce synchronized audio-video clips of up to 20 seconds, covering capabilities such as text/image/video-to-video conversion, keyframe-based transitions, and multilingual dialogue. In early benchmarks at 720p resolution for 10-second clips, Flux3 showed strong competitiveness, outperforming Luma Ray3.2 (93% win rate) and Runway Gen-4.5 (77% win rate), while maintaining a slight edge over top-tier models like Seedance2.0 and Gemini Omni Flash.
In addition, the company developed the video action model Flux-mimic in collaboration with Mimic Robotics for the robotics sector, which has already begun production task testing at an Audi factory. BFL has adopted a phased release strategy: Flux3Video is available now, with Flux3Image and the open-source weight "Flux3Dev" to follow shortly. By overcoming the limitations of single-modal information, Flux3 is accelerating AI's progress toward a "world model" capable of perception and action in the physical world.
Related article
Swedish AI Startup Lovable Eyes $13.2 Billion Valuation After Major Funding Round
As AI-driven coding tools gain traction, Swedish startup Lovable has secured a major funding round. The company aims to raise $3 billion, potentially boosting its valuation to $13.2 billion—double the $6.6 billion recorded last December. Menlo Ventur
Google Tests Remy AI Agent for Gemini as Focus Shifts to User Control
According to Business Insider, Google is testing Remy, a new AI personal agent for Gemini. This tool aims to execute tasks on behalf of users, streamlining both professional workflows and daily routines.Currently, Remy is undergoing testing in an int
How to fix Core Web Vitals for better SEO rankings
Streamline Report Card Comments with AI ToolsIntroductionAI Tools for Generating Report Card CommentsMagic SchoolAlmanac AIChat GPTUsing Magic School to Generate Report Card CommentsLogging into Magic SchoolSelecting the Report Card Comments ToolCust
Related Special Topic Recommendations
Comments (0)
0/500
On July 23, German AI startup Black Forest Labs officially launched the multimodal foundation model Flux3, marking a significant leap forward in generative video technology. Built on the Self-Flow architecture, the model integrates dedicated codecs for image, video, audio, and motion, enabling unified understanding and generation across both physical and digital environments.

As the first model to support native audio generation, Flux3 can produce synchronized audio-video clips of up to 20 seconds, covering capabilities such as text/image/video-to-video conversion, keyframe-based transitions, and multilingual dialogue. In early benchmarks at 720p resolution for 10-second clips, Flux3 showed strong competitiveness, outperforming Luma Ray3.2 (93% win rate) and Runway Gen-4.5 (77% win rate), while maintaining a slight edge over top-tier models like Seedance2.0 and Gemini Omni Flash.
In addition, the company developed the video action model Flux-mimic in collaboration with Mimic Robotics for the robotics sector, which has already begun production task testing at an Audi factory. BFL has adopted a phased release strategy: Flux3Video is available now, with Flux3Image and the open-source weight "Flux3Dev" to follow shortly. By overcoming the limitations of single-modal information, Flux3 is accelerating AI's progress toward a "world model" capable of perception and action in the physical world.
Swedish AI Startup Lovable Eyes $13.2 Billion Valuation After Major Funding Round
As AI-driven coding tools gain traction, Swedish startup Lovable has secured a major funding round. The company aims to raise $3 billion, potentially boosting its valuation to $13.2 billion—double the $6.6 billion recorded last December. Menlo Ventur
How to fix Core Web Vitals for better SEO rankings
Streamline Report Card Comments with AI ToolsIntroductionAI Tools for Generating Report Card CommentsMagic SchoolAlmanac AIChat GPTUsing Magic School to Generate Report Card CommentsLogging into Magic SchoolSelecting the Report Card Comments ToolCust











