Black Forest Labs has introduced Flux 3, a new multimodal foundation model capable of generating videos with native audio up to 20 seconds in length, according to The Decoder. This advancement marks a notable step in AI-generated multimedia content, combining visual and audio elements natively within a single model.

The Decoder also reports that internal testing by Black Forest Labs places Flux 3 slightly ahead of the current market leader, Seedance 2.0, highlighting its competitive edge in the evolving field of AI video generation. Flux 3’s ability to produce short, fully synchronized audiovisual clips could have multiple applications across creative and commercial sectors.

As Japan continues to expand its investment in AI and digital content technologies, innovations like Flux 3 may influence both local media production and broader market strategies in FX, crypto, and equities sectors where multimedia content is increasingly vital.