{"@context":"https://schema.org","@type":"NewsArticle","generatedAt":"2026-07-28T06:20:51.496Z","headline":"FLUX 3 - Real World Models： Towards Multimodal Flow Models as the Backbone of Visual Intelligence. FLUX 3， our new multimodal frontier model， jointly learns from images， video， and audio to build one representation of the world. Now available in Early Access. July 23， 2026 Read more","description":"FLUX 3, our new multimodal frontier model, jointly learns from images, video, and audio to build one representation of the world. Now available in Early Access.","url":"https://www.aioga.com/news/cms3dpvqp0aorro3fx8vn2b5w/","mainEntityOfPage":"https://www.aioga.com/news/cms3dpvqp0aorro3fx8vn2b5w/","datePublished":"2026-07-22T16:00:00.000Z","dateModified":"2026-07-22T16:00:00.000Z","inLanguage":"zh-CN","publisher":{"@type":"NewsMediaOrganization","name":"Aioga","url":"https://www.aioga.com"},"citation":["https://bfl.ai/blog/flux-3","https://aihot.virxact.com/items/cms3dpvqp0aorro3fx8vn2b5w"],"canonicalUrl":"https://www.aioga.com/news/cms3dpvqp0aorro3fx8vn2b5w/","directAnswer":{"@type":"Answer","text":"Black Forest Labs宣布推出多模态基础模型FLUX 3，并开放Early Access。该模型在统一架构中共同学习图像、视频和音频，目标是形成对现实世界的综合表示。","url":"https://www.aioga.com/news/cms3dpvqp0aorro3fx8vn2b5w/","dateCreated":"2026-07-22T16:00:00.000Z","author":{"@type":"Organization","@id":"https://www.aioga.com/authors/aioga-editorial/#editorial-team","name":"Aioga Editorial Team","url":"https://www.aioga.com/authors/aioga-editorial/"}},"evidence":[{"@type":"CreativeWork","name":"Black Forest Labs：Blog（网页） source article","url":"https://bfl.ai/blog/flux-3","datePublished":"2026-07-22T16:00:00.000Z","provider":{"@type":"Organization","name":"Black Forest Labs：Blog（网页）","url":"https://bfl.ai/blog/flux-3"}},{"@type":"CreativeWork","name":"AIHot archive record","url":"https://aihot.virxact.com/items/cms3dpvqp0aorro3fx8vn2b5w","datePublished":"2026-07-22T16:00:00.000Z","provider":{"@type":"Organization","name":"AIHot","url":"https://aihot.virxact.com/items/cms3dpvqp0aorro3fx8vn2b5w"}}],"aggregationSource":"Black Forest Labs：Blog（网页）","originalPublisher":{"name":"Black Forest Labs：Blog（网页）","url":"https://bfl.ai/blog/flux-3"},"article":{"id":"cms3dpvqp0aorro3fx8vn2b5w","slug":"cms3dpvqp0aorro3fx8vn2b5w","url":"https://www.aioga.com/news/cms3dpvqp0aorro3fx8vn2b5w/","title":"FLUX 3 - Real World Models： Towards Multimodal Flow Models as the Backbone of Visual Intelligence. FLUX 3， our new multimodal frontier model， jointly learns from images， video， and audio to build one representation of the world. Now available in Early Access. July 23， 2026 Read more","title_en":"","summary":"FLUX 3, our new multimodal frontier model, jointly learns from images, video, and audio to build one representation of the world. Now available in Early Access.","source":"Black Forest Labs：Blog（网页）","sourceUrl":"https://bfl.ai/blog/flux-3","aiHotUrl":"https://aihot.virxact.com/items/cms3dpvqp0aorro3fx8vn2b5w","publishedAt":"2026-07-22T16:00:00.000Z","category":"模型更新","score":59,"selected":true,"articleBody":["FLUX 3 is now available in Early Access.","FLUX 3 is our new multimodal foundation model. It jointly learns from images, videos, and audio within a unified architecture, because what it needs to learn is not any one of these elements in isolation. Instead, a model must learn a representation of the world: how objects hold together, how things move, and how events sound.","No single modality provides a complete description. Each is a projection of the same underlying reality, captured by different sensors, each of which loses some information in the process. Images capture spatial structures and relationships at a specific point in time. Videos restore the dimension of time and reveal temporal dynamics and physical laws. Audio reveals causal relationships between mechanical phenomena and acoustics that vision alone cannot detect. Language links these perceptions to goals, abstractions, and instructions.","Learn from one and you get a good model of that projection. Learn from all of them at once and their mutual constraints tell you more: the sound has to match the impact, the motion has to obey the mass, the future has to follow from the past. The modalities stop being separate and start being evidence about one underlying reality.","FLUX 3 is our first model built entirely on that principle, and a checkpoint on our mission to develop real-world visual intelligence: models that perceive, predict, and act across physical and digital environments. Early results in content creation and physical AI suggest it is the right path.","FLUX 3 builds on Self-Flow,：https://bfl.ai/research/self-flow our approach for efficiently aligning multimodal generation and understanding within the same underlying architecture. Based on this approach, we significantly scaled up compute and data resources to train FLUX 3 across video, images, and audio at the same time.","Self-Flow vs. Flow Matching (FM). Left: generation error (Fréchet distance) per modality, each normalized to FM = 100 (lower is better). Right: success rate on manipulation tasks averaged over four task groups through finetuning (higher is better).","As a result, FLUX 3 is capable of mixing modalities and generating images and video+audio jointly; both from pure text prompts as well as when providing input references such as images and video. We are highlighting a few of the model’s key capabilities below.","FLUX 3 can create highly diverse videos with audio up to 20 seconds in length in a single generation.","Its core capabilities include the following (all outputs come with native audio generation):","For the preliminary analysis below, we generated 10-second text-to-video clips in 720p with audio.","Evaluations are early and we expect further improvements","As the model and the harness around it are still in development, these results are preliminary, and we expect further improvements during the early access phase. Across early evaluations, FLUX 3 was preferred over Grok Imagine Video in up to 69% of comparisons, Kling v3 Pro in 60%, Happy Horse v1 in 59%, Happy Horse 1.1 in 57%, Seedance 2.0 and Gemini Omni Flash in 52%. FLUX 3 was preferred over Runway Gen-4.5 in 77% of comparisons and over Luma Ray 3.2 in 93% of comparisons.","While still in development, FLUX 3 Video is already particularly strong in capturing human facial expressions, associating sounds with physical events, and multilingual capabilities. Furthermore, these capabilities can be combined to create sequences lasting several minutes, where visual references help ensure that the characters remain consistent across all scenes.","FLUX 3 Video is now available in Early Access here：https://bfl.ai/models/flux-3","FLUX 3 can synthesize and edit images in a wide variety of styles, aspect ratios, and resolutions. In preliminary evaluations conducted during midtraining, FLUX 3 already shows a significant improvement over earlier versions of FLUX: its ability to handle complex prompts and text generation has improved significantly. The model produces a wide range of output styles (see the following samples), and is able to render high-accuracy text in multiple languages.","As with video evaluations, these are preliminary results, and we expect further improvements before release. We will open up an early access phase for FLUX 3 Image in the following weeks.","FLUX 3's world understanding extends to action prediction. We have taken two routes to it: integrating native action prediction into FLUX 3 directly, scaling up our initial work in Self-Flow; and using the pretrained video backbone as a dynamics-aware foundation that specialized action models can be finetuned from with limited task-specific data.","For the second, mimic robotics was one of the first partners to gain early access to FLUX 3. Together we developed FLUX-mimic, a video-action model combining the FLUX 3 backbone with mimic's expertise in robot learning for dexterous manipulation and production deployment. Read our thesis on why physical AI and content creation run on the same foundation, and how it's being tested on real production tasks at Audi.：https://bfl.ai/blog/flux-3-mimic","Over the next few weeks and months, we will make the following capabilities available, each after an early access phase for ensuring smooth rollout, collecting feedback and rigorous safety-testing. All capabilities are built from the same underlying multimodal flow matching model. These capabilities and models include:","We will also release more technical details on the underlying approach.","Request early access here ：https://tally.so/r/44d9NX","We are only beginning to scratch the surface of versatile, capable, unified multimodal models, and what they will enable. From interactive image & video editing, simulation to computer use and physical AI, the frontier is wide open. While we gradually roll out these new capabilities, we are already working on the next generation models. Our goal is to unify perceptual, action and language prediction in the same unified model.","If you are interested in exploring and building with FLUX 3, get in touch here. If you are interested in contributing to our mission, join us! We are hiring：https://bfl.ai/careers in Germany and the US."],"articleImages":[{"sourceUrl":"https://bfl.ai/_next/image?url=https%3A%2F%2Fcdn.sanity.io%2Fimages%2F2gpum2i6%2Fproduction%2F8cfc4af7a44825034e3bc938927315b68d6696bd-1900x1264.png&w=3840&q=75","alt":"FLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligence.","afterParagraph":0,"url":"/media/articles/cmryn9g0u033yrolggxnwywgy/0a3b4423ff043b3e.avif"},{"sourceUrl":"https://bfl.ai/_next/image?url=https%3A%2F%2Fcdn.sanity.io%2Fimages%2F2gpum2i6%2Fproduction%2F3deccc05050d243a8c6c7cbc704ec3a07f53df75-3040x1256.png&w=3840&q=75","alt":"","afterParagraph":4,"url":"/media/articles/cmryn9g0u033yrolggxnwywgy/0bb5015868b6a327.avif"},{"sourceUrl":"https://bfl.ai/_next/image?url=https%3A%2F%2Fcdn.sanity.io%2Fimages%2F2gpum2i6%2Fproduction%2F84c6a074bfe2ea965f869b948a51df493ec3637a-2240x1152.png&w=3840&q=75","alt":"","afterParagraph":5,"url":"/media/articles/cmryn9g0u033yrolggxnwywgy/015a769de0179590.avif"},{"sourceUrl":"https://bfl.ai/_next/image?url=https%3A%2F%2Fcdn.sanity.io%2Fimages%2F2gpum2i6%2Fproduction%2F30bb456fef266f4b6115a9f209d4b907108ae09e-2880x1800.png&w=3840&q=75","alt":"","afterParagraph":10,"url":"/media/articles/cmryn9g0u033yrolggxnwywgy/14b28030b8894fb7.avif"},{"sourceUrl":"https://bfl.ai/_next/image?url=https%3A%2F%2Fcdn.sanity.io%2Fimages%2F2gpum2i6%2Fproduction%2F67e6dbce8f40db2163330184d35c8dfe4d80bb32-3400x3659.png&w=3840&q=75","alt":"","afterParagraph":15,"url":"/media/articles/cmryn9g0u033yrolggxnwywgy/3920aa04ddaa1397.avif"},{"sourceUrl":"https://bfl.ai/_next/image?url=%2F_next%2Fstatic%2Fmedia%2Fbg-footer.c5582a35.png&w=3840&q=75&dpl=dpl_EZ5bN1KXojJ5NJroAyRfEJ9mxZ7G","alt":"","afterParagraph":23,"url":"/media/articles/cmryn9g0u033yrolggxnwywgy/bd35fd4194d3e52a.avif"}],"mediaStatus":"ok","articleBodyZh":["FLUX 3 现已提供早期访问。","FLUX 3 是我们新的多模态基础模型。它在统一架构中同时从图像、视频和音频中学习，因为它需要学习的不是这些元素的任何一个孤立部分。而是模型必须学习世界的表示：物体如何保持完整，事物如何运动，以及事件听起来如何。","没有单一模态能提供完整的描述。每种模态都是同一底层现实的投影，由不同的传感器捕捉，每种传感器在过程中都会丢失一些信息。图像捕捉特定时间点的空间结构和关系。视频恢复时间维度并揭示时间动态和物理规律。音频揭示机械现象与声学之间的因果关系，这是单靠视觉无法检测的。语言将这些感知与目标、抽象和指令联系起来。","从单一模态学习，你会得到该投影的良好模型。从所有模态同时学习，它们的相互约束让你知道更多：声音必须与撞击匹配，运动必须遵守质量规律，未来必须从过去推导出来。模态不再是分离的，而是关于同一底层现实的证据。","FLUX 3 是我们第一个完全基于这一原理构建的模型，也是我们开发现实世界视觉智能使命的一个里程碑：能够在物理和数字环境中感知、预测和行动的模型。内容创作和物理人工智能的初步结果表明，这是正确的路径。","FLUX 3 建立在 Self-Flow（https://bfl.ai/research/self-flow）的基础上，这是我们在相同底层架构中高效对齐多模态生成和理解的方法。基于这一方法，我们显著扩大了计算和数据资源，以便同时跨视频、图像和音频训练 FLUX 3。","Self-Flow 与 Flow Matching (FM) 对比。左图：每种模态的生成误差（Fréchet 距离），每个模态归一化为 FM = 100（值越低越好）。右图：微调后四组任务中操作任务的平均成功率（值越高越好）。","因此，FLUX 3 能够混合多种模态并联合生成图像和视频+音频；既可以从纯文本提示生成，也可以在提供图像和视频等输入参考时生成。我们在下面重点介绍了模型的一些关键能力。","FLUX 3 可以在单次生成中创建长度最长为 20 秒、带音频的高度多样化视频。","其核心能力包括以下内容（所有输出均包含原生音频生成）：","在下面的初步分析中，我们生成了带音频的 10 秒文本到视频剪辑，分辨率为 720p。","评估仍处于早期阶段，我们预期会有进一步的改进。","由于模型及其相关工具尚在开发中，这些结果是初步的，我们预期在早期访问阶段会有进一步改进。在早期评估中，FLUX 3 在最多 69% 的比较中被优于 Grok Imagine Video，在 60% 的比较中优于 Kling v3 Pro，在 59% 的比较中优于 Happy Horse v1，在 57% 的比较中优于 Happy Horse 1.1，在 52% 的比较中优于 Seedance 2.0 和 Gemini Omni Flash。FLUX 3 在 77% 的比较中优于 Runway Gen-4.5，在 93% 的比较中优于 Luma Ray 3.2。","虽然仍在开发中，FLUX 3 Video 已经在捕捉人类面部表情、将声音与物理事件关联以及多语言能力方面表现出特别强的能力。此外，这些能力可以结合起来创建持续几分钟的序列，其中视觉参考有助于确保角色在所有场景中保持一致。","FLUX 3 Video 现已通过早期访问提供：https://bfl.ai/models/flux-3","FLUX 3 可以合成和编辑各种风格、纵横比和分辨率的图像。在中期训练期间进行的初步评估显示，FLUX 3 已经比早期版本有了显著提升：其处理复杂提示和文本生成的能力显著提高。模型可以生成多种输出风格（见下列样例），并能够用多种语言渲染高精度文本。","与视频评估一样，这些是初步结果，我们预期在发布前会有进一步改进。我们将在未来几周为 FLUX 3 图像开启早期访问阶段。","FLUX 3 对世界的理解已经扩展到动作预测。我们采取了两条路线实现这一点：一是将原生动作预测直接整合到 FLUX 3 中，扩大我们在 Self-Flow 中的初步工作；二是使用预训练的视频骨干网络作为具有动态感知的基础，之后可以用有限的任务特定数据对专业动作模型进行微调。","在第二条路线中，Mimic Robotics 是最早获得 FLUX 3 早期访问权限的合作伙伴之一。我们共同开发了 FLUX-mimic，这是一种视频动作模型，将 FLUX 3 的骨干网络与 Mimic 在机器人学习、灵巧操作和生产部署的专业知识结合起来。阅读我们的论文，了解为什么物理 AI 和内容创作运行在相同的基础上，以及它如何在奥迪的实际生产任务中进行测试：https://bfl.ai/blog/flux-3-mimic","在接下来的数周和数月内，我们将提供以下功能，每一项功能在正式发布前都会经过早期访问阶段，以确保顺利推出、收集反馈并进行严格的安全测试。所有功能都基于相同的多模态流匹配模型。这些功能和模型包括：","我们还将发布有关基础方法的更多技术细节。","在此申请早期访问：https://tally.so/r/44d9NX","我们才刚刚开始触及多功能、高能力、统一的多模态模型的表面，以及它们将带来的可能性。从交互式图像和视频编辑、模拟，到计算机使用和物理 AI，前沿领域正广阔无垠。在逐步推出这些新功能的同时，我们已经在开发下一代模型。我们的目标是在同一个统一模型中实现感知、动作和语言预测的统一。","如果您有兴趣探索和构建 FLUX 3，请在此联系我们。如果您有兴趣参与我们的使命，请加入我们！我们正在德国和美国招聘：https://bfl.ai/careers"],"translationStatus":"translated","bodyOrigin":"source-page","editorial":{"summary":"Black Forest Labs宣布推出多模态基础模型FLUX 3，并开放Early Access。该模型在统一架构中共同学习图像、视频和音频，目标是形成对现实世界的综合表示。","background":"材料称，图像提供空间结构，视频补充时间变化，音频揭示视觉难以捕捉的机械现象与声学关系，语言则连接感知、目标、抽象概念和指令。FLUX 3建立在Self-Flow方法之上。","viewpoint":"Aioga判断，FLUX 3的核心看点不是单独增加模态，而是尝试让不同模态相互约束，以支持对物体、运动和事件的联合建模。不过，现有材料仅披露方向与早期结果，尚不足以判断实际性能。","implications":"材料显示，FLUX 3能够混合模态，并联合生成图像以及视频和音频，可由纯文本提示驱动，也可使用图像或视频作为输入参考。其对内容创作和物理AI的潜在影响值得关注，但具体应用效果仍需观察。","nextStep":"值得关注Early Access阶段公开的实际案例、评测和使用边界，尤其是图像、视频与音频协同生成是否稳定，以及其在内容创作和物理AI场景中的表现。材料目前未提供用户规模、性能数字或商业化安排。","evidenceRefs":["title","summary","articleBody","source"],"status":"published","aiGenerated":true,"autoApproved":true,"generatedBy":"aioga-editorial:gpt-5.6-sol","reviewedBy":"aioga-editorial-review:gpt-5.6-sol","generatedAt":"2026-07-27T15:54:26.661Z","sourceHash":"36631dc68de85b62","review":{"approved":true,"groundedness":95,"clarity":91,"duplicationRisk":12,"blockingIssues":[],"notes":["可将“联合生成图像以及视频和音频”改为“生成图像，并联合生成视频与音频”，以减少语义歧义。","“Early Access阶段公开的实际案例”属于前瞻性关注点，建议理解为待观察事项，而非来源已确认的安排。"]},"validation":{"passed":true,"mode":"ai-auto","revisions":0,"checks":["schema","length","source-attribution","low-source-overlap","no-html","independent-ai-review"]}},"tags":["模型更新","Black Forest Labs：Blog（网页）"],"translations":{"zh-CN":{"title":"Flux 3 多模态前沿模型发布，支持图像/视频/音频联合学习","summary":"Flux 3 多模态前沿模型发布，通过联合学习图像、视频和音频构建统一的世界表征。该模型现已开放 Early Access。","category":"模型更新","source":"bfl.ai","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Flux 3 多模态前沿模型发布，支持图像/视频/音频联合学习 - Aioga AI资讯","description":"Flux 3 多模态前沿模型发布，通过联合学习图像、视频和音频构建统一的世界表征。该模型现已开放 Early Access。","url":"https://www.aioga.com/news/cms3dpvqp0aorro3fx8vn2b5w/"},"en":{"title":"Flux 3 multimodal cutting-edge model released, supporting joint learning of images/videos/audio","summary":"Flux 3 multi-modal frontier model released, building a unified world representation through joint learning of images, videos, and audio. The model is now open for Early Access.","category":"Models","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Flux 3 multimodal cutting-edge model released, supporting joint learning of images/videos/audio - Aioga AI News","description":"Flux 3 multi-modal frontier model released, building a unified world representation through joint learning of images, videos, and audio. The model is now open for Early Access.","url":"https://www.aioga.com/en/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:21:39.991Z"},"ja":{"title":"Flux 3 多モーダル先端モデルが発表され、画像/動画/音声の統合学習をサポート","summary":"Flux 3 多モーダル最先端モデルが公開され、画像、動画、音声を共同で学習して統一された世界表現を構築します。このモデルは現在、Early Accessで利用可能です。","category":"モデル更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Flux 3 多モーダル先端モデルが発表され、画像/動画/音声の統合学習をサポート - Aioga AIニュース","description":"Flux 3 多モーダル最先端モデルが公開され、画像、動画、音声を共同で学習して統一された世界表現を構築します。このモデルは現在、Early Accessで利用可能です。","url":"https://www.aioga.com/ja/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:21:45.887Z"},"ko":{"title":"Flux 3 다중 모드 첨단 모델 발표, 이미지/비디오/오디오 공동 학습 지원","summary":"Flux 3 다중 모달 최첨단 모델 발표, 이미지, 비디오 및 오디오를 통합 학습하여 통합된 세계 표현 구축. 이 모델은 현재 Early Access를 오픈함.","category":"모델 업데이트","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Flux 3 다중 모드 첨단 모델 발표, 이미지/비디오/오디오 공동 학습 지원 - Aioga AI 뉴스","description":"Flux 3 다중 모달 최첨단 모델 발표, 이미지, 비디오 및 오디오를 통합 학습하여 통합된 세계 표현 구축. 이 모델은 현재 Early Access를 오픈함.","url":"https://www.aioga.com/ko/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:22:24.781Z"},"es":{"title":"Se publica el modelo de vanguardia multimodal Flux 3, que admite el aprendizaje conjunto de imágenes/videos/audio","summary":"El modelo avanzado multimodal Flux 3 ha sido lanzado, construyendo una representación unificada del mundo mediante el aprendizaje conjunto de imágenes, videos y audio. Este modelo ya está disponible en Acceso Anticipado.","category":"Modelos","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Se publica el modelo de vanguardia multimodal Flux 3, que admite el aprendizaje conjunto de imágenes/videos/audio - Aioga Noticias de IA","description":"El modelo avanzado multimodal Flux 3 ha sido lanzado, construyendo una representación unificada del mundo mediante el aprendizaje conjunto de imágenes, videos y audio. Este modelo...","url":"https://www.aioga.com/es/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:22:28.917Z"},"fr":{"title":"Publication du modèle de pointe multimodal Flux 3, prenant en charge l'apprentissage conjoint d'images/vidéos/audio","summary":"Le modèle de pointe multimodal Flux 3 a été publié, construisant une représentation unifiée du monde grâce à l'apprentissage conjoint des images, vidéos et audios. Ce modèle est désormais disponible en Early Access.","category":"Modèles","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Publication du modèle de pointe multimodal Flux 3, prenant en charge l'apprentissage conjoint d'images/vidéos/audio - Aioga Actualités IA","description":"Le modèle de pointe multimodal Flux 3 a été publié, construisant une représentation unifiée du monde grâce à l'apprentissage conjoint des images, vidéos et audios. Ce modèle est dé...","url":"https://www.aioga.com/fr/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:23:13.365Z"},"de":{"title":"Flux 3 multimodales Spitzenmodell veröffentlicht, unterstützt gemeinsames Lernen von Bildern/Videos/Audio","summary":"Flux 3 multimodales Spitzenmodell veröffentlicht, erstellt durch gemeinsames Lernen von Bildern, Videos und Audio eine einheitliche Weltrepräsentation. Das Modell ist jetzt im Early Access verfügbar.","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Flux 3 multimodales Spitzenmodell veröffentlicht, unterstützt gemeinsames Lernen von Bildern/Videos/Audio - Aioga KI-News","description":"Flux 3 multimodales Spitzenmodell veröffentlicht, erstellt durch gemeinsames Lernen von Bildern, Videos und Audio eine einheitliche Weltrepräsentation. Das Modell ist jetzt im Earl...","url":"https://www.aioga.com/de/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:23:05.084Z"},"pt-BR":{"title":"Lançamento do modelo multimodal avançado Flux 3, suportando aprendizado conjunto de imagens/vídeos/áudio","summary":"O modelo de ponta multimodal Flux 3 foi lançado, construindo uma representação unificada do mundo através do aprendizado conjunto de imagens, vídeos e áudios. Este modelo já está disponível em Acesso Antecipado.","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Lançamento do modelo multimodal avançado Flux 3, suportando aprendizado conjunto de imagens/vídeos/áudio - Aioga Notícias de IA","description":"O modelo de ponta multimodal Flux 3 foi lançado, construindo uma representação unificada do mundo através do aprendizado conjunto de imagens, vídeos e áudios. Este modelo já está d...","url":"https://www.aioga.com/pt-BR/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:23:55.714Z"},"ru":{"title":"Выпущена мультимодальная передовая модель Flux 3, поддерживающая совместное обучение с изображениями, видео и аудио","summary":"Flux 3 представлен как передовая мультимодальная модель, которая через совместное обучение изображений, видео и аудио строит единое представление мира. Эта модель теперь доступна в раннем доступе.","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Выпущена мультимодальная передовая модель Flux 3, поддерживающая совместное обучение с изображениями, видео и аудио - Aioga Новости ИИ","description":"Flux 3 представлен как передовая мультимодальная модель, которая через совместное обучение изображений, видео и аудио строит единое представление мира. Эта модель теперь доступна в...","url":"https://www.aioga.com/ru/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:23:52.458Z"},"ar":{"title":"تم إصدار نموذج Flux 3 متعدد الوسائط المتقدم، يدعم التعلم المشترك للصور/الفيديو/الصوت","summary":"تم إطلاق نموذج Flux 3 متعدد الوسائط الرائد، من خلال التعلم المشترك للصور والفيديو والصوت لبناء تمثيل موحد للعالم. أصبح النموذج متاحًا الآن للوصول المبكر.","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"تم إصدار نموذج Flux 3 متعدد الوسائط المتقدم، يدعم التعلم المشترك للصور/الفيديو/الصوت - Aioga أخبار الذكاء الاصطناعي","description":"تم إطلاق نموذج Flux 3 متعدد الوسائط الرائد، من خلال التعلم المشترك للصور والفيديو والصوت لبناء تمثيل موحد للعالم. أصبح النموذج متاحًا الآن للوصول المبكر.","url":"https://www.aioga.com/ar/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:24:42.228Z"},"hi":{"title":"Flux 3 बहु-प्रकार की अग्रणी मॉडल का विमोचन, छवि/वीडियो/ऑडियो संयुक्त शिक्षा का समर्थन करता है","summary":"Flux 3 बहु-मोडल अग्रणी मॉडल जारी किया गया, जो छवि, वीडियो और ऑडियो को संयुक्त रूप से सीखकर एक एकीकृत विश्व प्रतिनिधित्व बनाता है। यह मॉडल अब प्रारंभिक पहुंच के लिए खुला है।","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Flux 3 बहु-प्रकार की अग्रणी मॉडल का विमोचन, छवि/वीडियो/ऑडियो संयुक्त शिक्षा का समर्थन करता है - Aioga AI समाचार","description":"Flux 3 बहु-मोडल अग्रणी मॉडल जारी किया गया, जो छवि, वीडियो और ऑडियो को संयुक्त रूप से सीखकर एक एकीकृत विश्व प्रतिनिधित्व बनाता है। यह मॉडल अब प्रारंभिक पहुंच के लिए खुला है।","url":"https://www.aioga.com/hi/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:24:45.904Z"},"it":{"title":"Rilasciato Flux 3, modello all'avanguardia multimodale, supporta l'apprendimento congiunto di immagini/video/audio","summary":"Il modello all'avanguardia multimodale Flux 3 è stato rilasciato, costruendo una rappresentazione unificata del mondo attraverso l'apprendimento combinato di immagini, video e audio. Il modello è ora disponibile in Early Access.","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Rilasciato Flux 3, modello all'avanguardia multimodale, supporta l'apprendimento congiunto di immagini/video/audio - Aioga Notizie IA","description":"Il modello all'avanguardia multimodale Flux 3 è stato rilasciato, costruendo una rappresentazione unificata del mondo attraverso l'apprendimento combinato di immagini, video e audi...","url":"https://www.aioga.com/it/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:25:29.200Z"},"nl":{"title":"Flux 3 multimodaal geavanceerd model uitgebracht, ondersteunt gezamenlijk leren van afbeeldingen/video/audio","summary":"Flux 3 multi-modale geavanceerde model uitgebracht, bouwt een uniforme wereldrepresentatie door gezamenlijk leren van afbeeldingen, video's en audio. Het model is nu beschikbaar via Early Access.","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Flux 3 multimodaal geavanceerd model uitgebracht, ondersteunt gezamenlijk leren van afbeeldingen/video/audio - Aioga AI-nieuws","description":"Flux 3 multi-modale geavanceerde model uitgebracht, bouwt een uniforme wereldrepresentatie door gezamenlijk leren van afbeeldingen, video's en audio. Het model is nu beschikbaar vi...","url":"https://www.aioga.com/nl/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:25:30.150Z"},"tr":{"title":"Flux 3 çok modlu ileri düzey model yayınlandı, görüntü/video/ses birleşik öğrenmeyi destekliyor","summary":"Flux 3 çok modlu öncü model yayınlandı, görüntü, video ve sesi birleştirerek birleşik bir dünya temsili oluşturuyor. Bu model artık Erken Erişim için açık.","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Flux 3 çok modlu ileri düzey model yayınlandı, görüntü/video/ses birleşik öğrenmeyi destekliyor - Aioga AI Haberleri","description":"Flux 3 çok modlu öncü model yayınlandı, görüntü, video ve sesi birleştirerek birleşik bir dünya temsili oluşturuyor. Bu model artık Erken Erişim için açık.","url":"https://www.aioga.com/tr/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:26:09.627Z"},"vi":{"title":"Flux 3 ra mắt mô hình tiên tiến đa phương thức, hỗ trợ học kết hợp hình ảnh/video/âm thanh","summary":"Mô hình tiên tiến đa phương thức Flux 3 được ra mắt, xây dựng biểu diễn thế giới thống nhất thông qua việc học kết hợp hình ảnh, video và âm thanh. Mô hình này hiện đã mở truy cập sớm (Early Access).","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Flux 3 ra mắt mô hình tiên tiến đa phương thức, hỗ trợ học kết hợp hình ảnh/video/âm thanh - Tin tức AI Aioga","description":"Mô hình tiên tiến đa phương thức Flux 3 được ra mắt, xây dựng biểu diễn thế giới thống nhất thông qua việc học kết hợp hình ảnh, video và âm thanh. Mô hình này hiện đã mở truy cập...","url":"https://www.aioga.com/vi/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:26:16.023Z"},"id":{"title":"Flux 3 model terdepan multimodal dirilis, mendukung pembelajaran gabungan gambar/video/audio","summary":"Flux 3, model canggih multimodal, diluncurkan, membangun representasi dunia yang terpadu melalui pembelajaran gabungan gambar, video, dan audio. Model ini sekarang telah tersedia untuk Early Access.","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Flux 3 model terdepan multimodal dirilis, mendukung pembelajaran gabungan gambar/video/audio - Berita AI Aioga","description":"Flux 3, model canggih multimodal, diluncurkan, membangun representasi dunia yang terpadu melalui pembelajaran gabungan gambar, video, dan audio. Model ini sekarang telah tersedia u...","url":"https://www.aioga.com/id/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:26:54.216Z"},"th":{"title":"Flux 3 เปิดตัวโมเดลแนวหน้าหลายโหมด รองรับการเรียนรู้ร่วมของภาพ/วิดีโอ/เสียง","summary":"Flux 3 โมเดลแนวหน้าหลายโหมดเปิดตัว ผ่านการเรียนรู้ร่วมกันของภาพ วิดีโอ และเสียง เพื่อสร้างการแทนโลกแบบรวม โมเดลนี้เปิดให้เข้าถึงช่วง Early Access แล้ว","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Flux 3 เปิดตัวโมเดลแนวหน้าหลายโหมด รองรับการเรียนรู้ร่วมของภาพ/วิดีโอ/เสียง - ข่าว AI Aioga","description":"Flux 3 โมเดลแนวหน้าหลายโหมดเปิดตัว ผ่านการเรียนรู้ร่วมกันของภาพ วิดีโอ และเสียง เพื่อสร้างการแทนโลกแบบรวม โมเดลนี้เปิดให้เข้าถึงช่วง Early Access แล้ว","url":"https://www.aioga.com/th/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:27:13.209Z"},"pl":{"title":"Wydano model Flux 3 wielomodalny, wspierający wspólne uczenie się obrazów/wideo/dźwięku","summary":"Opublikowano Flux 3, wielomodalny model nowej generacji, który poprzez wspólne uczenie obrazów, wideo i dźwięku tworzy zunifikowaną reprezentację świata. Model jest już dostępny w ramach Early Access.","category":"模型更新","source":"Hacker News 热门（buzzing.cc 中文翻译）","aggregationSource":"Hacker News 热门（buzzing.cc 中文翻译）","pageTitle":"Wydano model Flux 3 wielomodalny, wspierający wspólne uczenie się obrazów/wideo/dźwięku - Aioga Wiadomości AI","description":"Opublikowano Flux 3, wielomodalny model nowej generacji, który poprzez wspólne uczenie obrazów, wideo i dźwięku tworzy zunifikowaną reprezentację świata. Model jest już dostępny w...","url":"https://www.aioga.com/pl/news/cms3dpvqp0aorro3fx8vn2b5w/","contentTranslated":true,"sourceHash":"41d9a316ebe0faa4","translatedAt":"2026-07-27T01:27:55.938Z"}}}}