Daily Tech Briefing
AI 科技速览
每天 5 分钟内学习 AI。获取最新的人工智能新闻,理解其重要性,并学习如何将其应用于您的工作。
Latent Space · 2026/7/24 04:30:12
![[AINews] Black Forest Labs FLUX 3 - Multimodal Flow Models that beat Seedance 2.0, Gemini Omni and Grok Imagine, and FLUX-mimic video-action robotics model](https://substackcdn.com/image/fetch/$s_!3n0x!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F__ss-rehost__tw-video-preview-13_2080308957898481664.jpg)
[AINews] Black Forest Labs FLUX 3 - Multimodal Flow Models that beat Seedance 2.0, Gemini Omni and Grok Imagine, and FLUX-mimic video-action robotics model
AI 中文解读
Black Forest Labs今天放出大招,正式推出FLUX 3视频生成模型。就在OpenAI和Anthropic忙着打语音助手口水仗的当口,这个开源模型直接抢走了全场风头,性能全面对标甚至超越Seedance 2.0、Gemini Omni等一线闭源产品。
简单来说,FLUX 3就像一个全能视频导演。你只要输入一段文字,就能生成带原生声音的视频;上传一张照片,照片就能动起来;给它一段老视频,它还能把主角抽出来放进全新场景。更酷的是,它可以按照你指定的关键画面自动补全中间过程,把零散片段拼接成完整故事,还能搞定动画片、纪实录像、电影大场面等多种风格,甚至能生成精准的字体和动态设计。普通人想做出专业级视频,门槛一下子被拉到了最低。
最值得关注的是,团队宣布面向开发者的开放权重版本即将上线。这意味着全球开发者都能基于这套模型开发自己的AI视频工具。对于普通用户来说,今后创作短视频、动画,可能就像用手机拍照一样简单,而且成本极低。另外团队还透露了名为FLUX3-mimic的机器人模型,能通过视频学习人类动作,未来机器人模仿人类做事将不再是科幻片桥段。
Thursdays are the heaviest days for AI releases, and even though OpenAI scored a victory over Anthropic in launching the new ChatGPT Voice (consumer) and OpenAI Presence (enterprise) and getting more impressions than Claude Voice today (a completely accidental coincidence in timing, we are sure), neither seem as monumental as BFL’s launch of FLUX 3 Video today:We last covered BFL in our very well received Anjney Midha podcast:$5000 w…","cta":null,"showBylines":true,"showDescription":true,"showImage":true,"size":"sm","isEditorNode":true,"title":"The Professor of Outputmaxxing — Anjney Midha, AMP","publishedBylines":[],"post_date":"2026-06-18T17:30:00.811Z","cover_image":"https://substack-video.s3.amazonaws.com/video_upload/post/202359797/8dbbb3fa-e808-473c-af72-b9aee4fe0026/transcoded-1781652240.png","cover_image_alt":null,"canonical_url":"https://www.latent.space/p/anj","section_name":null,"video_upload_id":null,"id":202359797,"type":"podcast","reaction_count":22,"comment_count":4,"publication_id":1084089,"publication_name":"Latent.Space","publication_logo_url":"https://substackcdn.com/image/fetch/$s_!DbYa!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F73b0838a-bd14-46a1-801c-b6a2046e5c1e_1130x1130.png","belowTheFold":false,"youtube_url":null,"show_links":null,"feed_url":null}">Most GenMedia people will remember the BFL homepage when they initially launched Flux 1 in 2024, hinting at video models next, with their logo in a forest. Well, 2 years later, it’s finally real:The blogpost outlines Self Flow, covering ALL their modalities together with strong preference claims: “Its core capabilities include the following (all outputs come with native audio generation):Text-to-video generation.Image-to-video generation, either continuing from a starting frame (“animation”) or using images as visual references.Video-to-video generation from a reference clip, carrying central elements of a source video - for instance the same character - into a new scene or context.Generative video-audio continuation from input video and audio.Keyframe-to-video generation for controlled transitions between defined moments.Multilingual dialogue.A broad range of visual styles and aspect ratios, extending far beyond conventional cinematic output.Agentic chaining of individual clips into longer, multi-shot sequences.High style diversity -- FLUX 3 Video easily handles ranges of styles from candid camcorder footage to animation and cinematics.Strong typography generation and animated designs.”Some of the above are SOTA features from other frontier lab models, like we discussed in our Grok Imagine pod, so the community has very much been put on notice that there has now been independent, perhaps SOTA, reproduction of these capabilities, with an open weights Dev version on the way.As if this release wasn’t enough, the team also announced FLUX3-mimic, which proves that the FLUX 3 model is learning a sufficient world model capable of driving robots…@mimicrobotics was one of the first partners to gain early access to FLUX 3. Together we developed FLUX-mimic, a video-action model combining the FLUX 3 backbone with mimic's expertise in robot learning for dexterous","username":"bfl_ai","name":"Black Forest Labs","profile_image_url":"https://pbs.substack.com/profile_images/1954888731053142016/NDyG-4-j_normal.jpg","date":"2026-07-23T15:08:16.000Z","photos":[],"quoted_tweet":{},"reply_count":2,"retweet_count":7,"like_count":138,"impression_count":14471,"expanded_url":null,"video_url":null,"video_preview_media_key":null,"belowTheFold":true}" data-component-name="Twitter2ToDOM">… and predicting their impact in real factory settings…AI News for 7/22/2026-7/23/2026. We checked 12 subreddits, 544 Twitters and no further Discords. AINews’ website lets you search all past issues. As a reminder, AINews is now a section of Latent Space.
分享
阅读原文 ↗