Google 发布 Gemini Omni 1.1 Flash:40 秒场景延伸 + 4K 视频生成 + stateful session,按秒计费 0.03-0.30 美元

Google 发布 Gemini Omni 1.1 Flash:40 秒场景延伸 + 4K 视频生成 + stateful session,按秒计费 0.03-0.30 美元

qidao123.com企服评测-应用市场2026-08-29T09:00:00.000+080013
Google 8-27 GA 上线 Gemini Omni 1.1 Flash:40 秒场景延伸(10 秒前情分析)、首尾帧插值、360p 草稿、4K 放大、多轮 stateful session;GA 定价 360p $0.03/秒、720p $0.10/秒、1080p $0.15/秒、4K $0.30/秒;Arena 评测文生视频第一。
【新闻原文】Gemini Omni 1.1 Flash lets you build with more control — Google 推出生产级视频生成模型 【来源:Google 官方博客 2026-08-27】 Today, we're introducing Gemini Omni 1.1 Flash, a new suite of creative controls and generative video capabilities to support developers. Gemini Omni brought real-world reasoning to generative creation, and today's updates make Omni 1.1 production-ready for professional use via the Gemini API in Google AI Studio. Whether you're building generative video workflows, creative tools, or media editing software, these updates make generative video more controllable, faster to iterate on, and polished for real-world deployment. Here's a look at what's new: **Extend scenes for longer storytelling.** Scene extension allows you to take an existing video and continue generating footage seamlessly from where it left off. With Omni 1.1, the model can now analyze up to 10 seconds of prior context — a leap from previous models that only referenced the final second. The result is improved visual consistency and narrative adherence, letting you build longer stories or branch into new creative directions. You can extend videos in 10-second increments up to a total cumulative length of 40 seconds. **First and last frame interpolation.** Developers can now specify the starting and ending frames of a shot, and the model will generate the motion between them, creating smooth, professional camera movements and transitions. **Faster prototyping with 360p drafts.** Use 360p previews to iterate faster and save money during your creative process. You can generate lightweight previews at 360p, then upscale your favorites to 720p, 1080p or 4K. 360p generation can be up to 60 percent faster and cost one third as much as standard 720p generation. **Crisp 4K upscaling.** Upscale your final projects to 4K resolution for a polished, professional look. Output resolutions include 360p, 720p, 1080p and 4K; 4K is achieved through upscaling, and 720p is the default output. **Multi-turn stateful sessions.** The model supports multi-turn interactions with stateful session IDs, meaning an agent can plan a sequence of shots, extend scenes in controlled increments, evaluate intermediate outputs in 360p draft mode, and only commit to expensive high-resolution rendering once the sequence meets quality thresholds. The model ID is gemini-omni-1.1-flash. The earlier gemini-omni-flash-preview endpoint is scheduled for deprecation on September 30, 2026. Access is available via the Gemini API in Google AI Studio, the Gemini Enterprise Agent Platform, and for eligible AI Plus, Pro and Ultra subscribers in Google Flow and the Gemini app. Pricing per second of generated video: 360p $0.03; 720p $0.10; 1080p $0.15; 4K $0.30. In independent Arena human evaluation, Gemini Omni 1.1 Flash ranked first in text-to-video generation, and second in image-to-video (behind the open-source MiniMax H3). Google says the new control features — scene extension, frame interpolation, draft mode, and stateful multi-turn sessions — transform generative video from a creative toy into an addressable component within autonomous agent architectures. 【企服观察】Gemini Omni 1.1 Flash 把"AI 视频"从 Demo 变组件:Google 在抢 Agent 时代的"视频 API 权" qidao123.com 企服评测-应用市场观察 8月27日 Google 把 Gemini Omni 升级到 1.1 Flash GA(正式商用),表面是“又一款视频生成模型”,实际上 Google 在用这款产品做一件更大的事——把生成式视频变成 Agent 工作流里的"可调用 API"。 一、四项核心升级里,藏着 Agent 时代的秘密。逐条看: - 1. **10 秒上下文 → 40 秒场景延伸**:老模型只能看最后一秒,新模型能看 10 秒前情。这意味着 Agent 可以在"剧情一致性"这个老难题上做更长链的拆解,不用每 3 秒重新生成。 - 2. **首尾帧插值**:开发者指定"第一帧是这个画面,最后帧是那个画面",模型生成中间过程。这给 Agent 提供了"按关键帧规划镜头语言"的能力——这是电影/广告行业的基本功,过去只有真人导演能做。 - 3. **360p 草稿 + 4K 放大**:360p 生成速度提升 60%、成本只有 720p 的 1/3。Agent 可以先低成本跑多版草稿 → 评估 → 再 upscale 到 4K 终稿。 - 4. **多轮 stateful session**:这是最关键的一条。Agent 可以"持有"一段视频的会话 ID,连续调用 5-10 次,模型能记住前几次的镜头、人物、光照。这把视频生成从"一次性 API"变成了"可长程编辑的组件"。 二、定价策略:按秒计费,让 Agent 真的用得起。 - 360p $0.03/秒(约 ¥4.78/10 秒) - 720p $0.10/秒(约 ¥15.94/10 秒) - 1080p $0.15/秒(约 ¥23.92/10 秒) - 4K $0.30/秒(约 ¥47.83/10 秒) - 一段 30 秒的 1080p 视频,成本 4.5 美元;一段 30 秒的 4K 视频,成本 9 美元。这个价格,对广告投放、电商素材、社交短视频这些场景来说,是真的可以进生产环境的。 三、为什么 Google 这次"不再发 Demo,直接给 API"?因为视频生成赛道 2026 年的胜负手,不在"谁的演示更炫",而在"谁能最快被 Agent 调用"。Google 这一波操作,至少压对了三件事: - 1. **抢 Agent 工作流的"视频 API 权"**:当 Anthropic、OpenAI、Salesforce 的 Agent 开始做"营销活动自动化"时,必然要调用视频生成能力。Google 现在把 API 做好、价格定低、文档给齐,等于提前锁定了"被集成"的位置。 - 2. **压住 Runway、Pika、Luma 这些垂直玩家**:Runway Gen-4 单价 0.5 美元/秒,OpenAI Sora 0.3 美元/秒,Google 直接干到 0.03-0.30 美元/秒——把价格地板打穿。垂直玩家要么降价(毛利撑不住),要么打差异化(被 Google 4K 跨级压制)。 - 3. **为 Gemini Enterprise 拉新**:Omni 1.1 Flash 直接进 Google AI Studio + Gemini Enterprise Agent Platform,相当于给企业版 AI 套件"加了一个杀手级能力"——这对冲 Anthropic Claude 和 OpenAI ChatGPT Enterprise 的冲击力。 四、对中国视频生成厂商的三点真话: - 1. **快手可灵、字节豆包、阿里通义、MiniMax 必须在 90 天内推出"按秒计费 + stateful session"**。这是 Gemini 1.1 Flash 立下的新标准,错过这个窗口期的产品,会被企业 Agent 客户直接划到"上一代"。 - 2. **首尾帧插值是视频 Agent 的核心接口**。可灵和 MiniMax H3 都有类似能力,但要尽快把"指定首尾帧 + 中间自动生成"做成标准 API 调用,而不是埋在交互界面里。 - 3. **4K 不是越高越好,360p 草稿才重要**。Gemini 的定价模型(360p 草稿便宜、4K 终稿贵)会成为行业新范式——国内厂商的"全 1080p/4K 输出"路线,会被 Agent 时代的"草稿+终稿"两段式定价压低毛利。 五、对 ToB 厂商的四个具体场景建议: - 1. **跨境电商**:把 Gemini Omni 1.1 Flash 接入 Shopify 插件,自动生成 30 秒短视频广告,30 秒成本 9 美元 = 60 多元人民币,比真人拍摄便宜 95%。 - 2. **B2B SaaS 营销**:用 360p 草稿给客户做"我的产品能拍 100 种风格"的批量 demo,再用 4K 终稿做最终投放——获客成本直接降一半。 - 3. **企业内部培训**:把入职手册 / 合规培训 / 销售话术输入 Agent,让 Agent 一次性生成 5-10 段 1080p 短视频,比外采视频公司便宜 90% 以上。 - 4. **金融行业投顾**:让 Agent 用 Gemini Omni 1.1 Flash 自动生成"行业研报可视化视频",每天 100 条投顾短视频,总成本不到 1000 美元。 六、我们的判断:Omni 1.1 Flash 真正的护城河是"被集成"而非"被体验"。Google 这次 GA 走得比所有人预期的都早、定价都比所有人预期的都低——这说明 Google 已经把"AI 视频"从"研发故事"切换到"基础设施故事"。当一款 AI 模型按秒计费、可以被 Agent 批量调用、定价低于 0.1 美元/秒的时候,它就具备了"水电煤气"的属性——而这正是 Google 在 Gemini 1.0 Pro/1.5 Pro/2.0/2.5 上一直想做到、但始终被 OpenAI Sora 和 Runway Gen-3 抢了风头的位置。 未来 90 天,看三件事: - 1)OpenAI Sora 2 是否被迫降价、是否同步支持 stateful session; - 2)快手可灵 / 字节豆包 / 阿里通义是否在 Q4 把"按秒计费"做成默认; - 3)国内的 Agent 厂商(智谱、Anthropic 中国版、阿里通义、百度搭子)是否在 2026 年底前把"视频 Agent"做成标配。 - 这三件事如果都发生,2027 年的"AI Agent 内容工厂"会比所有人想象的更早成熟。

好内容,需要你的鼓励

评论

请评价真实、客观、有价值的内容
0/500

请回复有价值的信息,无意义的评论将很快被删除,账号将被禁止发言。

还没有评论,期待你的第一条客观评价
qidao123.com:企服评测·应用市场·商务社交-IT产业互联网
微信订阅号微信订阅号
微信服务号微信服务号
微信客服微信客服(加群)
小程序小程序
H5H5
关于我们
商务合作
本站浏览量:3672576 ©Copyright 2022-2026 杭州祥升科技有限公司 版权所有浙ICP备20004199号