# Google 發布 Gemini Omni 1.1 Flash，支援可控影片延伸與首尾影格生成

> 📖 本站完整內容索引（documentation index）：[llms.txt](/llms.txt)

> 原作者：Google (@Google) · 策展與摘要：EasyVibeCoding · 平台：X (Twitter) · 熱度：🔥🔥 · 日期：2026-09-01

> 原始來源：https://x.com/Google/status/2094513065018499355

## 證據與延伸閱讀

- [Google 發布 Gemini Omni 1.1 Flash，支援可控影片延伸與首尾影格生成。](https://blog.google/innovation-and-ai/technology/developers-tools/build-with-gemini-omni-1-1-flash) — 官方文件 · 最後核對：2026-09-01 · 支持主張：The official post adds up to 10 seconds of prior-scene context and 10-second extensions up to 40 seconds, 360p draft generation, high-resolution output, and up to three seconds of reference video. It also documents a Gemini API interaction using previous_interaction_id and identifies Google AI Studio, the Gemini Enterprise Agent Platform/API, Google Flow, and the Gemini app as rollout surfaces.
- [blog.google/innovation-and-ai/technology/developers-tools](https://blog.google/innovation-and-ai/technology/developers-tools/build-with-gemini-omni-1-1-flash/) — 官方文件
- [truescho.com/en/blog/gemini-omni-1-1-flash-2026](https://truescho.com/en/blog/gemini-omni-1-1-flash-2026)

## 中文摘要

Google 發布 Gemini Omni 1.1 Flash，支援可控影片延伸與首尾影格生成。

**核心更新**　Google DeepMind 產品經理 Anish Nangia 與 Alisa Fortin 在 2026 年 8 月 27 日的官方文章中介紹 Gemini Omni 1.1 Flash，定位為 production-ready 的開發者更新。這次重點不只是生成影片，而是讓開發者能更精準地控制影片如何延續、轉場、迭代與提升畫質。官方說明可參考[〈Build with Gemini Omni 1.1 Flash〉](https://blog.google/innovation-and-ai/technology/developers-tools/build-with-gemini-omni-1-1-flash)。

**場景延伸**　Gemini Omni 1.1 Flash 能分析既有影片最多 10 秒的前段內容，再從場景結尾繼續生成後續片段。Google 將此描述為相較 Veo 只參考最後一秒的進展，目標是提升視覺一致性與敘事銜接，讓創作者可以延長故事，或從既有片段分支出新的創作方向。

- 影片可用 10 秒為一個增量延伸，累計總長度最多 40 秒。
- Gemini API 可透過 `previous_interaction_id` 指定前一次影片互動，接續生成下一段內容。
- 官方示例涵蓋鏡頭拉遠、dolly-zoom、snap-zoom，以及圍繞靜止角色進行 360 度環繞等連續鏡頭。

官方文件中的 API 互動範例如下：

```python
from google import genai
client = genai.Client()
interaction = client.interactions.create(
model="gemini-omni-1.1-flash",
previous_interaction_id=previous_video_interaction.id,
input=[
{"type": "text", "text": "Continue the scene."}
],
response_format={
"resolution": "360p",
},
)
```

**首尾影格控制**　開發者可以指定一個鏡頭的起始影格與結束影格，讓模型在兩個關鍵影格之間生成連續影片。這項功能特別適合需要控制攝影機路徑的內容，例如複雜的環繞、推近、拉遠、快速搖攝或無縫循環。官方示例要求維持同一個連續鏡頭、避免跳接，並展示從鼓手轉向薩克斯風手與舞者、再回到原場景的轉場設計。

**低成本迭代**　Gemini Omni 1.1 Flash 支援以 360p 產生輕量預覽，方便開發者在分鏡、提示與鏡頭設計階段快速測試。Google 表示，360p 預覽相較標準 720p 生成速度最多快 60%，成本則是 720p 的三分之一，適合一次產生 3 至 4 個草稿、每次只調整一個因素，再並排比較不同版本。不過，這兩項數字是 Google 以 360p 與 720p 的系統吞吐量所做的比較，並非獨立 benchmark，不能直接視為所有環境都能達到的效能。

**高解析度與影片參考**　完成創意探索後，模型可輸出 1080p 或 4K 影片，讓低解析度草稿與正式製作之間形成分階段流程。多模態輸入則可加入最長 3 秒的影片參考，協助維持角色、動作與視覺脈絡。官方示例讓狗、章魚與熊分別模仿不同舞者影片的古典舞、嘻哈與霹靂舞，並在同一個開放空間中完成一鏡到底的表演；這顯示影片參考可用於動作條件與角色一致性，但不代表模型能在所有片段中穩定重現相同品質。

**產品與部署範圍**　Gemini Omni 1.1 Flash 正在 Google 的開發者生態系統中推出，開發者可從下列管道開始測試或規劃整合：

- Google AI Studio：直接試用 Gemini Omni 1.1。
- Gemini Enterprise Agent Platform：透過 Agent Platform API 建置企業應用。
- Gemini API：將場景延伸、影片參考與升頻功能整合進自有工作流程。
- Google Flow：全球 Google AI Plus、Pro 與 Ultra 訂閱者自當日起可使用 Omni 1.1。
- Gemini app：全球 Google AI Plus、Pro 與 Ultra 訂閱者可使用場景延伸功能。

Google 也提供官方文件、cookbook 與 prompting guides，協助開發者整合上述控制功能。既有的產品採用案例包括 Adobe Firefly、Figma Weave、GMI Cloud 與 Runway；Figma Weave 創意總監 Itay Schiff 認為，延伸、更豐富的參考素材與 4K 解析度，讓團隊從「生成影片」進一步走向「導演影片」。GMI Cloud 行銷副總裁 Louisa Guo 則強調細節準確度，認為對教育與解說內容而言，可靠性比單一功能更重要。

**實際限制**　Google 宣稱 Gemini Omni 1.1 Flash 能帶來更好的視覺一致性與敘事遵循，但目前提供的媒體畫面只展示一個停車場場景中、以合成方式呈現的黑色兔狀物件；單一影格只能證明示例媒體存在，無法驗證長距離一致性、時間連續性、模型身分或完整影片行為。因此，開發者在規劃正式整合時，仍應分別核對 Google AI Studio、Gemini Enterprise Agent Platform、Google Flow 與 Gemini app 的實際可用範圍，並以自身素材測試場景延伸、轉場穩定性與 4K 輸出品質。

## 標籤

新產品, 教學資源, SDK, Gemini Omni 1.1 Flash, Google DeepMind, Gemini API, Python, Anish Nangia, Alisa Fortin
