diff --git a/docs.json b/docs.json
index 0d9b37054..56cd57ad6 100644
--- a/docs.json
+++ b/docs.json
@@ -4099,7 +4099,8 @@
{
"group": "Google",
"pages": [
- "zh/tutorials/partner-nodes/google/gemini-omni-flash"
+ "zh/tutorials/partner-nodes/google/gemini-omni-flash",
+ "zh/tutorials/partner-nodes/google/gemini-omni-flash/workflow"
]
},
{
@@ -7257,7 +7258,8 @@
{
"group": "Google",
"pages": [
- "ja/tutorials/partner-nodes/google/gemini-omni-flash"
+ "ja/tutorials/partner-nodes/google/gemini-omni-flash",
+ "ja/tutorials/partner-nodes/google/gemini-omni-flash/workflow"
]
},
{
@@ -10428,7 +10430,8 @@
{
"group": "Google",
"pages": [
- "ko/tutorials/partner-nodes/google/gemini-omni-flash"
+ "ko/tutorials/partner-nodes/google/gemini-omni-flash",
+ "ko/tutorials/partner-nodes/google/gemini-omni-flash/workflow"
]
},
{
diff --git a/ja/tutorials/partner-nodes/google/gemini-omni-flash.mdx b/ja/tutorials/partner-nodes/google/gemini-omni-flash.mdx
index 4b404f10b..3b2cda29e 100644
--- a/ja/tutorials/partner-nodes/google/gemini-omni-flash.mdx
+++ b/ja/tutorials/partner-nodes/google/gemini-omni-flash.mdx
@@ -1,97 +1,68 @@
---
title: "Gemini Omni Flash: 会話型ビデオ生成"
-description: "Gemini Omni Flashは、Googleのマルチモーダルビデオモデルです。パートナーノードを通じてComfyUIで利用でき、自然言語でビデオを生成・編集できます。"
+description: "Gemini Omni Flash 1.1は、Googleのマルチモーダルビデオモデルです。パートナーノードを通じてComfyUIで利用でき、自然言語でビデオを生成・編集できます。"
sidebarTitle: "Gemini Omni Flash"
-translationSourceHash: 9a3c952e
+translationSourceHash: 277b18df
translationFrom: tutorials/partner-nodes/google/gemini-omni-flash.mdx
translationBlockHashes:
- "_intro": 3b6973dc
- "What Gemini Omni Flash offers": d0140bc2
- "Workflows": 753af4cb
- "Get started": 64517938
+ "_intro": 5b9c88f4
+ "What Gemini Omni Flash 1.1 is good at": 92599abb
+ "Omni Flash 1.1 vs Omni Flash (preview)": ba3e2a73
+ "Pricing": da96966c
+ "Use it in ComfyUI": dae97960
---
import ReqHint from "/snippets/ja/tutorials/partner-nodes/req-hint.mdx";
import UpdateReminder from "/snippets/ja/tutorials/update-reminder.mdx";
-Gemini Omni Flashは、Google DeepMindの高品質でコスト効率の高いビデオ生成および会話型編集モデルです。Google I/O 2026でGemini Omniファミリーの一部として初めて発表され、Geminiのマルチモーダル推論とネイティブビデオ作成を組み合わせ、開発者が自然な会話を通じてビデオを生成、編集、リミックスできるようにします。
+Gemini Omni Flashは、Google DeepMindの会話型ビデオ生成・編集モデルです。Google I/O 2026で初めて発表されたGemini Omniファミリーの一部です。Geminiのマルチモーダル推論とネイティブなビデオ作成を組み合わせ、自然言語で欲しい内容を記述し、参照画像やビデオを添付すると、同期されたオーディオ付きのクリップを生成します。現在のバージョンであるGemini Omni Flash 1.1は2026年8月27日に一般提供(GA)され、キーフレーム補間、360p/4K出力オプション、シーン拡張が追加されました。
-## Gemini Omni Flashが提供する機能
+## Gemini Omni Flash 1.1が得意なこと
- **会話型ビデオ編集**: 自然言語を使用してビデオを洗練・編集します。キャラクターの入れ替え、シーンの再照明、アングルの変更、オブジェクトの追加・削除が可能で、オリジナルのオーディオとビデオトラックは保持されます。
-- **マルチモーダル入力**: テキスト、画像、ビデオ入力を組み合わせて生成をガイドします。すべてのビデオ出力に同期されたオーディオをネイティブ生成します。
+- **マルチモーダル入力**: テキスト、画像(最大14枚)、ビデオ(最大3本、各10秒)を組み合わせて生成をガイドします。すべての出力にネイティブのオーディオトラックが含まれます。
+- **キーフレーム補間**: `image_to_video`タスクでは、開始フレームとオプションの終了フレームを添付すると、その間の映像を生成します。
+- **参照からビデオ生成**: ``のようなタグで参照画像を役割に紐付け、画像の中のキャラクター、製品、オブジェクトを新しいシーンに登場させます。画像自体がフレームとして使われることはありません。
+- **シーン拡張**: `extend`タスクはクリップに最大10秒の新しい映像を追加します。直近10秒のコンテキストを読み取り、キャラクターと動きの一貫性を保ちながら、累計約40秒まで延長できます。
+- **解像度コントロール**: 360pで下書きし(コストは720pの約3分の1)、720p、1080p、4Kで再レンダリングします。16:9と9:16のアスペクト比に対応しています。
- **世界知識とシミュレーション**: 物理理解とGeminiの歴史、科学、文化的背景に関する知識を組み合わせ、フォトリアリズムを超えた意味のあるストーリーテリングを実現します。
- **テキストとアクションの同期**: ビデオに直接読みやすいテキストやグラフィックスをレンダリングし、動きのあるタイポグラフィを画面上の動きと同期させます。
-- **料金**: ビデオ出力1秒あたり0.10ドル。Veo 3.1 Fastと同じ料金設定です。
-## ワークフロー
+## Omni Flash 1.1 と Omni Flash(プレビュー)の比較
-### テキストから動画へ
+ComfyUIノードは、モデルのドロップダウンで両方のバージョンを提供しています:
-
-
- Comfy Cloudで開く
-
-
- JSONをダウンロードするか、テンプレートライブラリで「Gemini Omni Flash」を検索
-
-
+| | Gemini Omni Flash 1.1 | Gemini Omni Flash(プレビュー) |
+|---|---|---|
+| モデルID | `gemini-omni-1.1-flash` | `gemini-omni-flash-preview` |
+| 状態 | 一般提供中(2026年8月27日) | 公開プレビュー、2026年9月30日に提供終了 |
+| 解像度 | 360p、720p、1080p、4K | 固定720p、24 FPS |
+| タスク制御 | 明示的な`task_type`セレクター: auto、text_to_video、image_to_video、reference_to_video、edit、extend | 添付メディアから推論 |
+| 料金 | 秒あたり10.19〜91.76クレジット(解像度による) | 秒あたり30.57クレジット |
-
-
-自然言語のプロンプトから映画的なビデオを生成します。テキストによる説明を、世界認識に基づく動き、照明、音声を備えたビデオ出力に変換します。ソーシャルメディアコンテンツ作成、迅速なビデオプロトタイピング、反復的なビジュアルストーリーテリングに最適です。
-
-### 画像から動画へ
-
-
-
- Comfy Cloudで開く
-
-
- JSONをダウンロードするか、テンプレートライブラリで「Gemini Omni Flash」を検索
-
-
- このワークフローで使用する例の入力画像を取得
-
-
- 2つ目のサンプル入力画像を取得
-
-
-
-
-
-Gemini Omni Flashを使用して2枚の画像からビデオを生成します。自然言語のプロンプトを解釈して、再生時間とアスペクト比を制御します。短いブランドクリップ、ダイナミックなソーシャルメディアコンテンツ、会話型プロンプトによる反復的なビデオ編集に最適です。
-
-### ビデオ編集
+
+プレビュー版エンドポイント(`gemini-omni-flash-preview`)は2026年9月30日以降に動作しなくなる予定です。それまでにワークフローをOmni Flash 1.1モデルオプションへ切り替えてください。
+
-
-
- Comfy Cloudで開く
-
-
- JSONをダウンロードするか、テンプレートライブラリで「Gemini Omni Flash」を検索
-
-
- このワークフローで使用する例の入力ビデオを取得
-
-
+## 料金
-
+合計クレジット = **(秒あたりクレジット)× `duration`**。Gemini Omni Flash 1.1は出力解像度によって課金されます:
-Gemini Omni Flashを使用して、自然言語でビデオを編集します。単一の入力ビデオを、説明文に基づいて1つの編集済み出力に変換します。プロンプトで再生時間とアスペクト比を指定します。ソーシャルメディアでの素早いリミックス、映画的なシーン調整、反復的なビデオの洗練に最適です。
+| 解像度 | 秒あたりクレジット |
+| :--------- | :------------ |
+| 360p | 10.19 |
+| 720p | 30.57 |
+| 1080p | 45.87 |
+| 4K | 91.76 |
-## はじめる
+プレビューモデルは秒あたり30.57クレジットの均一料金です。完全な料金表は[パートナーノードの料金ページ](/ja/tutorials/partner-nodes/pricing)を参照してください。
-1. ComfyUIを最新バージョンにアップデートする
-2. キャンバスをダブルクリックし、「Gemini Omni Flash」ノードを検索する
-3. またはテンプレートライブラリから既製のワークフローを使用する
-4. 入力タイプ(テキスト、画像、ビデオ)に合ったワークフローを選択する
-5. プロンプトを入力して生成する
+## ComfyUIで使う
-
-最良の結果を得るには、Gemini Omni FlashをNano Banana 2 Liteと組み合わせて使用してください。高速度で画像を生成し、その後Gemini Omni Flashでアニメーション化してビデオにします。
-
+
+テキストからビデオ、画像からビデオ、参照からビデオ、ビデオ編集、ビデオ拡張のワークフローを、ComfyUIローカルまたはComfy Cloudで実行します
+
diff --git a/ja/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx b/ja/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
new file mode 100644
index 000000000..1e789d6e7
--- /dev/null
+++ b/ja/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
@@ -0,0 +1,135 @@
+---
+title: "ComfyUI で Gemini Omni Flash ワークフローを実行"
+description: "Gemini Omni Flashのテキストからビデオ、画像からビデオ、参照からビデオ、ビデオ編集、ビデオ拡張ワークフローを、ComfyUIローカルまたはComfy Cloudで実行する方法"
+sidebarTitle: "ワークフロー"
+translationSourceHash: 14cd9be3
+translationFrom: tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
+translationBlockHashes:
+ "_intro": 5ee3af24
+ "Workflows": 36bbfae6
+ "Get started": 02809da9
+---
+
+import ReqHint from "/snippets/ja/tutorials/partner-nodes/req-hint.mdx";
+import UpdateReminder from "/snippets/ja/tutorials/update-reminder.mdx";
+
+このガイドでは、ComfyUIでGemini Omni Flashワークフローを実行する方法を説明します。モデルの概要と機能については、[Gemini Omni Flashの概要](/ja/tutorials/partner-nodes/google/gemini-omni-flash)を参照してください。
+
+
+
+
+
+Gemini Omni Flash 1.1ワークフローにはComfyUI 0.34.2以降が必要です。ノードのモデルドロップダウンで**Omni Flash 1.1**を選択するとGA版モデルを使用できます。**Omni Flash**オプションはプレビューモデルを実行しますが、2026年9月30日に提供終了予定です。
+
+
+## ワークフロー
+
+### テキストから動画へ(Omni Flash 1.1)
+
+
+
+ Comfy Cloudで開く
+
+
+ JSONをダウンロードするか、テンプレートライブラリで「Gemini Omni 1.1」を検索
+
+
+
+
+
+
+
+Gemini Omni Flash 1.1で自然言語プロンプトから映画的なビデオを生成します。プロンプトに直接希望の長さ(3〜10秒)を記述し、ノードでアスペクト比と出力解像度を選択します。16:9または9:16、下書きは360p、最終レンダリングは最大4K。すべてのクリップに生成されたオーディオトラックが含まれます。
+
+### 画像から動画へ(Omni Flash 1.1)
+
+
+
+ Comfy Cloudで開く
+
+
+ JSONをダウンロードするか、テンプレートライブラリで「Gemini Omni 1.1」を検索
+
+
+ このワークフローで使用する例の入力画像を取得
+
+
+
+
+
+
+
+Gemini Omni Flash 1.1で画像をアニメーション化します。`image_to_video`タスクでは、1枚目の画像が開始フレーム、オプションの2枚目が終了フレームになります。モデルがその間の映像を生成するため、カメラオービット、ズームトランジション、ループクリップが予測可能になります。
+
+### 参照からビデオ生成(Omni Flash 1.1)
+
+
+
+ Comfy Cloudで開く
+
+
+ JSONをダウンロードするか、テンプレートライブラリで「Gemini Omni 1.1」を検索
+
+
+ 1つ目のサンプル参照画像を取得
+
+
+ 2つ目のサンプル参照画像を取得
+
+
+
+
+
+
+
+最大14枚の参照画像から特定の被写体を取り込んだビデオを生成します。参照モードでは、``のようなタグで各画像を役割に紐付け、プロンプトでそのタグを参照します。画像の中のキャラクター、製品、オブジェクトがシーンに登場し、画像自体がフレームとして使われることはありません。キャラクター参照とスタイル参照を組み合わせると、ブランドに一貫したコンテンツを作成できます。
+
+### ビデオ編集(Omni Flash 1.1)
+
+
+
+ Comfy Cloudで開く
+
+
+ JSONをダウンロードするか、テンプレートライブラリで「Gemini Omni 1.1」を検索
+
+
+ このワークフローで使用する例の入力ビデオを取得
+
+
+
+
+
+
+
+Gemini Omni Flash 1.1で自然言語によるビデオ編集を行います。`edit`タスクでは、ノードは入力ビデオを1本だけ受け取り(10秒以内)、指示に基づいて書き換えます。背景の入れ替え、シーンのスタイル変更、要素の追加・削除が可能です。`edit`と`extend`タスクは入力ビデオのアスペクト比を保持します。シンプルなプロンプトが最も効果的で、「他はそのままに」を追加すると一貫性が最大化されます。
+
+### ビデオ拡張(Omni Flash 1.1)
+
+
+
+ Comfy Cloudで開く
+
+
+ JSONをダウンロードするか、テンプレートライブラリで「Gemini Omni 1.1」を検索
+
+
+ このワークフローで使用する例の入力ビデオを取得
+
+
+
+
+
+`extend`タスクで既存のビデオを1ステップあたり最大10秒ずつ拡張し、累計約40秒のストーリーを構築できます。モデルは直近10秒のコンテキストを分析し、シーンが中断した箇所から続く際にキャラクター、動き、オーディオの一貫性を保ちます。参照画像を添付すると、物語の途中に新しいキャラクターを登場させることもできます。拡張は新しいコンテンツをクリップの末尾にのみ追加しますが、シームレスな切り替えのために最後のソースフレームが修正される場合があります。
+
+## はじめる
+
+1. ComfyUIを最新バージョンにアップデートする(Omni Flash 1.1ワークフローには0.34.2以降が必要)
+2. キャンバスをダブルクリックし、「Gemini Omni Flash」ノードを検索する
+3. またはテンプレートライブラリから既製のワークフローを使用する
+4. 入力タイプ(テキスト、画像、ビデオ)に合ったワークフローを選択する
+5. プロンプトを入力して生成する
+
+
+最良の結果を得るには、Gemini Omni FlashをNano Banana 2 Liteと組み合わせて使用してください。高速度で画像を生成し、その後Gemini Omni Flashでアニメーション化してビデオにします。
+
diff --git a/ko/tutorials/partner-nodes/google/gemini-omni-flash.mdx b/ko/tutorials/partner-nodes/google/gemini-omni-flash.mdx
index bffd11e6a..41ca9aa12 100644
--- a/ko/tutorials/partner-nodes/google/gemini-omni-flash.mdx
+++ b/ko/tutorials/partner-nodes/google/gemini-omni-flash.mdx
@@ -1,97 +1,68 @@
---
title: "Gemini Omni Flash: 대화형 비디오 생성"
-description: "Gemini Omni Flash를 사용하여 자연어로 비디오를 생성하고 편집하세요. Google의 멀티모달 비디오 모델로, ComfyUI에서 파트너 노드를 통해 사용 가능합니다."
+description: "Gemini Omni Flash 1.1을 사용하여 자연어로 비디오를 생성하고 편집하세요. Google의 멀티모달 비디오 모델로, ComfyUI에서 파트너 노드를 통해 사용 가능합니다."
sidebarTitle: "Gemini Omni Flash"
-translationSourceHash: 9a3c952e
+translationSourceHash: 277b18df
translationFrom: tutorials/partner-nodes/google/gemini-omni-flash.mdx
translationBlockHashes:
- "_intro": 3b6973dc
- "What Gemini Omni Flash offers": d0140bc2
- "Workflows": 753af4cb
- "Get started": 64517938
+ "_intro": 5b9c88f4
+ "What Gemini Omni Flash 1.1 is good at": 92599abb
+ "Omni Flash 1.1 vs Omni Flash (preview)": ba3e2a73
+ "Pricing": da96966c
+ "Use it in ComfyUI": dae97960
---
import ReqHint from "/snippets/ko/tutorials/partner-nodes/req-hint.mdx";
import UpdateReminder from "/snippets/ko/tutorials/update-reminder.mdx";
-Gemini Omni Flash는 Google DeepMind의 고품질, 비용 효율적인 비디오 생성 및 대화형 편집 모델입니다. Google I/O 2026에서 Gemini Omni 제품군의 일부로 처음 소개되었으며, Gemini의 멀티모달 추론과 네이티브 비디오 생성을 결합하여 개발자가 자연어 대화를 통해 비디오를 생성, 편집 및 리믹스할 수 있도록 합니다.
+Gemini Omni Flash는 Google DeepMind의 대화형 비디오 생성 및 편집 모델로, Google I/O 2026에서 처음 공개된 Gemini Omni 제품군의 일부입니다. Gemini의 멀티모달 추론과 네이티브 비디오 생성을 결합하여, 자연어로 원하는 내용을 설명하고 참조 이미지나 비디오를 첨부하면 동기화된 오디오가 포함된 클립을 생성합니다. 현재 버전인 Gemini Omni Flash 1.1은 2026년 8월 27일에 정식 출시(GA)되었으며, 키프레임 보간, 360p/4K 출력 옵션, 장면 연장 기능이 추가되었습니다.
-## Gemini Omni Flash가 제공하는 기능
+## Gemini Omni Flash 1.1이 뛰어난 분야
-- **대화형 비디오 편집**: 자연어를 사용하여 비디오를 다듬고 편집하세요. 캐릭터 교체, 장면 조명 변경, 각도 변경, 객체 추가 또는 제거를 수행하면서 원본 오디오 및 비디오 트랙을 유지합니다.
-- **멀티모달 입력**: 텍스트, 이미지 및 비디오 입력을 결합하여 생성을 안내합니다. 모든 비디오 출력에 동기화된 오디오를 기본적으로 생성합니다.
+- **대화형 비디오 편집**: 자연어를 사용하여 비디오를 다듬고 편집합니다. 캐릭터 교체, 장면 조명 변경, 각도 변경, 객체 추가 또는 제거를 수행하면서 원본 오디오 및 비디오 트랙을 유지합니다.
+- **멀티모달 입력**: 텍스트, 이미지(최대 14개), 비디오(최대 3개, 각 10초)를 결합하여 생성을 안내합니다. 모든 출력에 네이티브 오디오 트랙이 포함됩니다.
+- **키프레임 보간**: `image_to_video` 작업에서 시작 프레임과 선택적인 종료 프레임을 첨부하면, 그 사이의 영상을 생성합니다.
+- **참조 기반 비디오 생성**: `` 같은 태그로 참조 이미지를 역할에 바인딩하여, 이미지 속 캐릭터, 제품, 사물을 새 장면에 등장시킵니다. 이미지 자체는 프레임으로 사용되지 않습니다.
+- **장면 연장**: `extend` 작업은 클립에 최대 10초의 새 영상을 추가합니다. 마지막 10초의 컨텍스트를 분석하여 캐릭터와 동작의 일관성을 유지하며, 누적 약 40초까지 연장할 수 있습니다.
+- **해상도 제어**: 360p로 초안을 작성한 후(비용은 720p의 약 3분의 1) 720p, 1080p, 4K로 다시 렌더링합니다. 16:9 및 9:16 화면 비율을 지원합니다.
- **세계 지식 및 시뮬레이션**: 물리 이해와 Gemini의 역사, 과학 및 문화적 맥락에 대한 지식을 결합하여 사실적 표현을 넘어 의미 있는 스토리텔링을 가능하게 합니다.
- **텍스트 및 동작 동기화**: 읽기 쉬운 텍스트와 그래픽을 비디오에 직접 렌더링하여 동적 타이포그래피를 화면 움직임과 동기화합니다.
-- **가격**: 비디오 출력 초당 $0.10이며, Veo 3.1 Fast 가격과 동일합니다.
-## 워크플로
+## Omni Flash 1.1과 Omni Flash(프리뷰) 비교
-### 텍스트 기반 비디오 생성
+ComfyUI 노드는 모델 드롭다운에서 두 버전을 모두 제공합니다:
-
-
- Comfy Cloud에서 열기
-
-
- JSON 다운로드 또는 템플릿 라이브러리에서 "Gemini Omni Flash" 검색
-
-
+| | Gemini Omni Flash 1.1 | Gemini Omni Flash(프리뷰) |
+|---|---|---|
+| 모델 ID | `gemini-omni-1.1-flash` | `gemini-omni-flash-preview` |
+| 상태 | 정식 출시(2026년 8월 27일) | 공개 프리뷰, 2026년 9월 30일 종료 예정 |
+| 해상도 | 360p, 720p, 1080p, 4K | 고정 720p, 24 FPS |
+| 작업 제어 | 명시적인 `task_type` 선택기: auto, text_to_video, image_to_video, reference_to_video, edit, extend | 첨부된 미디어에서 추론 |
+| 가격 | 초당 10.19~91.76 크레딧(해상도별) | 초당 30.57 크레딧 |
-
-
-자연어 프롬프트로 시네마틱 비디오를 생성합니다. 텍스트 설명을 세계 인식 모션, 조명 및 사운드가 포함된 비디오 출력으로 변환합니다. 소셜 미디어 콘텐츠 생성, 빠른 비디오 프로토타이핑 및 반복적인 시각적 스토리텔링에 이상적입니다.
-
-### 이미지 기반 비디오 생성
-
-
-
- Comfy Cloud에서 열기
-
-
- JSON 다운로드 또는 템플릿 라이브러리에서 "Gemini Omni Flash" 검색
-
-
- 이 워크플로의 예제 입력 이미지 가져오기
-
-
- 두 번째 예제 입력 이미지 가져오기
-
-
-
-
-
-Gemini Omni Flash를 사용하여 두 이미지로 비디오를 생성합니다. 자연어 프롬프트를 해석하여 재생 시간과 화면 비율을 제어합니다. 짧은 브랜드 클립, 다이나믹한 소셜 미디어 콘텐츠 제작 및 대화형 프롬프트를 통한 반복적인 비디오 편집에 적합합니다.
-
-### 비디오 편집
+
+프리뷰 엔드포인트(`gemini-omni-flash-preview`)는 2026년 9월 30일 이후에 작동이 중단될 예정입니다. 그 전에 워크플로를 Omni Flash 1.1 모델 옵션으로 전환하세요.
+
-
-
- Comfy Cloud에서 열기
-
-
- JSON 다운로드 또는 템플릿 라이브러리에서 "Gemini Omni Flash" 검색
-
-
- 이 워크플로의 예제 입력 비디오 가져오기
-
-
+## 가격
-
+총 크레딧 = **(초당 크레딧) × `duration`**. Gemini Omni Flash 1.1은 출력 해상도에 따라 과금됩니다:
-Gemini Omni Flash를 사용하여 자연어로 비디오를 편집합니다. 하나의 입력 비디오를 설명 지침에 따라 하나의 편집된 출력으로 변환합니다. 프롬프트에서 재생 시간과 화면 비율을 지정합니다. 빠른 소셜 미디어 리믹스, 시네마틱 장면 조정 및 반복적인 비디오 다듬기에 이상적입니다.
+| 해상도 | 초당 크레딧 |
+| :--------- | :------------ |
+| 360p | 10.19 |
+| 720p | 30.57 |
+| 1080p | 45.87 |
+| 4K | 91.76 |
-## 시작하기
+프리뷰 모델은 초당 30.57 크레딧의 단일 요금이 적용됩니다. 전체 요금표는 [파트너 노드 가격 페이지](/ko/tutorials/partner-nodes/pricing)를 참조하세요.
-1. ComfyUI를 최신 버전으로 업데이트하세요.
-2. 캔버스를 더블 클릭하고 "Gemini Omni Flash" 노드를 검색하세요.
-3. 또는 템플릿 라이브러리에서 바로 사용할 수 있는 워크플로를 사용하세요.
-4. 입력 유형(텍스트, 이미지 또는 비디오)에 맞는 워크플로를 선택하세요.
-5. 프롬프트를 입력하고 생성하세요.
+## ComfyUI에서 사용하기
-
-최상의 결과를 위해 Gemini Omni Flash를 Nano Banana 2 Lite와 결합하세요: 고속으로 이미지를 생성한 다음, Gemini Omni Flash를 사용하여 비디오로 애니메이션화하세요.
-
+
+텍스트 기반 비디오 생성, 이미지 기반 비디오 생성, 참조 기반 비디오 생성, 비디오 편집, 비디오 연장 워크플로를 ComfyUI 로컬 또는 Comfy Cloud에서 실행합니다
+
diff --git a/ko/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx b/ko/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
new file mode 100644
index 000000000..a868c45e6
--- /dev/null
+++ b/ko/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
@@ -0,0 +1,135 @@
+---
+title: "ComfyUI에서 Gemini Omni Flash 워크플로 실행"
+description: "Gemini Omni Flash의 텍스트 기반 비디오 생성, 이미지 기반 비디오 생성, 참조 기반 비디오 생성, 비디오 편집, 비디오 연장 워크플로를 ComfyUI 로컬 또는 Comfy Cloud에서 실행하는 방법"
+sidebarTitle: "워크플로"
+translationSourceHash: 14cd9be3
+translationFrom: tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
+translationBlockHashes:
+ "_intro": 5ee3af24
+ "Workflows": 36bbfae6
+ "Get started": 02809da9
+---
+
+import ReqHint from "/snippets/ko/tutorials/partner-nodes/req-hint.mdx";
+import UpdateReminder from "/snippets/ko/tutorials/update-reminder.mdx";
+
+이 가이드는 ComfyUI에서 Gemini Omni Flash 워크플로를 실행하는 방법을 설명합니다. 모델의 개요와 기능은 [Gemini Omni Flash 개요](/ko/tutorials/partner-nodes/google/gemini-omni-flash)를 참조하세요.
+
+
+
+
+
+Gemini Omni Flash 1.1 워크플로에는 ComfyUI 0.34.2 이상이 필요합니다. 노드의 모델 드롭다운에서 **Omni Flash 1.1**을 선택하면 정식 출시 모델을 사용할 수 있습니다. **Omni Flash** 옵션은 프리뷰 모델을 실행하며, 2026년 9월 30일에 종료될 예정입니다.
+
+
+## 워크플로
+
+### 텍스트 기반 비디오 생성(Omni Flash 1.1)
+
+
+
+ Comfy Cloud에서 열기
+
+
+ JSON 다운로드 또는 템플릿 라이브러리에서 "Gemini Omni 1.1" 검색
+
+
+
+
+
+
+
+Gemini Omni Flash 1.1로 자연어 프롬프트에서 시네마틱 비디오를 생성합니다. 프롬프트에 원하는 길이(3~10초)를 직접 지정하고, 노드에서 화면 비율과 출력 해상도를 선택합니다. 16:9 또는 9:16, 초안은 360p, 최종 렌더링은 최대 4K까지 가능합니다. 모든 클립에는 생성된 오디오 트랙이 포함됩니다.
+
+### 이미지 기반 비디오 생성(Omni Flash 1.1)
+
+
+
+ Comfy Cloud에서 열기
+
+
+ JSON 다운로드 또는 템플릿 라이브러리에서 "Gemini Omni 1.1" 검색
+
+
+ 이 워크플로의 예제 입력 이미지 가져오기
+
+
+
+
+
+
+
+Gemini Omni Flash 1.1로 이미지를 애니메이션화합니다. `image_to_video` 작업에서는 첫 번째 첨부 이미지가 시작 프레임이 되고, 선택적인 두 번째 이미지가 종료 프레임이 됩니다. 모델이 그 사이의 영상을 생성하므로 카메라 궤도 회전, 줌 전환, 루프 클립을 예측 가능하게 만들 수 있습니다.
+
+### 참조 기반 비디오 생성(Omni Flash 1.1)
+
+
+
+ Comfy Cloud에서 열기
+
+
+ JSON 다운로드 또는 템플릿 라이브러리에서 "Gemini Omni 1.1" 검색
+
+
+ 첫 번째 예제 참조 이미지 가져오기
+
+
+ 두 번째 예제 참조 이미지 가져오기
+
+
+
+
+
+
+
+최대 14개의 참조 이미지에서 특정 피사체를 반영한 비디오를 생성합니다. 참조 모드에서는 `` 같은 태그로 각 이미지를 역할에 바인딩하고 프롬프트에서 해당 태그를 참조합니다. 이미지 속 캐릭터, 제품, 사물이 장면에 등장하며, 이미지 자체는 프레임으로 사용되지 않습니다. 캐릭터 참조와 스타일 참조를 결합하면 브랜드에 일관된 콘텐츠를 만들 수 있습니다.
+
+### 비디오 편집(Omni Flash 1.1)
+
+
+
+ Comfy Cloud에서 열기
+
+
+ JSON 다운로드 또는 템플릿 라이브러리에서 "Gemini Omni 1.1" 검색
+
+
+ 이 워크플로의 예제 입력 비디오 가져오기
+
+
+
+
+
+
+
+Gemini Omni Flash 1.1로 자연어 기반 비디오 편집을 수행합니다. `edit` 작업에서는 노드가 입력 비디오를 정확히 하나만 받고(10초 이하) 지시에 따라 다시 만듭니다. 배경 교체, 장면 스타일 변경, 요소 추가 및 제거가 가능합니다. `edit` 및 `extend` 작업은 입력 비디오의 화면 비율을 유지합니다. 간결한 프롬프트가 가장 효과적이며, "나머지는 그대로 유지"를 추가하면 일관성이 극대화됩니다.
+
+### 비디오 연장(Omni Flash 1.1)
+
+
+
+ Comfy Cloud에서 열기
+
+
+ JSON 다운로드 또는 템플릿 라이브러리에서 "Gemini Omni 1.1" 검색
+
+
+ 이 워크플로의 예제 입력 비디오 가져오기
+
+
+
+
+
+`extend` 작업으로 기존 비디오를 단계당 최대 10초씩 연장하여, 누적 약 40초의 스토리를 구축할 수 있습니다. 모델은 마지막 10초의 컨텍스트를 분석하여, 장면이 중단된 지점부터 이어질 때 캐릭터, 동작, 오디오의 일관성을 유지합니다. 참조 이미지를 첨부하면 스토리 중간에 새 캐릭터를 등장시킬 수도 있습니다. 연장은 새 콘텐츠를 클립 끝에만 추가하지만, 원활한 전환을 위해 마지막 소스 프레임이 수정될 수 있습니다.
+
+## 시작하기
+
+1. ComfyUI를 최신 버전으로 업데이트하세요(Omni Flash 1.1 워크플로에는 0.34.2 이상 필요)
+2. 캔버스를 더블 클릭하고 "Gemini Omni Flash" 노드를 검색하세요.
+3. 또는 템플릿 라이브러리에서 바로 사용할 수 있는 워크플로를 사용하세요.
+4. 입력 유형(텍스트, 이미지 또는 비디오)에 맞는 워크플로를 선택하세요.
+5. 프롬프트를 입력하고 생성하세요.
+
+
+최상의 결과를 위해 Gemini Omni Flash를 Nano Banana 2 Lite와 결합하세요: 고속으로 이미지를 생성한 다음, Gemini Omni Flash를 사용하여 비디오로 애니메이션화하세요.
+
diff --git a/tutorials/partner-nodes/google/gemini-omni-flash.mdx b/tutorials/partner-nodes/google/gemini-omni-flash.mdx
index 21f707e1e..4ba73c8f9 100644
--- a/tutorials/partner-nodes/google/gemini-omni-flash.mdx
+++ b/tutorials/partner-nodes/google/gemini-omni-flash.mdx
@@ -1,21 +1,53 @@
---
title: "Gemini Omni Flash: Conversational Video Generation"
-description: "Generate and edit videos through natural language using Gemini Omni Flash, Google's multimodal video model, available in ComfyUI through Partner Nodes"
+description: "Generate and edit videos through natural language using Gemini Omni Flash 1.1, Google's multimodal video model, available in ComfyUI through Partner Nodes"
sidebarTitle: "Overview"
---
-Gemini Omni Flash is Google DeepMind's high-quality, cost-efficient video generation and conversational editing model. First introduced at Google I/O 2026 as part of the Gemini Omni family, it combines Gemini's multimodal reasoning with native video creation, enabling developers to generate, edit, and remix videos through natural conversation.
+Gemini Omni Flash is Google DeepMind's conversational video generation and editing model, part of the Gemini Omni family first introduced at Google I/O 2026. It combines Gemini's multimodal reasoning with native video creation: you describe what you want in plain language, attach reference images or videos, and the model generates a clip with synchronized audio. The current version, Gemini Omni Flash 1.1, reached general availability on August 27, 2026 and adds keyframe interpolation, 360p/4K output options, and scene extension.
-## What Gemini Omni Flash is good at
+## What Gemini Omni Flash 1.1 is good at
-- **Conversational video editing**: Refine and edit videos using natural language: swap characters, relight scenes, alter angles, add or remove objects while maintaining original audio and video tracks
-- **Multimodal input**: Combine text, images, and video inputs to guide generation. Natively generates synchronized audio with every video output
+- **Conversational video editing**: Refine and edit videos using natural language: swap characters, relight scenes, alter angles, add or remove objects while maintaining the original audio and video tracks
+- **Multimodal input**: Combine text, images (up to 14), and videos (up to 3, 10 seconds each) to guide generation. Every output carries a native audio track
+- **Keyframe interpolation**: With the `image_to_video` task, attach a starting frame and an optional ending frame, and the model generates the footage in between
+- **Reference to video**: Bind reference images to roles with tags like `` so characters, products, and objects from your images appear in a new scene without the image itself being used as a frame
+- **Scene extension**: The `extend` task appends up to 10 seconds of new footage to a clip, reading the last 10 seconds of context for character and motion consistency, up to about 40 seconds cumulatively
+- **Resolution control**: Draft at 360p (about one third of the 720p cost), then re-render at 720p, 1080p, or 4K. 16:9 and 9:16 aspect ratios are supported
- **World knowledge and simulation**: Combines physics understanding with Gemini's knowledge of history, science, and cultural context, enabling meaningful storytelling beyond photorealism
- **Text and action synchronization**: Render legible text and graphics directly into video, syncing kinetic typography with on-screen movements
-- **Pricing**: $0.10 per second of video output, matching Veo 3.1 Fast pricing
+
+## Omni Flash 1.1 vs Omni Flash (preview)
+
+The ComfyUI node offers both versions in the model dropdown:
+
+| | Gemini Omni Flash 1.1 | Gemini Omni Flash (preview) |
+|---|---|---|
+| Model ID | `gemini-omni-1.1-flash` | `gemini-omni-flash-preview` |
+| Status | Generally available (August 27, 2026) | Public preview, retiring September 30, 2026 |
+| Resolution | 360p, 720p, 1080p, 4K | Fixed 720p, 24 FPS |
+| Task control | Explicit `task_type` selector: auto, text_to_video, image_to_video, reference_to_video, edit, extend | Inferred from attached media |
+| Pricing | 10.19 to 91.76 credits per second by resolution | 30.57 credits per second |
+
+
+The preview endpoint (`gemini-omni-flash-preview`) is scheduled to stop working after September 30, 2026. Switch workflows to the Omni Flash 1.1 model option before then.
+
+
+## Pricing
+
+Total credits = **(credits / sec) × `duration`**. Gemini Omni Flash 1.1 bills by output resolution:
+
+| Resolution | Credits / sec |
+| :--------- | :------------ |
+| 360p | 10.19 |
+| 720p | 30.57 |
+| 1080p | 45.87 |
+| 4K | 91.76 |
+
+The preview model bills a flat 30.57 credits per second. See the [partner node pricing page](/tutorials/partner-nodes/pricing) for the full rate table.
## Use it in ComfyUI
-Run the text-to-video, image-to-video, and video edit workflows in ComfyUI, locally or on Comfy Cloud
+Run the text-to-video, image-to-video, reference-to-video, video edit, and video extend workflows in ComfyUI, locally or on Comfy Cloud
diff --git a/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx b/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
index 649e0a51b..769d20d89 100644
--- a/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
+++ b/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
@@ -1,6 +1,6 @@
---
title: "Gemini Omni Flash Workflows in ComfyUI"
-description: "How to run the Gemini Omni Flash text-to-video, image-to-video, and video edit workflows in ComfyUI, locally or on Comfy Cloud"
+description: "How to run the Gemini Omni Flash text-to-video, image-to-video, reference-to-video, video edit, and video extend workflows in ComfyUI, locally or on Comfy Cloud"
sidebarTitle: "Workflow"
---
@@ -12,65 +12,113 @@ This guide shows how to run the Gemini Omni Flash workflows in ComfyUI. For what
+
+The Gemini Omni Flash 1.1 workflows require ComfyUI 0.34.2 or later. In the node, select **Omni Flash 1.1** from the model dropdown to use the GA model; the **Omni Flash** option runs the preview model, which is scheduled to retire on September 30, 2026.
+
+
## Workflows
-### Text to Video
+### Text to Video (Omni Flash 1.1)
-
+
Open in Comfy Cloud
-
- Download JSON or search "Gemini Omni Flash" in Template Library
+
+ Download JSON or search "Gemini Omni 1.1" in Template Library
-
+
+
+
-Generate cinematic video from natural language prompts. Transform text descriptions into video output with world-aware motion, lighting, and sound. Ideal for social media content creation, rapid video prototyping, and iterative visual storytelling.
+Generate cinematic video from natural language prompts with Gemini Omni Flash 1.1. Describe the desired length (3 to 10 seconds) directly in the prompt, and pick the aspect ratio and output resolution in the node: 16:9 or 9:16, 360p for cheap drafts, up to 4K for final renders. Every clip includes a generated audio track.
-### Image to Video
+### Image to Video (Omni Flash 1.1)
-
+
Open in Comfy Cloud
-
- Download JSON or search "Gemini Omni Flash" in Template Library
+
+ Download JSON or search "Gemini Omni 1.1" in Template Library
-
+
Get the example input image for this workflow
-
- Get the second example input image
+
+
+
+
+
+
+Animate an image with Gemini Omni Flash 1.1. With the `image_to_video` task, the first attached image becomes the starting frame and an optional second image becomes the ending frame: the model generates the footage in between, which makes camera orbits, zoom transitions, and looping clips predictable.
+
+### Reference to Video (Omni Flash 1.1)
+
+
+
+ Open in Comfy Cloud
+
+
+ Download JSON or search "Gemini Omni 1.1" in Template Library
+
+
+ Get the first example reference image
+
+
+ Get the second example reference image
+
+
+
+
+
+
+
+Generate video that incorporates specific subjects from up to 14 reference images. In reference mode, bind each image to a role with tags like `` and refer to the tags in your prompt: characters, products, and objects from your images appear in the scene while the image itself is never used as a frame. Combine character references with style references for brand-consistent content.
+
+### Video Edit (Omni Flash 1.1)
+
+
+
+ Open in Comfy Cloud
+
+
+ Download JSON or search "Gemini Omni 1.1" in Template Library
+
+
+ Get the example input video for this workflow
-
+
+
+
-Generate a video from two images using Gemini Omni Flash. Interpret natural language prompts to control duration and aspect ratio. Perfect for creating short brand clips, dynamic social media content, and iterative video edits through conversational prompting.
+Edit videos with natural language using Gemini Omni Flash 1.1. With the `edit` task, the node takes exactly one input video (10 seconds or less) and rewrites it based on your instructions: swap backgrounds, restyle scenes, add or remove elements. The `edit` and `extend` tasks keep the aspect ratio of the input video. Simple prompts work best; adding "keep everything else the same" maximizes consistency.
-### Video Edit
+### Video Extend (Omni Flash 1.1)
-
+
Open in Comfy Cloud
-
- Download JSON or search "Gemini Omni Flash" in Template Library
+
+ Download JSON or search "Gemini Omni 1.1" in Template Library
-
+
Get the example input video for this workflow
-
+
-Edit videos with natural language using Gemini Omni Flash. Transform a single input video into one edited output based on your descriptive instructions. Specify the duration and aspect ratio in your prompt. Ideal for quick social media remixes, cinematic scene adjustments, and iterative video refinements.
+Extend an existing video by up to 10 seconds per step with the `extend` task, building stories up to about 40 seconds total. The model analyzes the last 10 seconds of context, keeping characters, motion, and audio coherent as the scene continues from where it left off. Optionally attach reference images to introduce new characters mid-story. Extension appends new content only at the end of the clip, but the model may revise the final source frames to make the transition seamless.
## Get started
-1. Update ComfyUI to the latest version
+1. Update ComfyUI to the latest version (0.34.2 or later for the Omni Flash 1.1 workflows)
2. Double-click the canvas and search for "Gemini Omni Flash" nodes
3. Or go to the Template Library to use the ready-to-go workflows
4. Choose the workflow that matches your input type (text, image, or video)
diff --git a/zh/tutorials/partner-nodes/google/gemini-omni-flash.mdx b/zh/tutorials/partner-nodes/google/gemini-omni-flash.mdx
index 421832703..6b00d8401 100644
--- a/zh/tutorials/partner-nodes/google/gemini-omni-flash.mdx
+++ b/zh/tutorials/partner-nodes/google/gemini-omni-flash.mdx
@@ -1,97 +1,68 @@
---
title: "Gemini Omni Flash:对话式视频生成"
-description: "通过合作节点在 ComfyUI 中使用 Google 的多模态视频模型 Gemini Omni Flash,以自然语言生成和编辑视频"
+description: "通过合作节点在 ComfyUI 中使用 Google 的多模态视频模型 Gemini Omni Flash 1.1,以自然语言生成和编辑视频"
sidebarTitle: "Gemini Omni Flash"
-translationSourceHash: 9a3c952e
+translationSourceHash: 277b18df
translationFrom: tutorials/partner-nodes/google/gemini-omni-flash.mdx
translationBlockHashes:
- "_intro": 3b6973dc
- "What Gemini Omni Flash offers": d0140bc2
- "Workflows": 753af4cb
- "Get started": 64517938
+ "_intro": 5b9c88f4
+ "What Gemini Omni Flash 1.1 is good at": 92599abb
+ "Omni Flash 1.1 vs Omni Flash (preview)": ba3e2a73
+ "Pricing": da96966c
+ "Use it in ComfyUI": dae97960
---
import ReqHint from "/snippets/zh/tutorials/partner-nodes/req-hint.mdx";
import UpdateReminder from "/snippets/zh/tutorials/update-reminder.mdx";
-Gemini Omni Flash 是 Google DeepMind 推出的高质量、经济高效的视频生成与对话式编辑模型。该模型于 Google I/O 2026 作为 Gemini Omni 家族成员首次亮相,将 Gemini 的多模态推理能力与原生的视频创建功能结合,使开发者能够通过自然对话生成、编辑和重新混合视频。
+Gemini Omni Flash 是 Google DeepMind 的对话式视频生成与编辑模型,属于 Google I/O 2026 上首次亮相的 Gemini Omni 家族。它将 Gemini 的多模态推理能力与原生视频创建相结合:用自然语言描述你想要的内容,附上参考图像或视频,模型即可生成带同步音频的片段。当前版本 Gemini Omni Flash 1.1 已于 2026 年 8 月 27 日正式发布(GA),新增了关键帧插值、360p/4K 输出选项和场景延续功能。
-## Gemini Omni Flash 提供的功能
+## Gemini Omni Flash 1.1 的优势
- **对话式视频编辑**:使用自然语言精炼和编辑视频:替换角色、重新布光、改变角度、添加或移除物体,同时保留原始音视频轨道
-- **多模态输入**:结合文本、图像和视频输入来引导生成。每次输出视频时原生生成同步音频
+- **多模态输入**:结合文本、图像(最多 14 张)和视频(最多 3 个,每个 10 秒)来引导生成。每次输出都带有原生音频轨道
+- **关键帧插值**:使用 `image_to_video` 任务时,附上起始帧和可选的结束帧,模型会生成两者之间的画面
+- **参考生成视频**:使用 `` 等标签将参考图像绑定到角色,让图像中的角色、产品和物体出现在新场景中,而图像本身不会作为画面帧使用
+- **场景延续**:`extend` 任务可向片段追加最多 10 秒的新画面,模型会读取最后 10 秒的上下文以保持角色和动作的一致性,累计可延续约 40 秒
+- **分辨率控制**:先用 360p 起草(成本约为 720p 的三分之一),再以 720p、1080p 或 4K 重新渲染。支持 16:9 和 9:16 两种画面比例
- **世界知识与模拟**:将物理理解与 Gemini 在历史、科学及文化背景方面的知识相结合,实现超越照片真实感的有意义叙事
- **文本与动作同步**:直接在视频中渲染清晰文本和图形,使动态排版与屏幕上的运动同步
-- **定价**:每秒钟视频输出 $0.10,与 Veo 3.1 Fast 定价一致
-## 工作流
+## Omni Flash 1.1 与 Omni Flash(预览版)对比
-### 文本转视频
+ComfyUI 节点在模型下拉菜单中同时提供两个版本:
-
-
- 在 Comfy Cloud 中打开
-
-
- 下载 JSON,或在模板库中搜索“Gemini Omni Flash”
-
-
+| | Gemini Omni Flash 1.1 | Gemini Omni Flash(预览版) |
+|---|---|---|
+| 模型 ID | `gemini-omni-1.1-flash` | `gemini-omni-flash-preview` |
+| 状态 | 已正式发布(2026 年 8 月 27 日) | 公开预览,2026 年 9 月 30 日退役 |
+| 分辨率 | 360p、720p、1080p、4K | 固定 720p、24 FPS |
+| 任务控制 | 显式的 `task_type` 选择器:auto、text_to_video、image_to_video、reference_to_video、edit、extend | 根据附加媒体自动推断 |
+| 定价 | 每秒 10.19 至 91.76 积分(按分辨率) | 每秒 30.57 积分 |
-
-
-根据自然语言提示生成电影级视频。将文本描述转换为具有世界感知的运动、光照和声音的视频输出。非常适合社交媒体内容创作、快速视频原型制作以及迭代式视觉叙事。
-
-### 图像转视频
-
-
-
- 在 Comfy Cloud 中打开
-
-
- 下载 JSON,或在模板库中搜索“Gemini Omni Flash”
-
-
- 获取此工作流的示例输入图像
-
-
- 获取第二张示例输入图像
-
-
-
-
-
-使用 Gemini Omni Flash 从两张图像生成视频。解释自然语言提示以控制时长和画面比例。非常适合制作简短品牌剪辑、动态社交媒体内容,以及通过对话式提示进行迭代视频编辑。
-
-### 视频编辑
+
+预览版端点(`gemini-omni-flash-preview`)计划于 2026 年 9 月 30 日之后停止服务。请在此之前将工作流切换到 Omni Flash 1.1 模型选项。
+
-
-
- 在 Comfy Cloud 中打开
-
-
- 下载 JSON,或在模板库中搜索“Gemini Omni Flash”
-
-
- 获取此工作流的示例输入视频
-
-
+## 定价
-
+总积分 = **(每秒积分)× `duration`**。Gemini Omni Flash 1.1 按输出分辨率计费:
-使用 Gemini Omni Flash 以自然语言编辑视频。根据描述性指令将单个输入视频转换为经过编辑的输出。在提示中指定时长和画面比例。非常适合快速社交媒体混剪、电影场景调整以及迭代视频精修。
+| 分辨率 | 每秒积分 |
+| :--------- | :------------ |
+| 360p | 10.19 |
+| 720p | 30.57 |
+| 1080p | 45.87 |
+| 4K | 91.76 |
-## 开始使用
+预览版模型按每秒 30.57 积分的统一费率计费。完整费率表请参阅[合作伙伴节点定价页面](/zh/tutorials/partner-nodes/pricing)。
-1. 将 ComfyUI 更新到最新版本
-2. 双击画布,搜索“Gemini Omni Flash”节点
-3. 或者进入模板库,使用现成的工作流
-4. 选择与输入类型(文本、图像或视频)匹配的工作流
-5. 输入提示并生成
+## 在 ComfyUI 中使用
-
-为获得最佳效果,可将 Gemini Omni Flash 与 Nano Banana 2 Lite 组合使用:先高速生成图像,再用 Gemini Omni Flash 将它们动画化为视频。
-
+
+在 ComfyUI 本地或 Comfy Cloud 上运行文本转视频、图像转视频、参考生成视频、视频编辑和视频延续工作流
+
diff --git a/zh/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx b/zh/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
new file mode 100644
index 000000000..a6bc38271
--- /dev/null
+++ b/zh/tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
@@ -0,0 +1,135 @@
+---
+title: "Gemini Omni Flash 工作流(ComfyUI)"
+description: "如何在 ComfyUI 本地或 Comfy Cloud 上运行 Gemini Omni Flash 的文本转视频、图像转视频、参考生成视频、视频编辑和视频延续工作流"
+sidebarTitle: "工作流"
+translationSourceHash: 14cd9be3
+translationFrom: tutorials/partner-nodes/google/gemini-omni-flash/workflow.mdx
+translationBlockHashes:
+ "_intro": 5ee3af24
+ "Workflows": 36bbfae6
+ "Get started": 02809da9
+---
+
+import ReqHint from "/snippets/zh/tutorials/partner-nodes/req-hint.mdx";
+import UpdateReminder from "/snippets/zh/tutorials/update-reminder.mdx";
+
+本指南介绍如何在 ComfyUI 中运行 Gemini Omni Flash 工作流。关于模型的功能与优势,请参阅 [Gemini Omni Flash 概览](/zh/tutorials/partner-nodes/google/gemini-omni-flash)。
+
+
+
+
+
+Gemini Omni Flash 1.1 工作流需要 ComfyUI 0.34.2 或更高版本。在节点中,从模型下拉菜单选择 **Omni Flash 1.1** 即可使用正式版模型;**Omni Flash** 选项运行的是预览版模型,计划于 2026 年 9 月 30 日退役。
+
+
+## 工作流
+
+### 文本转视频(Omni Flash 1.1)
+
+
+
+ 在 Comfy Cloud 中打开
+
+
+ 下载 JSON,或在模板库中搜索“Gemini Omni 1.1”
+
+
+
+
+
+
+
+使用 Gemini Omni Flash 1.1 从自然语言提示生成电影级视频。直接在提示中描述所需时长(3 到 10 秒),并在节点中选择画面比例和输出分辨率:16:9 或 9:16,起草用 360p,成片最高 4K。每个片段都包含生成的音频轨道。
+
+### 图像转视频(Omni Flash 1.1)
+
+
+
+ 在 Comfy Cloud 中打开
+
+
+ 下载 JSON,或在模板库中搜索“Gemini Omni 1.1”
+
+
+ 获取此工作流的示例输入图像
+
+
+
+
+
+
+
+使用 Gemini Omni Flash 1.1 让图像动起来。使用 `image_to_video` 任务时,第一张附加图像作为起始帧,可选的第二张图像作为结束帧:模型会生成两者之间的画面,让环绕镜头、变焦转场和循环片段变得可预测。
+
+### 参考生成视频(Omni Flash 1.1)
+
+
+
+ 在 Comfy Cloud 中打开
+
+
+ 下载 JSON,或在模板库中搜索“Gemini Omni 1.1”
+
+
+ 获取第一张示例参考图像
+
+
+ 获取第二张示例参考图像
+
+
+
+
+
+
+
+使用最多 14 张参考图像中特定主体生成视频。在参考模式下,使用 `` 等标签将每张图像绑定到角色,并在提示中引用这些标签:图像中的角色、产品和物体将出现在场景中,而图像本身不会作为画面帧使用。将角色参考与风格参考结合,可保持品牌内容的一致性。
+
+### 视频编辑(Omni Flash 1.1)
+
+
+
+ 在 Comfy Cloud 中打开
+
+
+ 下载 JSON,或在模板库中搜索“Gemini Omni 1.1”
+
+
+ 获取此工作流的示例输入视频
+
+
+
+
+
+
+
+使用 Gemini Omni Flash 1.1 以自然语言编辑视频。使用 `edit` 任务时,节点只接受一个输入视频(10 秒以内),并根据你的指令重写它:替换背景、改变场景风格、添加或移除元素。`edit` 和 `extend` 任务会保持输入视频的画面比例。提示词越简洁效果越好;加上“其余保持不变”可以最大化一致性。
+
+### 视频延续(Omni Flash 1.1)
+
+
+
+ 在 Comfy Cloud 中打开
+
+
+ 下载 JSON,或在模板库中搜索“Gemini Omni 1.1”
+
+
+ 获取此工作流的示例输入视频
+
+
+
+
+
+使用 `extend` 任务将现有视频每步延续最多 10 秒,累计可构建约 40 秒的故事。模型会分析最后 10 秒的上下文,在场景从中断处继续时保持角色、动作和音频的连贯。可选择附上参考图像,在故事中途引入新角色。延续只在片段末尾追加新内容,但为了让过渡无缝,模型可能会修改最后的源帧。
+
+## 开始使用
+
+1. 将 ComfyUI 更新到最新版本(Omni Flash 1.1 工作流需要 0.34.2 或更高版本)
+2. 双击画布,搜索“Gemini Omni Flash”节点
+3. 或者进入模板库,使用现成的工作流
+4. 选择与输入类型(文本、图像或视频)匹配的工作流
+5. 输入提示并生成
+
+
+为获得最佳效果,可将 Gemini Omni Flash 与 Nano Banana 2 Lite 组合使用:先高速生成图像,再用 Gemini Omni Flash 将它们动画化为视频。
+