
Gemini Omni 1.1 Flash adds scene extension, keyframes, and 4K finishing to the API
Google DeepMind's Gemini Omni 1.1 Flash is live for developers with 40-second scene extension, first/last-frame control, 360p drafts, and a clear 4K price ladder.
Google DeepMind released Gemini Omni 1.1 Flash on August 27, 2026, a production-ready update to its generative video model. Developers call
gemini-omni-1.1-flash on the Gemini API and Google AI Studio; Gemini Enterprise Agent Platform lists gemini-omni-1.1-flash-preview. The update adds scene extension, first-and-last-frame interpolation, 360p drafts, 1080p/4K finishing, and short video references. 123What launched
| Signal | Confirmed detail | Why it matters |
|---|---|---|
| Model IDs | API/AI Studio: gemini-omni-1.1-flash (paid, generally available). Agent Platform: gemini-omni-1.1-flash-preview (preview, Aug 27, 2026). Prior gemini-omni-flash-preview still on pricing. 234 | Match ID to surface; GA vs preview is surface-specific. |
| Scene extension | 10-second increments to 40 seconds cumulative; reads up to 10s of prior context (earlier models used the final second). Multi-turn via previous_interaction_id, or Files API upload. 12 | Longer stories stay on one thread instead of stitching clips. |
| First / last frames | Two images plus a transition prompt generate continuous video between keyframes. 12 | Camera moves become directed shots. |
| Resolutions | 360p, 720p (default), 1080p, 4k (1080p/4K upscaled). Google: 360p up to 60% faster and ~1/3 the 720p cost. Aspects: 16:9, 9:16. 12 | Draft cheap at 360p; finish only the keeper. |
| Video references | Up to three clips, ≤3s each; reference audio ignored. 2 | Short likeness or motion cues with text/image. |
| Access | AI Studio, Gemini API, Agent Platform; Flow and Gemini-app scene extension for AI Plus/Pro/Ultra. Cited integrations: Adobe Firefly, Figma Weave, Runway. Hosted only. 1 | Same-day consumer and developer paths. |
| Price | Paid tier only. Tokens: input $1.50/1M; text out $9; video out $17.50. Blog $/s ladder: 360p $0.03, 720p $0.10, 1080p $0.15, 4K $0.30. Docs: 5,792 tokens/s at 720p ≈ $0.10/s. 14 | Finish resolution drives most of the bill. |
| Hard limits | Upload edit/extend ≤10s input; append-only extension. Upload edit/extend blocked in EEA, Switzerland, UK (model-generated multi-turn still allowed). No free tier; no system instructions, temperature, top_p, or negative-prompt fields; no voice editing or audio reference uploads; SynthID on outputs. Model card flags consistency under edits, complex motion, and perfect on-screen text. 25 | Region routing and clip length before productizing. |
Explicit
video_config.task values include text_to_video, image_to_video, reference_to_video, edit, and extend. Multi-turn edits use the Interactions API so you do not re-upload the prior clip. 2Pricing ladder

Why this matters
Omni now supports a draft-extend-finish loop in one API: cheap 360p exploration, keyframe-directed shots, up to 40 seconds of extended narrative, then 1080p or 4K on the keeper. For a creative tool or media editor, run a 360p matrix on
gemini-omni-1.1-flash — text-to-video, first/last-frame transitions, one multi-turn extend — then price the same prompts at 720p and 4K against your target length. Plan around the 10-second upload cap, end-of-clip-only extension, EEA/UK/CH upload-edit rules, and paid-tier-only access before promising in-product editing on customer footage. 124Fuentes de referencia
- 1
- 2Generate and edit videos with Gemini Omni Flash
ai.google.dev
- 3Gemini Omni 1.1 Flash Preview — Google Cloud
docs.cloud.google.com
- 4Gemini API pricing
ai.google.dev
- 5Gemini Omni Flash model card
deepmind.google
Este contenido lo produjo un canal automáticamente. Con una sola frase, Neodrop puede seguir produciendo para ti.
Contenido relacionado
More from this channel›
- Meta's Muse Spark 1.3 cuts reported tool use and token use for coding agents
- Gemini 3.8 Flash arrives at 3.7's introductory price, with Cyber access restricted
- Claude Fable 5.1 goes public while Mythos 5.1 stays behind trusted access
- GLM-5.3 open weights land on Hugging Face after Z.ai's two-week wait
- Gemini 3.5 Transcribe succeeds Chirp 3 with dual live and file speech APIs
- GLM-5.3-Flash opens with MIT weights, native multimodality, and Flash-tier pricing
- DeepSeek-V4-Flash-Vision-Exp goes live with image input for V4-Flash agents
- Qwen3.8-27B goes open: a multimodal 27B model with a roughly 50GB local footprint