Compare

Gemini Omni Flash 1.1 Reference to Video — and what?

Pricing per million tokens, context window, capabilities — pulled from each provider's public docs. All 1 are available via the same AIgateway OpenAI-compatible endpoint; flip the model string to switch.

Search1/4
Gemini Omni Flash 1.1 Reference to Video
google/gemini-omni-flash/v1.1/reference-to-video
Provider
Google
Family
Modality
video
Context window
Max output
Released
License
Proprietary
Per output second
$0.030 /sec
Tools
Streaming
Vision
JSON mode
Reasoning
Prompt caching
Batch API
Try it
View model →
Gemini Omni Flash 1.1 Reference to Video
google/gemini-omni-flash/v1.1/reference-to-video
Full spec →

Gemini Omni Flash 1.1 is Google's multimodal video model. This endpoint generates video from combined multimodal references, images, videos and text together. Reasoning across all inputs to produce a single coherent result, with characters retaining their face, clothing, and voice throughout

Strengths
  • Text-to-video generation
  • Cinematic motion
Use cases
AdsStoryboardsDemos