Compare

Wan 3.0 Prime (Text To Video) vs Gemini 3.7 Flash

Pricing per million tokens, context window, capabilities — pulled from each provider's public docs. All 2 are available via the same AIgateway OpenAI-compatible endpoint; flip the model string to switch.

Search2/4
Wan 3.0 Prime (Text To Video)
alibaba/wan-3.0-prime
Gemini 3.7 Flash
google/gemini-3.7-flash
Provider
Alibaba
Google
Family
Gemini 3
Modality
video
text
Context window
1,048,576 tok
Max output
65,536 tok
Released
2026-09-01
2026-08-14
License
Open-weight
Proprietary
Input price
$0.750 /1M
Output price
$3.75 /1M
Per output second
$0.050 /sec
Tools
yes
Streaming
yes
Vision
yes
JSON mode
yes
Reasoning
Prompt caching
Batch API
Try it
View model →
Open in playground →
Wan 3.0 Prime (Text To Video)
alibaba/wan-3.0-prime
Full spec →

Wan 3.0 Prime Text-to-Video transforms written prompts into polished videos with accelerated generation, fluid motion, strong scene fidelity, and coherent visual storytelling. Built for fast creative iteration, it brings complex ideas to life while preserving visual detail and cinematic consistency throughout each shot.

Strengths
  • Text-to-video generation
  • Cinematic motion
Use cases
AdsStoryboardsDemos
Gemini 3.7 Flash
google/gemini-3.7-flash
Full spec →

Gemini 3.7 Flash is a highly capable, natively multimodal reasoning model optimized for agentic workflows and real-world tasks.

Strengths
  • General-purpose chat
  • Long context
  • Tool use
Use cases
ChatbotsContent generationAgentic workflows
SWITCH BETWEEN THEM

One key, all 2, one line different.

from openai import OpenAI

client = OpenAI(
    base_url="https://api.aigateway.sh/v1",
    api_key="sk-aig-...",
)

# Wan 3.0 Prime (Text To Video)
client.chat.completions.create(
    model="alibaba/wan-3.0-prime",
    messages=[{"role":"user","content":"hello"}],
)

# Gemini 3.7 Flash
client.chat.completions.create(
    model="google/gemini-3.7-flash",
    messages=[{"role":"user","content":"hello"}],
)
Get an AIgateway keyRun an eval on these →