Compare

Qwen3.8 Max vs Gemini Omni Flash

Pricing per million tokens, context window, capabilities — pulled from each provider's public docs. All 2 are available via the same AIgateway OpenAI-compatible endpoint; flip the model string to switch.

Search2/4
Qwen3.8 Max
alibaba/qwen3.8-max
Gemini Omni Flash
google/gemini-omni-flash
Provider
Alibaba
Google
Family
Qwen
Modality
text
video
Context window
1,000,000 tok
Max output
128,000 tok
Released
2026-08-03
2026-09-09
License
Open-weight
Proprietary
Input price
$2.00 /1M
Output price
$6.00 /1M
Per output second
$1.50 /sec
Tools
yes
Streaming
yes
Vision
yes
JSON mode
yes
Reasoning
yes
Prompt caching
Batch API
Try it
Open in playground →
View model →
Qwen3.8 Max
alibaba/qwen3.8-max
Full spec →

Alibaba's Qwen 3.8 Max is a 2.4-trillion-parameter MoE flagship built for professional-grade coding and long-horizon autonomous work, capable of delivering complete, production-grade projects spanning 10+ days across legal, financial, design, and other specialized domains. Native visual understanding of images and extended video runs through the full plan-execute-verify cycle, served via DashScope's OpenAI-compatible endpoint.

Strengths
  • General-purpose chat
  • Long context
  • Tool use
Use cases
ChatbotsContent generationAgentic workflows
Gemini Omni Flash
google/gemini-omni-flash
Full spec →

Preview high-performance multimodal video generation and editing model with conversational controls and generated audio.

Strengths
  • Text-to-video generation
Use cases
PromosDemosStoryboards

Compare with another

Claude Fable 5.1 vs Qwen3.8 Max
anthropic/claude-fable-5.1 · alibaba/qwen3.8-max
Claude Fable 5.1 vs Gemini Omni Flash
anthropic/claude-fable-5.1 · google/gemini-omni-flash
Gemini 3.8 Flash vs Qwen3.8 Max
google/gemini-3.8-flash · alibaba/qwen3.8-max
Gemini 3.8 Flash vs Gemini Omni Flash
google/gemini-3.8-flash · google/gemini-omni-flash
Claude Opus 5 vs Qwen3.8 Max
anthropic/claude-opus-5 · alibaba/qwen3.8-max
Claude Opus 5 vs Gemini Omni Flash
anthropic/claude-opus-5 · google/gemini-omni-flash
Glm-5.3 vs Qwen3.8 Max
zai-org/glm-5.3 · alibaba/qwen3.8-max
Glm-5.3 vs Gemini Omni Flash
zai-org/glm-5.3 · google/gemini-omni-flash
Grok 4.6 vs Qwen3.8 Max
xai/grok-4.6 · alibaba/qwen3.8-max
SWITCH BETWEEN THEM

One key, all 2, one line different.

from openai import OpenAI

client = OpenAI(
    base_url="https://api.aigateway.sh/v1",
    api_key="sk-aig-...",
)

# Qwen3.8 Max
client.chat.completions.create(
    model="alibaba/qwen3.8-max",
    messages=[{"role":"user","content":"hello"}],
)

# Gemini Omni Flash
client.chat.completions.create(
    model="google/gemini-omni-flash",
    messages=[{"role":"user","content":"hello"}],
)
Get an AIgateway keyRun an eval on these →