Compare

Qwen3.8 Max vs Kimi-K2.6

Pricing per million tokens, context window, capabilities — pulled from each provider's public docs. All 2 are available via the same AIgateway OpenAI-compatible endpoint; flip the model string to switch.

Search2/4
Qwen3.8 Max
alibaba/qwen3.8-max
Kimi-K2.6
moonshot/kimi-k2.6
Provider
Alibaba
Moonshot
Family
Qwen
Kimi
Modality
text
text
Context window
1,000,000 tok
262,144 tok
Max output
4,096 tok
16,384 tok
Released
2026-08-03
2026-04-20
License
Open-weight
Open-weight
Input price
$2.00 /1M
$0.950 /1M
Output price
$6.00 /1M
$4.00 /1M
Cache read
$0.160 /1M
Tools
yes
yes
Streaming
yes
yes
Vision
yes
yes
JSON mode
yes
yes
Reasoning
yes
yes
Prompt caching
yes
Batch API
Try it
Open in playground →
Open in playground →
Qwen3.8 Max
alibaba/qwen3.8-max
Full spec →

Alibaba's Qwen 3.8 Max is a 2.4-trillion-parameter MoE flagship built for professional-grade coding and long-horizon autonomous work, capable of delivering complete, production-grade projects spanning 10+ days across legal, financial, design, and other specialized domains. Native visual understanding of images and extended video runs through the full plan-execute-verify cycle, served via DashScope's OpenAI-compatible endpoint.

Strengths
  • General-purpose chat
  • Long context
  • Tool use
Use cases
ChatbotsContent generationAgentic workflows
Kimi-K2.6
moonshot/kimi-k2.6
Full spec →

Kimi K2.6 is a frontier-scale open-source 1T parameter model with a 262.1k context window, multi-turn tool calling, vision inputs, and structured outputs for agentic workloads.

Strengths
  • Frontier-scale 1T parameters, open-weight
  • ~10× cheaper than Opus
  • Multi-turn tool calling + vision
Use cases
Code agentsLong-document reasoningCost-conscious production

Benchmarks

Qwen3.8 Max
Kimi-K2.6
HumanEval
92.7
SWE-Bench
68.2

Source: each provider's published benchmarks. Higher is better. Run an eval to compare on your own data.

Compare with another

Kimi-K2.6 vs Kimi-K2.7-Code
moonshot/kimi-k2.6 · moonshot/kimi-k2.7-code
Grok 4.6 vs Qwen3.8 Max
xai/grok-4.6 · alibaba/qwen3.8-max
Grok 4.6 vs Kimi-K2.6
xai/grok-4.6 · moonshot/kimi-k2.6
Claude Opus 5 vs Qwen3.8 Max
anthropic/claude-opus-5 · alibaba/qwen3.8-max
Claude Opus 5 vs Kimi-K2.6
anthropic/claude-opus-5 · moonshot/kimi-k2.6
Qwen3.8 Max vs DeepSeek V4 Pro
alibaba/qwen3.8-max · deepseek/deepseek-v4-pro
Qwen3.8 Max vs Kimi K3
alibaba/qwen3.8-max · moonshot/kimi-k3
Qwen3.8 Max vs GPT-5.6 Sol
alibaba/qwen3.8-max · openai/gpt-5.6-sol
Qwen3.8 Max vs Claude Fable 5
alibaba/qwen3.8-max · anthropic/claude-fable-5
SWITCH BETWEEN THEM

One key, all 2, one line different.

from openai import OpenAI

client = OpenAI(
    base_url="https://api.aigateway.sh/v1",
    api_key="sk-aig-...",
)

# Qwen3.8 Max
client.chat.completions.create(
    model="alibaba/qwen3.8-max",
    messages=[{"role":"user","content":"hello"}],
)

# Kimi-K2.6
client.chat.completions.create(
    model="moonshot/kimi-k2.6",
    messages=[{"role":"user","content":"hello"}],
)
Get an AIgateway keyRun an eval on these →