Pricing per million tokens, context window, capabilities — pulled from each provider's public docs. All 2 are available via the same AIgateway OpenAI-compatible endpoint; flip the model string to switch.
Qwen 3.8 27B is a 27-billion-parameter instruction-tuned language model from Alibaba's Qwen family, designed for vision, efficient general-purpose text generation and agentic workloads.
fal's H3 Max is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality
from openai import OpenAI
client = OpenAI(
base_url="https://api.aigateway.sh/v1",
api_key="sk-aig-...",
)
# Qwen3.8-27b
client.chat.completions.create(
model="qwen/qwen3.8-27b",
messages=[{"role":"user","content":"hello"}],
)
# MiniMax H3 Max Text to Video
client.chat.completions.create(
model="minimax/h3-max",
messages=[{"role":"user","content":"hello"}],
)