Pricing per million tokens, context window, capabilities — pulled from each provider's public docs. All 2 are available via the same AIgateway OpenAI-compatible endpoint; flip the model string to switch.
Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for high-throughput, low-cost execution for subagent tasks and document parsing.
MiniMax's M3 language model with frontier coding and agentic capabilities, a 1M token context window, and multilingual support.
from openai import OpenAI
client = OpenAI(
base_url="https://api.aigateway.sh/v1",
api_key="sk-aig-...",
)
# Gemini 3.5 Flash-Lite
client.chat.completions.create(
model="google/gemini-3.5-flash-lite",
messages=[{"role":"user","content":"hello"}],
)
# MiniMax M3
client.chat.completions.create(
model="minimax/m3",
messages=[{"role":"user","content":"hello"}],
)