Model catalog
2 models, one key, metered per million tokens
Rates below are list prices in USD per 1M tokens. Your prepaid credits cover every model — pick per request, no per-vendor setup.
| Model | Provider | Context | Input / 1M | Output / 1M | |
|---|---|---|---|---|---|
|
Gemini 1.5 Pro
Long context
Two-million-token context window for whole-codebase and long-video reasoning.
|
2000K | $1.25 | $5 | use → | |
|
Gemini 1.5 Flash
Cost-efficient multimodal model with a one-million-token window.
|
1000K | $0.075 | $0.3 | use → |
OpenAI-compatible
Point any OpenAI SDK, LangChain or LlamaIndex client at the gateway. Keep your code, change one URL.
Failover included
A rate-limited provider is retried against spare capacity — transparently, with the same request id.
Need a model we don't list?
Tell us at [email protected] and we will wire it into the gateway.