Gateway operational · 16 models live · GPU capacity on demand Read the docs →
LINGYUNS gateway
Model catalog

1 models, one key, metered per million tokens

Rates below are list prices in USD per 1M tokens. Your prepaid credits cover every model — pick per request, no per-vendor setup.

Model Provider Context Input / 1M Output / 1M
Gemini 1.5 Flash
Cost-efficient multimodal model with a one-million-token window.
Google 1000K $0.075 $0.3 use →

OpenAI-compatible

Point any OpenAI SDK, LangChain or LlamaIndex client at the gateway. Keep your code, change one URL.

Failover included

A rate-limited provider is retried against spare capacity — transparently, with the same request id.

Need a model we don't list?

Tell us at [email protected] and we will wire it into the gateway.