Gateway operational · 16 models live · GPU capacity on demand Read the docs →
LINGYUNS gateway
Model catalog

2 models, one key, metered per million tokens

Rates below are list prices in USD per 1M tokens. Your prepaid credits cover every model — pick per request, no per-vendor setup.

Model Provider Context Input / 1M Output / 1M
Claude 3.5 Sonnet
Anthropic flagship — best-in-class writing, analysis and agentic coding.
Anthropic 200K $3 $15 use →
Claude 3.5 Haiku
Fastest Claude — near-instant responses for chat and extraction.
Anthropic 200K $0.8 $4 use →

OpenAI-compatible

Point any OpenAI SDK, LangChain or LlamaIndex client at the gateway. Keep your code, change one URL.

Failover included

A rate-limited provider is retried against spare capacity — transparently, with the same request id.

Need a model we don't list?

Tell us at [email protected] and we will wire it into the gateway.