Gateway operational · 16 models live · GPU capacity on demand Read the docs →
LINGYUNS gateway

Service Level Agreement

Last updated September 17, 2026

This SLA describes the availability of the Lingyuns gateway and GPU instances, and the remedies available when we miss the target.

1. Gateway availability

Target: 99.9% monthly availability for the inference gateway, measured per calendar month across all successfully authenticated requests except those failing due to your content, rate limits or upstream model deprecations.

2. Service credits

Below 99.9%: 10% of the affected month's token spend credited. Below 99.0%: 25%. Below 95%: 50%. Credits are applied to your balance within 30 days of a verified claim.

3. GPU instances

GPU instances carry a 99.5% monthly availability target counted from the moment the instance reports ready. Unplanned downtime is credited pro-rata in minutes.

4. Exclusions

Scheduled maintenance (announced 48h ahead), force majeure, your software stack, network paths outside our control, and free trial or beta capacity are excluded from credit calculations.

5. Claims

Open a claim within 30 days of the incident by emailing [email protected] with request IDs and timestamps. Approved credits are the sole remedy under this SLA.

Questions about this document? Write to [email protected] or use the contact form. Our registered office is 242 N Ash Street, Fruita, CO 81521, United States.