For providers

Put your endpoints in the router pool.

llmcloud.ai routes millions of requests across upstream providers. If your endpoints win on price, latency, or quality, we send you traffic and pay commissions monthly.

Developer distribution
Instant reach into every app using our SDK — no per-app integrations.
Fair routing
Requests are matched on capability, price, and live health. No pay-to-win.
Zero rewrite
If you speak OpenAI-compatible /v1/chat/completions, you're ready.

Requirements

OpenAI-compatible API
/v1/chat/completions, /v1/embeddings, streaming, and tool calls.
Public health endpoint
GET /health returning region, model list, and version.
≥99.5% 30-day uptime
Measured from our probes across 3 regions.
Rate-limit headers
Standard X-RateLimit-* headers so we can back off cleanly.
Per-model pricing manifest
JSON manifest we can poll for input/output/image pricing.
Signed webhooks
Optional but recommended for usage reconciliation.

How onboarding works

  1. 01
    Apply
    Fill out the form below. We reply within 3 business days.
  2. 02
    Sandbox
    You get a scoped test key. We run compatibility, streaming, and tool-use suites.
  3. 03
    Load & latency
    We benchmark p50/p95 tps and TTFT across 3 regions for 48h.
  4. 04
    Publish
    Approved endpoints appear in the catalog and enter the router pool.
  5. 05
    Settle
    Monthly reports plus programmatic access to per-model usage and commissions.

Apply

Fields below build your provider profile draft. Nothing goes live until you review it.

By submitting you agree to our provider terms and data-processing addendum.