For providers
Put your endpoints in the router pool.
llmcloud.ai routes millions of requests across upstream providers. If your endpoints win on price, latency, or quality, we send you traffic and pay commissions monthly.
Developer distribution
Instant reach into every app using our SDK — no per-app integrations.
Fair routing
Requests are matched on capability, price, and live health. No pay-to-win.
Zero rewrite
If you speak OpenAI-compatible /v1/chat/completions, you're ready.
Requirements
✓OpenAI-compatible API
/v1/chat/completions, /v1/embeddings, streaming, and tool calls.
✓Public health endpoint
GET /health returning region, model list, and version.
✓≥99.5% 30-day uptime
Measured from our probes across 3 regions.
✓Rate-limit headers
Standard X-RateLimit-* headers so we can back off cleanly.
✓Per-model pricing manifest
JSON manifest we can poll for input/output/image pricing.
✓Signed webhooks
Optional but recommended for usage reconciliation.
How onboarding works
- 01ApplyFill out the form below. We reply within 3 business days.
- 02SandboxYou get a scoped test key. We run compatibility, streaming, and tool-use suites.
- 03Load & latencyWe benchmark p50/p95 tps and TTFT across 3 regions for 48h.
- 04PublishApproved endpoints appear in the catalog and enter the router pool.
- 05SettleMonthly reports plus programmatic access to per-model usage and commissions.
Apply
Fields below build your provider profile draft. Nothing goes live until you review it.