List on OpenRouter
Integrate your inference API with the OpenRouter network.
OpenRouter routes requests across 80+ providers to serve 10M+ developers. We review every application to maintain API reliability and performance standards across the network.
How the network works
Unified API Surface
Your endpoints are accessible to 10M+ developers through a single OpenAI-compatible API — no additional integration work on their side.
Performance-Based Routing
Requests are routed based on latency, throughput, uptime, and price. Providers that perform well receive proportionally more traffic.
Public Performance Metrics
TTFT, throughput, and uptime are tracked publicly on every model page. These metrics are transparent to developers choosing providers.
Automated Payments
Usage-based billing handled via monthly invoicing. Token counts are reconciled automatically.
Uptime Monitoring
Endpoint reliability is continuously monitored. Providers maintaining 95%+ uptime retain standard routing priority.
Geographic Routing
Declare your datacenter locations in the /models endpoint. Routing respects geographic preferences and data residency requirements.
Technical requirements
All providers must meet these requirements before being considered for the network. Applications that don't meet these criteria will not be reviewed.
View full technical requirements →OpenAI-Compatible API
Your /chat/completions endpoint must be OpenAI-compatible, return usage tokens for both stream and non-stream requests, and support streaming.
List Models Endpoint
Publish a /models endpoint returning your available models with pricing, context length, max output tokens, supported features, and datacenter locations.
Automated Payment
Support monthly invoicing so OpenRouter can pay for inference without manual intervention.
Privacy & Data Policy
Have a published privacy policy and clear data retention terms. Providers must disclose whether prompts are logged and if data is used for training.
How it works
- 1
Submit your application
Provide details about your infrastructure, API endpoints, supported models, and data policies.
- 2
Technical review
Our team evaluates your API compatibility, endpoint reliability, pricing, and performance against network standards.
- 3
Integration & testing
Accepted providers are onboarded with test traffic to validate latency, throughput, and error handling.
- 4
Go live
Your models become available on the network and begin receiving production requests routed by performance and price.
FAQ
Can't find what you need?
Reach out to [email protected]
Submit your application
We review applications on a rolling basis. Due to high demand, not all providers will be accepted. Priority is given to providers that fill gaps in our current network.