Baseten
Model deployment with fast cold starts and per-minute billing, so idle time costs nothing.
One API key, hundreds of models, automatic failover. The easiest way to compare models on price before committing.
Usage-based plans are shown as Metered and are excluded from stack totals.
| Category | Infra |
|---|---|
| Company | OpenRouter |
| Free plan | No |
| Starting price | Usage-based |
| Top plan | Not verified |
| Pricing model | usage |
| API | Yes |
| Open weights | Not verified |
| Commercial use | Not verified |
| Rating | Not rated — we do not publish scores we cannot source. |
Last verified: 17 Aug 2026
Pricing source: openrouter.ai
Listing status: Unclaimed, verified
Are you the owner of OpenRouter? Claim this listing.
Model deployment with fast cold starts and per-minute billing, so idle time costs nothing.
Custom silicon delivering very high tokens-per-second on open models.
Serverless GPUs for your own inference and training jobs, billed per second with a free starter allowance.
Run open models locally on your own machine.
Managed Ray for distributed training and inference, either hosted or inside your own cloud account.
Token-metered serving for open-weight models at some of the lowest published rates, with no contracts or upfront cost.