AI Model Providers

Compare pricing and capabilities across leading AI providers

Model Aggregator11

Cloudflare Workers AI logo

Cloudflare Workers AI

44 models

Cloudflare Workers AI runs open models (Llama, Qwen, Mistral, Gemma, GPT-OSS, plus embeddings and speech) on Cloudflare's global edge network, billed per token in USD (also denominated in Neurons).

View details
Cursor logo

Cursor

23 models
View details
DeepInfra logo

DeepInfra

160 models

DeepInfra is a serverless inference marketplace running open models (Llama, Qwen, DeepSeek, GLM, Kimi and more) on managed GPUs, billed per token with no idle cost — among the cheapest hosted open-model APIs.

View details
Fireworks AI logo

Fireworks AI

27 models

Fireworks AI is a high-performance serverless inference platform for open and open-weight frontier models (DeepSeek, GLM, Qwen, Kimi, MiniMax, GPT-OSS), offering Standard and Priority throughput tiers with cached-input discounts.

View details
Groq logo

Groq

16 models
View details
Novita AI logo

Novita AI

155 models

Novita AI is an open-model inference aggregator with an OpenAI-compatible API, hosting DeepSeek, Llama, Qwen, GLM, Kimi and MiniMax models with per-token pricing and prompt caching.

View details
OpenCode logo

OpenCode

74 models

OpenCode sells access to the same models two ways. Zen is a pay-per-token API gateway for the OpenCode coding agent, carrying GPT, Claude, Gemini and Grok alongside open models (Qwen, DeepSeek, MiniMax, GLM, Kimi) — several of them free for a limited time. OpenCode Go is a $10/month subscription over the open models only, with its allowance published in dollars of usage.

View details
OpenRouter logo

OpenRouter

519 models
View details
SambaNova Cloud logo

SambaNova Cloud

7 models

SambaNova Cloud runs open models (DeepSeek, Llama, Qwen) on custom RDU hardware for record-fast token throughput, billed per token.

View details
SiliconFlow logo

SiliconFlow

56 models
View details
Volcano Ark logo

Volcano Ark

39 models
View details