AI Model Providers
Compare pricing and capabilities across leading AI providers
Model Creator15

AI21 Labs
AI21 Labs builds the Jamba family of hybrid SSM-Transformer models with long context windows, offered via its own API with per-token pricing.
View details →
Anthropic
Cohere
DeepSeek

MiniMax
Mistral

Moonshot
OpenAI
Perplexity

Qwen

Runway
Runway builds frontier video-generation models (Gen-4, Aleph) plus hosted partner models (Veo, Seedance), billed per second of generated video via its API.
View details →StepFun
StepFun (阶跃星辰) is a Chinese AI lab building the Step series of multimodal foundation models (text, vision and audio), offered via its own API with per-token pricing and prompt caching.
View details →xAI

Zhipu
Model Aggregator11
Cloudflare Workers AI
Cloudflare Workers AI runs open models (Llama, Qwen, Mistral, Gemma, GPT-OSS, plus embeddings and speech) on Cloudflare's global edge network, billed per token in USD (also denominated in Neurons).
View details →Cursor

DeepInfra
DeepInfra is a serverless inference marketplace running open models (Llama, Qwen, DeepSeek, GLM, Kimi and more) on managed GPUs, billed per token with no idle cost — among the cheapest hosted open-model APIs.
View details →Fireworks AI
Fireworks AI is a high-performance serverless inference platform for open and open-weight frontier models (DeepSeek, GLM, Qwen, Kimi, MiniMax, GPT-OSS), offering Standard and Priority throughput tiers with cached-input discounts.
View details →Groq

Novita AI
Novita AI is an open-model inference aggregator with an OpenAI-compatible API, hosting DeepSeek, Llama, Qwen, GLM, Kimi and MiniMax models with per-token pricing and prompt caching.
View details →OpenCode
OpenCode sells access to the same models two ways. Zen is a pay-per-token API gateway for the OpenCode coding agent, carrying GPT, Claude, Gemini and Grok alongside open models (Qwen, DeepSeek, MiniMax, GLM, Kimi) — several of them free for a limited time. OpenCode Go is a $10/month subscription over the open models only, with its allowance published in dollars of usage.
View details →OpenRouter

SambaNova Cloud
SambaNova Cloud runs open models (DeepSeek, Llama, Qwen) on custom RDU hardware for record-fast token throughput, billed per token.
View details →SiliconFlow

Volcano Ark
Cloud Platform3
Coding Tool13

Anthropic
CodeBuddy
Tencent's coding assistant. CodeBuddy and the WorkBuddy desktop agent bill against one shared credit balance, so a subscription covers both.
View details →Cursor
GitHub Copilot
GitHub's coding assistant, sold only as a subscription. Billing moved from premium requests to AI credits, published at 1 credit = $0.01 USD.
View details →
MiniMax

Moonshot
OpenAI
OpenCode
OpenCode sells access to the same models two ways. Zen is a pay-per-token API gateway for the OpenCode coding agent, carrying GPT, Claude, Gemini and Grok alongside open models (Qwen, DeepSeek, MiniMax, GLM, Kimi) — several of them free for a limited time. OpenCode Go is a $10/month subscription over the open models only, with its allowance published in dollars of usage.
View details →Qoder
Alibaba Cloud's coding assistant, renamed from 通义灵码 (Lingma) on 2026-05-20 and repriced onto a Credits quota the same day.
View details →
Qwen
TRAE
ByteDance's AI IDE. Paid tiers currently sell queue priority (速通) rather than an allowance; a credit-based plan was announced for 2026-07-30.
View details →
