modelprices.xyz
activenormalized LLM market data
Per-token prices, capability limits, and a price-change feed for 2,000+ LLMs across 70+ providers, normalized into single JSON tables and refreshed hourly.
Settled via Coinbase.
- Transactions · 30d
- 73
- Volume · 30d
- $0.67
- Unique buyers · 30d
- 8
- Uptime · 30d
- 100.0%
- Latency p50
- 113ms
- Reported calls · 30d
- 64
Endpoints (54 live)
GET/jamba-pricing— Jamba pricing table: what every Jamba model costs per token right now — Jamba 1.5 Large, Jamba Mini — from AI21 and every provider that resells or hosts them (Bedrock, Azure, Vertex, OpenRouter, Fireworks). Input, output, cache and batch USD per 1M tokens, sorted cheapest first, so you can compare Jamba inference cost across hosts and pick the cheapest place to run one. Normalized, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/gemini-pricing— Gemini pricing table: what every Gemini model costs per token right now — Gemini 3 Pro, Gemini 3 Flash, Gemini 2.5 — from Google and every provider that resells or hosts them (Bedrock, Azure, Vertex, OpenRouter, Fireworks). Input, output, cache and batch USD per 1M tokens, sorted cheapest first, so you can compare Gemini inference cost across hosts and pick the cheapest place to run one. Normalized, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/deepseek-pricing— DeepSeek pricing table: what every DeepSeek model costs per token right now — DeepSeek V4, DeepSeek R1, DeepSeek Coder — from DeepSeek and every provider that resells or hosts them (Bedrock, Azure, Vertex, OpenRouter, Fireworks). Input, output, cache and batch USD per 1M tokens, sorted cheapest first, so you can compare DeepSeek inference cost across hosts and pick the cheapest place to run one. Normalized, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/command-pricing— Command pricing table: what every Command model costs per token right now — Command A, Command R+, Command R — from Cohere and every provider that resells or hosts them (Bedrock, Azure, Vertex, OpenRouter, Fireworks). Input, output, cache and batch USD per 1M tokens, sorted cheapest first, so you can compare Command inference cost across hosts and pick the cheapest place to run one. Normalized, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/claude-pricing— Claude pricing table: what every Claude model costs per token right now — Claude 5, Claude Opus, Claude Sonnet, Claude Haiku — from Anthropic and every provider that resells or hosts them (Bedrock, Azure, Vertex, OpenRouter, Fireworks). Input, output, cache and batch USD per 1M tokens, sorted cheapest first, so you can compare Claude inference cost across hosts and pick the cheapest place to run one. Normalized, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices— Compare AI model prices side-by-side: a live price leaderboard of LLM token cost for GPT-5, Claude 5, Claude Sonnet, Gemini 3 Pro, Llama 4, DeepSeek V4, Grok 4, Mistral Large and 2,000+ more models across 70+ providers (OpenAI, Anthropic, Google, Meta, xAI). Inference cost per token — input, output, cache and batch USD per 1M tokens — ranked, normalized into one table, cross-checked across two sources, refreshed hourly. Find the cheapest model and estimate token budgets. (0.02 USDC on Base)GET/gemma-pricing— Gemma pricing table: what every Gemma model costs per token right now — Gemma 3, Gemma 2 — from Google and every provider that resells or hosts them (Bedrock, Azure, Vertex, OpenRouter, Fireworks). Input, output, cache and batch USD per 1M tokens, sorted cheapest first, so you can compare Gemma inference cost across hosts and pick the cheapest place to run one. Normalized, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/glm-pricing— GLM pricing table: what every GLM model costs per token right now — GLM-4.6, GLM-4.5 Air — from Zhipu AI and every provider that resells or hosts them (Bedrock, Azure, Vertex, OpenRouter, Fireworks). Input, output, cache and batch USD per 1M tokens, sorted cheapest first, so you can compare GLM inference cost across hosts and pick the cheapest place to run one. Normalized, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/vercel— Vercel AI Gateway pricing table: per-token cost of every AI model Vercel AI Gateway serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside Vercel AI Gateway and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/snowflake— Snowflake Cortex pricing table: per-token cost of every AI model Snowflake Cortex serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside Snowflake Cortex and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/sambanova— SambaNova pricing table: per-token cost of every AI model SambaNova serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside SambaNova and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/replicate— Replicate pricing table: per-token cost of every AI model Replicate serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside Replicate and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/openai— OpenAI pricing table: per-token cost of every AI model OpenAI serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside OpenAI and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/novita— Novita AI pricing table: per-token cost of every AI model Novita AI serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside Novita AI and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/mistral— Mistral AI pricing table: per-token cost of every AI model Mistral AI serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside Mistral AI and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/google— Google AI Studio pricing table: per-token cost of every AI model Google AI Studio serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside Google AI Studio and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/databricks— Databricks pricing table: per-token cost of every AI model Databricks serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside Databricks and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/cloudflare— Cloudflare Workers AI pricing table: per-token cost of every AI model Cloudflare Workers AI serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside Cloudflare Workers AI and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/bedrock— AWS Bedrock pricing table: per-token cost of every AI model AWS Bedrock serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside AWS Bedrock and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)GET/llm/prices/azure— Microsoft Azure pricing table: per-token cost of every AI model Microsoft Azure serves, in one call — input, output, cache and batch USD per 1M tokens, ranked cheapest first, with context window and capability flags joined in. Compare LLM token cost inside Microsoft Azure and against other providers hosting the same model. Normalized from public sources, cross-checked, refreshed hourly. (0.01 USDC on Base)
+34 more endpoints.
First seen · last seen · last active