What does price per 1M tokens mean?
It is the charge for one million input or output tokens. Your estimated cost is input tokens divided by one million times the input rate, plus output tokens divided by one million times the output rate.
Compare the live price you pay for every text model available on AnyModel. Filter by provider, sort input or output rates, and calculate the cost of your own token volume before choosing an API.
Search by model or provider, sort exact rates, then estimate a workload before you integrate.
Estimate API cost from input and output token counts. The result uses the live AnyModel rates in this table.
| Model | Provider | AnyModel per 1M The multiplier scales AnyModel's base rate: 5¢ per 1M tokens.Every rate in this column is final: base 5¢ × the model's multiplier.Your balance is held in tokens: billed = request tokens × the multiplier.Stronger models carry a higher multiplier; cheaper ones sit below ×1. | Vendor list input / output | Cheaper | Context |
|---|---|---|---|---|---|
GPT-5.6 SolPopularcx/gpt-5.6-sol | OpenAI | $0.20×4 | $5.00 / $30.00 | 25× cheaper | 372K |
Claude Opus 5cc/claude-opus-5 | Anthropic | $0.30×6 | $5.00 / $25.00 | 16× cheaper | 1M |
Claude Opus 4.8Popularcc/claude-opus-4-8 | Anthropic | $0.20×4 | $5.00 / $25.00 | 25× cheaper | 1M |
GLM-5.3-Flashglm/glm-5.3-flash | Zhipu | $0.025×0.5 | $0.15 / $0.50 | 5.9× cheaper | 1M |
Claude Sonnet 5cc/claude-sonnet-5 | Anthropic | $0.15×3 | $2.00 / $10.00 | 13× cheaper | 1M |
Kimi K3kmc/k3 | Moonshot | $0.15×3 | $3.00 / $15.00 | 19× cheaper | 256K |
GPT-5.5Popularcx/gpt-5.5 | OpenAI | $0.15×3 | $5.00 / $30.00 | 33× cheaper | 400K |
GPT-5.6 Lunacx/gpt-5.6-luna | OpenAI | $0.075×1.5 | $1.00 / $6.00 | 13× cheaper | 272K |
Claude Opus 4.7cc/claude-opus-4-7 | Anthropic | $0.20×4 | $5.00 / $25.00 | 25× cheaper | 1M |
GPT-5.6 TerraPopularcx/gpt-5.6-terra | OpenAI | $0.15×3 | $2.50 / $15.00 | 16× cheaper | 272K |
Grok 4.6Populargcli/grok-4.6 | xAI | $0.10×2 | $2.00 / $6.00 | 20× cheaper | 500K |
Claude Sonnet 4.6Popularag/claude-sonnet-4-6 | Anthropic | $0.12×2.4 | $3.00 / $15.00 | 25× cheaper | 1M |
GPT-5.4Popularcx/gpt-5.4 | OpenAI | $0.10×2 | $2.50 / $15.00 | 25× cheaper | 400K |
Claude Opus 4.6cc/claude-opus-4-6 | Anthropic | $0.20×4 | $5.00 / $25.00 | 25× cheaper | 1M |
GLM-5.3glm/glm-5.3 | Zhipu | $0.075×1.5 | $1.40 / $4.40 | 18× cheaper | 200K |
Kimi K2.7 Codekmc/kimi-for-coding | Moonshot | $0.075×1.5 | — | — | 256K |
Claude Haiku 4.5cc/claude-haiku-4-5-20251001 | Anthropic | $0.08×1.6 | $1.00 / $5.00 | 12× cheaper | 200K |
Free Models Auto Routeram/free | AnyModel | $0.00×0 | — | — | Varies |
GPT-5.4 minicx/gpt-5.4-mini | OpenAI | $0.075×1.5 | $0.75 / $4.50 | 9.9× cheaper | 400K |
MiniMax M3am/minimax-m3 | MiniMax | $0.00×0 | — | — | 512K |
Gemini 3.5 Flashag/gemini-3.5-flash-high | $0.03×0.6 | $1.50 / $9.00 | 50× cheaper | 1M | |
Grok 4.20 Multi-Agent ResearchPopularxai/grok-4.20-multi-agent-0309 | xAI | $0.10×2 | $1.25 / $2.50 | 12× cheaper | 1M |
Gemini 3.1 Flash-Liteag/gemini-3.1-flash-lite-preview | $0.03×0.6 | $0.25 / $1.50 | 8.3× cheaper | 1M | |
Nemotron 3 Ultraam/nemotron-3-ultra-550b-a55b | NVIDIA | $0.00×0 | — | — | 256K |
Qwen3.7 Maxqwen/qwen3.7-max | Qwen | $0.0125×0.25 | $2.50 / $7.50 | 200× cheaper | 1M |
Gemma 4 31Bam/gemma-4-31b-it | $0.00×0 | — | — | 128K | |
Kimi K3 1Mam/kimi-k3 | Moonshot | $0.05×1 | $3.00 / $15.00 | 60× cheaper | 1M |
Gemini 2.5 Flashag/gemini-2.5-flash | $0.03×0.6 | — | — | 1M | |
Gemini 2.5 Flash Liteag/gemini-2.5-flash-lite | $0.03×0.6 | — | — | 1M | |
Gemini 2.5 Proag/gemini-2.5-pro | $0.085×1.7 | — | — | 1M | |
Gemini 3 Flashag/gemini-3-flash | $0.03×0.6 | — | — | 1M | |
Gemini 3.6 Flashag/gemini-3.6-flash-high | $0.03×0.6 | — | — | 1M | |
Gemini 3.7 Flashag/gemini-3.7-flash-high | $0.03×0.6 | — | — | 1M | |
Gemini 3.1 Proag/gemini-pro-agent | $0.085×1.7 | — | — | 1M | |
GPT-OSS 120Bag/gpt-oss-120b-medium | OpenAI | $0.025×0.5 | — | — | 128K |
DiffusionGemma 26B A4B ITam/diffusiongemma-26b-a4b-it | $0.00×0 | — | — | 128K | |
GPT-OSS 20Bam/gpt-oss-20b | OpenAI | $0.00×0 | — | — | 128K |
Laguna XS 2.1am/laguna-xs-2.1 | Poolside | $0.00×0 | — | — | 128K |
Llama 3.2 11B Vision Instructam/llama-3.2-11b-vision-instruct | Meta | $0.00×0 | — | — | 128K |
Mistral Nemotronam/mistral-nemotron | Mistral AI | $0.00×0 | — | — | 128K |
Nemotron 3 Nano 30B A3Bam/nemotron-3-nano-30b-a3b | NVIDIA | $0.00×0 | — | — | 256K |
Nemotron 3 Nano Omni 30B A3B Reasoningam/nemotron-3-nano-omni-30b-a3b-reasoning | NVIDIA | $0.00×0 | — | — | 256K |
Nemotron 3 Super 120B A12Bam/nemotron-3-super-120b-a12b | NVIDIA | $0.00×0 | — | — | 256K |
Nemotron 3.5 Content Safetyam/nemotron-3.5-content-safety | NVIDIA | $0.00×0 | — | — | 32K |
Nemotron 3.5 Lightning 30B A3Bam/nemotron-3.5-lightning-30b-a3b | NVIDIA | $0.00×0 | — | — | 256K |
Riva Translate 4B Instruct v2am/riva-translate-4b-instruct-v2 | NVIDIA | $0.00×0 | — | — | 8K |
DeepSeek V4 Flashds/deepseek-v4-flash | DeepSeek | $0.0025×0.05 | — | — | 1M |
DeepSeek-V4-Prods/deepseek-v4-pro | DeepSeek | $0.0075×0.15 | — | — | 1M |
GLM 4.6Vglm/glm-4.6v | Zhipu | $0.015×0.3 | — | — | 128K |
GLM 4.7glm/glm-4.7 | Zhipu | $0.015×0.3 | — | — | 200K |
GLM 5glm/glm-5 | Zhipu | $0.025×0.5 | — | — | 200K |
GLM-5.1glm/glm-5.1 | Zhipu | $0.075×1.5 | — | — | 200K |
GLM-5.2glm/glm-5.2 | Zhipu | $0.075×1.5 | — | — | 200K |
Qwen3.6 Flashqwen/qwen3.6-flash | Qwen | $0.0125×0.25 | — | — | 1M |
Qwen 3.7 Plusqwen/qwen3.7-plus | Qwen | $0.0125×0.25 | — | — | 1M |
Qwen3.8 Maxqwen/qwen3.8-max | Qwen | $0.0125×0.25 | — | — | 1M |
Grok 4.20xai/grok-4.20-0309-reasoning | xAI | $0.05×1 | — | — | 1M |
Grok 4.3xai/grok-4.3 | xAI | $0.05×1 | — | — | 1M |
Grok 4.5xai/grok-4.5 | xAI | $0.10×2 | — | — | 500K |
Grok Build 0.1xai/grok-build-0.1 | xAI | $0.025×0.5 | — | — | 256K |
AnyModel prices are shown per 1M tokens. Most models use one flat rate; models with separate input/output billing show both rates. Vendor list prices are shown only when directly comparable. No subscription, no minimums.
Media is billed per unit of output, not per 1M tokens. A video job costs its duration times the chosen model and resolution rate, up to 15 seconds; an image costs the base rate scaled by quality and size.
| What | Unit | Balance tokens | ≈ price |
|---|---|---|---|
| Grok Imagine Video 480p | per second | 125,000 | $0.0063 |
| Grok Imagine Video 720p | per second | 175,000 | $0.0087 |
| Grok Imagine Video 1.5 480p | per second | 200,000 | $0.01 |
| Grok Imagine Video 1.5 720p | per second | 350,000 | $0.0175 |
| Grok Imagine Video 1.5 1080p | per second | 625,000 | $0.0313 |
| Image (gpt-image-2) | per image, 1024×1024, medium | 250,000 | $0.0125 |
Example: an 8-second 480p clip costs 1M tokens. The charge is booked once, when the provider accepts the job, and polling a running job is free. A job that ends without a video is returned in full — by itself for jobs started in Video Studio, and on the poll that first reports the failure for jobs created through the API.
Input and output rates matter separately. Retrieval, classification and long-document workloads are input-heavy; agents and content generation can spend much more on output.
The AnyModel columns show current customer rates per 1 million tokens. Vendor list prices are context only and may not include caching, batch discounts, regional taxes or special tiers.
Price is only one production variable. Test response quality, latency, rate limits and availability with the same prompts before moving a critical workload.
It is the charge for one million input or output tokens. Your estimated cost is input tokens divided by one million times the input rate, plus output tokens divided by one million times the output rate.
The AnyModel rates are resolved from the same current pricing catalog used by the service. Model availability and rates can change, so verify the page and dashboard before a production launch.
No. AnyModel uses a prepaid pay-as-you-go balance with no monthly platform fee or minimum spend.
Point any OpenAI-compatible client at AnyModel and switch models by name — no separate billing per provider. Fund one prepaid balance and pay only for what you use.