deepseek/deepseek-v4-flash-0731 max output tokens update
deepseek/deepseek-v4-flash-0731 changed max output tokens from 384000 to 65536.
AI model changelog
Track new AI models, removed models, price changes, context updates, status changes, and provider changes from AI Pricing Hub snapshots.
Latest changes
deepseek/deepseek-v4-flash-0731 changed max output tokens from 384000 to 65536.
deepseek/deepseek-v4-flash-0731 changed from $0.420 to $0.270 combined input plus output price per 1M tokens.
gryphe/mythomax-l2-13b changed from $0.120 to $0.190 combined input plus output price per 1M tokens.
meta-llama/llama-3.3-70b-instruct changed max output tokens from 128000 to 16384.
meta-llama/llama-3.3-70b-instruct changed from $0.530 to $0.420 combined input plus output price per 1M tokens.
mistralai/mistral-small-3.2-24b-instruct changed max output tokens from 0 to 16384.
mistralai/mistral-small-3.2-24b-instruct changed from $0.400 to $0.275 combined input plus output price per 1M tokens.
moonshotai/kimi-k2.6 changed from $4.95 to $3.07 combined input plus output price per 1M tokens.
openai/gpt-oss-20b changed from $0.170 to $0.160 combined input plus output price per 1M tokens.
~deepseek/deepseek-v4-flash-latest first appeared in the AI Pricing Hub provider snapshot.
qwen/qwen3.8-max first appeared in the AI Pricing Hub provider snapshot.
thinkingmachines/inkling-small first appeared in the AI Pricing Hub provider snapshot.
anthropic/claude-fable-5:batch is no longer active in the current pricing snapshot.
anthropic/claude-haiku-4.5:batch is no longer active in the current pricing snapshot.
anthropic/claude-opus-4.1:batch is no longer active in the current pricing snapshot.
anthropic/claude-opus-4.5:batch is no longer active in the current pricing snapshot.
anthropic/claude-opus-4.6:batch is no longer active in the current pricing snapshot.
anthropic/claude-opus-4.7:batch is no longer active in the current pricing snapshot.
anthropic/claude-opus-4.8:batch is no longer active in the current pricing snapshot.
anthropic/claude-sonnet-4.5:batch is no longer active in the current pricing snapshot.
anthropic/claude-sonnet-5:batch is no longer active in the current pricing snapshot.
google/gemini-2.5-flash-lite:batch is no longer active in the current pricing snapshot.
google/gemini-2.5-flash:batch is no longer active in the current pricing snapshot.
google/gemini-2.5-pro:batch is no longer active in the current pricing snapshot.
google/gemini-3-flash-preview:batch is no longer active in the current pricing snapshot.
google/gemini-3.1-flash-lite:batch is no longer active in the current pricing snapshot.
google/gemini-3.1-pro-preview:batch is no longer active in the current pricing snapshot.
google/gemini-3.5-flash-lite:batch is no longer active in the current pricing snapshot.
google/gemini-3.5-flash:batch is no longer active in the current pricing snapshot.
google/gemini-3.6-flash:batch is no longer active in the current pricing snapshot.
minimax/minimax-m3:batch is no longer active in the current pricing snapshot.
mistralai/devstral-2512 is no longer active in the current pricing snapshot.
openai/gpt-5-mini:batch is no longer active in the current pricing snapshot.
openai/gpt-5-nano:batch is no longer active in the current pricing snapshot.
openai/gpt-5:batch is no longer active in the current pricing snapshot.
openai/gpt-5.1-chat is no longer active in the current pricing snapshot.
openai/gpt-5.1:batch is no longer active in the current pricing snapshot.
openai/gpt-5.2:batch is no longer active in the current pricing snapshot.
openai/gpt-5.4-mini:batch is no longer active in the current pricing snapshot.
openai/gpt-5.4-nano:batch is no longer active in the current pricing snapshot.
openai/gpt-5.4:batch is no longer active in the current pricing snapshot.
openai/gpt-5.5:batch is no longer active in the current pricing snapshot.
qwen/qwen3-235b-a22b-2507 changed max output tokens from 16384 to 32768.
qwen/qwen3-235b-a22b-2507 changed from $0.640 to $0.747 combined input plus output price per 1M tokens.
qwen/qwen3-next-80b-a3b-instruct changed max output tokens from 262144 to 16384.
qwen/qwen3-next-80b-a3b-instruct changed from $1.20 to $1.19 combined input plus output price per 1M tokens.
qwen/qwen3-next-80b-a3b-thinking changed max output tokens from 32768 to 0.
qwen/qwen3-vl-235b-a22b-thinking changed from $4.40 to $4.93 combined input plus output price per 1M tokens.
qwen/qwen3.5-122b-a10b changed from $2.34 to $3.60 combined input plus output price per 1M tokens.
qwen/qwen3.6-27b changed max output tokens from 65536 to 131072.
qwen/qwen3.6-27b changed from $2.30 to $2.69 combined input plus output price per 1M tokens.
thedrummer/unslopnemo-12b changed max output tokens from 32768 to 1024000.
z-ai/glm-5.2 changed max output tokens from 128000 to 262144.
z-ai/glm-5.2 changed from $4.64 to $3.18 combined input plus output price per 1M tokens.
~moonshotai/kimi-latest changed max output tokens from 0 to 1048576.
~moonshotai/kimi-latest changed from $18.00 to $16.90 combined input plus output price per 1M tokens.
deepseek/deepseek-chat changed from $1.00 to $1.29 combined input plus output price per 1M tokens.
gpt-5.4-mini changed context window from 1,050,000 token context to 272,000 token context.
gpt-5.4-mini-2026-03-17 changed context window from 1,050,000 token context to 272,000 token context.
gpt-5.4-nano changed context window from 1,050,000 token context to 272,000 token context.
gpt-5.4-nano-2026-03-17 changed context window from 1,050,000 token context to 272,000 token context.
gpt-5.6-luna changed from $7.00 to $1.40 combined input plus output price per 1M tokens.
gpt-5.6-terra changed from $17.50 to $14.00 combined input plus output price per 1M tokens.
nvidia/nemotron-3-nano-30b-a3b changed max output tokens from 228000 to 262144.
nvidia/nemotron-3-ultra-550b-a55b changed max output tokens from 16384 to 0.
nvidia/nemotron-3-ultra-550b-a55b changed from $2.70 to $4.20 combined input plus output price per 1M tokens.
openai/gpt-5.6-luna changed from $3.50 to $0.700 combined input plus output price per 1M tokens.
openai/gpt-5.6-luna-pro changed from $3.50 to $0.700 combined input plus output price per 1M tokens.
openai/gpt-5.6-terra changed from $8.75 to $7.00 combined input plus output price per 1M tokens.
openai/gpt-5.6-terra-pro changed from $8.75 to $7.00 combined input plus output price per 1M tokens.
deepseek/deepseek-v4-flash-0731 first appeared in the AI Pricing Hub provider snapshot.
openai/gpt-5-codex is no longer active in the current pricing snapshot.
openai/o3-deep-research is no longer active in the current pricing snapshot.
openai/o4-mini-deep-research is no longer active in the current pricing snapshot.
poolside/laguna-s-2.1 changed from $0.300 to $0.270 combined input plus output price per 1M tokens.
qwen/qwen3-235b-a22b-thinking-2507 changed max output tokens from 32768 to 0.
qwen/qwen3-235b-a22b-thinking-2507 changed from $3.30 to $2.53 combined input plus output price per 1M tokens.
qwen/qwen3-vl-30b-a3b-instruct changed max output tokens from 16384 to 32768.
qwen/qwen3-vl-30b-a3b-instruct changed from $0.750 to $0.650 combined input plus output price per 1M tokens.
qwen/qwen3.7-max changed max output tokens from 65536 to 131072.
qwen/qwen3.7-plus changed max output tokens from 65536 to 131072.
google/gemini-3.1-flash-lite-image changed max output tokens from 66000 to 65536.
google/gemma-4-31b-it changed from $0.540 to $0.440 combined input plus output price per 1M tokens.
qwen/qwen-2.5-7b-instruct changed from $0.140 to $0.300 combined input plus output price per 1M tokens.
command-a-03-2025 first appeared in the AI Pricing Hub provider snapshot.
command-r-08-2024 first appeared in the AI Pricing Hub provider snapshot.
command-r-plus-08-2024 first appeared in the AI Pricing Hub provider snapshot.
command-r7b-12-2024 first appeared in the AI Pricing Hub provider snapshot.
Cohere first appeared in the tracked provider dataset.
deepseek-v4-flash first appeared in the AI Pricing Hub provider snapshot.
deepseek-v4-pro first appeared in the AI Pricing Hub provider snapshot.
DeepSeek first appeared in the tracked provider dataset.
llama-3.1-8b-instant first appeared in the AI Pricing Hub provider snapshot.
llama-3.3-70b-versatile first appeared in the AI Pricing Hub provider snapshot.
meta-llama/llama-prompt-guard-2-22m first appeared in the AI Pricing Hub provider snapshot.
meta-llama/llama-prompt-guard-2-86m first appeared in the AI Pricing Hub provider snapshot.
openai/gpt-oss-120b first appeared in the AI Pricing Hub provider snapshot.
openai/gpt-oss-20b first appeared in the AI Pricing Hub provider snapshot.
openai/gpt-oss-safeguard-20b first appeared in the AI Pricing Hub provider snapshot.
qwen/qwen3.6-27b first appeared in the AI Pricing Hub provider snapshot.
Groq first appeared in the tracked provider dataset.
~anthropic/claude-fable-latest first appeared in the AI Pricing Hub provider snapshot.
~anthropic/claude-haiku-latest first appeared in the AI Pricing Hub provider snapshot.
~anthropic/claude-opus-latest first appeared in the AI Pricing Hub provider snapshot.
~anthropic/claude-sonnet-latest first appeared in the AI Pricing Hub provider snapshot.
~google/gemini-flash-latest first appeared in the AI Pricing Hub provider snapshot.
~google/gemini-pro-latest first appeared in the AI Pricing Hub provider snapshot.
~moonshotai/kimi-latest first appeared in the AI Pricing Hub provider snapshot.
~openai/gpt-latest first appeared in the AI Pricing Hub provider snapshot.
~openai/gpt-mini-latest first appeared in the AI Pricing Hub provider snapshot.
~x-ai/grok-latest first appeared in the AI Pricing Hub provider snapshot.
ai21/jamba-large-1.7 first appeared in the AI Pricing Hub provider snapshot.
aion-labs/aion-2.0 first appeared in the AI Pricing Hub provider snapshot.
aion-labs/aion-3.0 first appeared in the AI Pricing Hub provider snapshot.
aion-labs/aion-3.0-mini first appeared in the AI Pricing Hub provider snapshot.
aion-labs/aion-rp-llama-3.1-8b first appeared in the AI Pricing Hub provider snapshot.
allenai/olmo-3-32b-think first appeared in the AI Pricing Hub provider snapshot.
amazon/nova-2-lite-v1 first appeared in the AI Pricing Hub provider snapshot.
amazon/nova-lite-v1 first appeared in the AI Pricing Hub provider snapshot.
amazon/nova-micro-v1 first appeared in the AI Pricing Hub provider snapshot.
Continue
Public API
Use static JSON endpoints for providers, models, rankings, history, market metrics, and changelog events.
Editorial information
2026-08-04
Methodology explains collection, validation, limitations, and update cadence.