isitdeprecated.com
collectors 5/5 ok last sync state ok models 2,537 events 7,802

AI API changes: August 2026

430 changes shipped without an announcement, and 239 that came with a published date. Compiled automatically from provider catalogs and endpoints; every figure below links to the model it came from.

77

capacity cuts

context or output shrank

221

price changes

same call, different bill

24

capabilities lost

239

announced

deprecations with a date

Unannounced, biggest first

ImpactProvider What changedDetected
truncationthedrummer-97% max out · thedrummer/unslopnemo-12b: max_output_tokens cut 1,024,000 to 32,768 inferred2026-08-24
truncationopenai-97% max out · openai/gpt-4.1-nano: max_output_tokens cut 942,818 to 32,7682026-08-30
truncationnvidia-96% max out · nvidia/nemotron-3-ultra-550b-a55b: max_output_tokens cut 461,059 to 16,3842026-08-28
truncationqwen-94% max out · qwen/qwen3-next-80b-a3b-instruct: max_output_tokens cut 262,144 to 16,384 +1 alias2026-08-19
costdeepseek15.00x price · deepseek/deepseek-v4-pro: cache_read_price_per_mtok $0.0036/Mtok to $0.05/Mtok2026-08-11
costdeepseek12.14x price · deepseek-v4-pro: cache_read_price_per_mtok $0.0036/Mtok to $0.04/Mtok +1 alias2026-08-21
costz-ai11.00x price · z-ai/glm-5.2: output_price_per_mtok $0.22/Mtok to $2.42/Mtok2026-08-10
costz-ai10.86x price · z-ai/glm-5.2: input_price_per_mtok $0.07/Mtok to $0.76/Mtok2026-08-10
costz-ai10.77x price · z-ai/glm-5.2: cache_read_price_per_mtok $0.01/Mtok to $0.14/Mtok2026-08-10
truncationnovita-91% ctx · novita/meta-llama/llama-3.3-70b-instruct: context_tokens cut 131,072 to 12,2882026-08-28
capabilityarcee-ailost frequency_penalty, logit_bias, logprobs, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, top_logprobs · arcee-ai/trinity-large-thinking: capabilities lost frequency_penalty, logit_bias, logprobs, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, top_logprobs inferred2026-08-31
costxai10.00x price · xai/grok-code-fast-1-0825: cache_read_price_per_mtok $0.02/Mtok to $0.20/Mtok +2 aliases2026-08-12
truncationnovita-90% max out · novita/meta-llama/llama-3.3-70b-instruct: max_output_tokens cut 120,000 to 12,2882026-08-28
truncationtogether_ai-88% max out · together_ai/zai-org/GLM-5.2: max_output_tokens cut 1,048,575 to 128,000 +1 alias2026-08-30
truncationgemini-88% ctx · gemini/gemini-omni-flash-preview: context_tokens cut 1,048,576 to 131,0722026-08-30
capabilitygooglelost frequency_penalty, logprobs, presence_penalty, repetition_penalty, stop, structured_outputs, top_k, top_logprobs · google/gemma-4-26b-a4b-it:free: capabilities lost frequency_penalty, logprobs, presence_penalty, repetition_penalty, stop, structured_outputs, top_k, top_logprobs inferred2026-08-20
truncationqwen-88% max out · qwen/qwen3-coder-30b-a3b-instruct: max_output_tokens cut 262,144 to 32,768 +2 aliases2026-08-13
truncationmeta-llama-87% max out · meta-llama/llama-3.3-70b-instruct: max_output_tokens cut 128,000 to 16,3842026-08-04
truncationmeta-86% max out · meta/muse-glimmer-30b: max_output_tokens cut 117,964 to 16,384 inferred2026-08-28
costmeta-llama7.10x price · meta-llama/llama-3.3-70b-instruct: input_price_per_mtok $0.10/Mtok to $0.71/Mtok2026-08-26
costqwen6.67x price · qwen/qwen3.5-397b-a17b: cache_read_price_per_mtok $0.04/Mtok to $0.30/Mtok2026-08-13
costdeepseek6.43x price · deepseek/deepseek-v4-flash-0731: cache_read_price_per_mtok $0.0028/Mtok to $0.02/Mtok2026-08-01
truncationmeta-llama-84% ctx · meta-llama/llama-guard-4-12b: context_tokens cut 1,048,576 to 163,8402026-08-26
costxai6.25x price · xai/grok-4-1-fast-non-reasoning-latest: input_price_per_mtok $0.20/Mtok to $1.25/Mtok +6 aliases2026-08-30
costdeepseek6.07x price · deepseek/deepseek-v4-pro-0813: cache_read_price_per_mtok $0.0036/Mtok to $0.02/Mtok2026-08-16
capabilitytencentlost frequency_penalty, presence_penalty, repetition_penalty, response_format, stop, structured_outputs · tencent/hy3-preview: capabilities lost frequency_penalty, presence_penalty, repetition_penalty, response_format, stop, structured_outputs inferred2026-08-03
truncationdeepseek-83% max out · deepseek/deepseek-v4-flash-0731: max_output_tokens cut 384,000 to 65,536 inferred2026-08-01
truncationdeepseek-80% max out · deepseek/deepseek-v3.1-terminus: max_output_tokens cut 163,840 to 32,7682026-08-20
costxai5.00x price · xai/grok-3-mini-beta: output_price_per_mtok $0.50/Mtok to $2.50/Mtok +9 aliases2026-08-30
costdeepinfra5.00x price · deepinfra/Gryphe/MythoMax-L2-13b: input_price_per_mtok $0.08/Mtok to $0.40/Mtok2026-08-28
costdeepseek5.00x price · deepseek-v4-flash: cache_read_price_per_mtok $0.0028/Mtok to $0.01/Mtok +1 alias2026-08-21
costxai5.00x price · xai/grok-code-fast-1-0825: input_price_per_mtok $0.20/Mtok to $1.00/Mtok +2 aliases2026-08-12
costdeepseek4.71x price · deepseek-v4-flash: output_price_per_mtok $0.28/Mtok to $1.32/Mtok +1 alias2026-08-21
costdeepseek4.55x price · deepseek-v4-pro: output_price_per_mtok $0.87/Mtok to $3.96/Mtok +1 alias2026-08-21
costdeepinfra4.44x price · deepinfra/Gryphe/MythoMax-L2-13b: output_price_per_mtok $0.09/Mtok to $0.40/Mtok2026-08-28
costz-ai4.19x price · z-ai/glm-5.2: output_price_per_mtok $0.89/Mtok to $3.74/Mtok2026-08-03
costz-ai4.19x price · z-ai/glm-5.2: cache_read_price_per_mtok $0.05/Mtok to $0.22/Mtok2026-08-03
costz-ai4.19x price · z-ai/glm-5.2: input_price_per_mtok $0.28/Mtok to $1.19/Mtok2026-08-03
costxai4.17x price · xai/grok-3-mini-beta: input_price_per_mtok $0.30/Mtok to $1.25/Mtok +2 aliases2026-08-30
truncationqwen-75% max out · qwen/qwen2.5-vl-72b-instruct: max_output_tokens cut 115,200 to 28,8002026-08-26
truncationqwen-75% max out · qwen/qwen3.5-122b-a10b: max_output_tokens cut 262,144 to 65,536 +8 aliases2026-08-23
truncation~deepseek-75% max out · ~deepseek/deepseek-v4-flash-latest: max_output_tokens cut 1,048,576 to 262,144 +1 alias inferred2026-08-23
truncationliquid-75% max out · liquid/lfm-2.5-2.6b:free: max_output_tokens cut 32,768 to 8,192 inferred2026-08-14
costxai4.00x price · xai/grok-4-1-fast-non-reasoning-latest: cache_read_price_per_mtok $0.05/Mtok to $0.20/Mtok +6 aliases2026-08-30
costgemini4.00x price · gemini-2.5-flash-preview-tts: output_price_per_mtok $2.50/Mtok to $10.00/Mtok +1 alias2026-08-28
costdeepinfra4.00x price · deepinfra/meta-llama/Meta-Llama-3.1-70B-Instruct-Turbo: input_price_per_mtok $0.10/Mtok to $0.40/Mtok2026-08-28
costqwen4.00x price · qwen/qwen3.6-27b: cache_read_price_per_mtok $0.03/Mtok to $0.12/Mtok +1 alias2026-08-20
truncationnovita-74% max out · novita/qwen/qwen2.5-7b-instruct: max_output_tokens cut 32,000 to 8,192 inferred2026-08-28
costdeepseek3.88x price · deepseek/deepseek-v4-pro: cache_read_price_per_mtok $0.03/Mtok to $0.14/Mtok2026-08-31
costdeepseek3.83x price · deepseek/deepseek-v4-pro: output_price_per_mtok $0.83/Mtok to $3.20/Mtok2026-08-31
costdeepseek3.83x price · deepseek/deepseek-v4-pro: input_price_per_mtok $0.42/Mtok to $1.60/Mtok2026-08-31
costnvidia3.53x price · nvidia/nemotron-3-super-120b-a12b: input_price_per_mtok $0.09/Mtok to $0.30/Mtok +1 alias2026-08-09
costmistral3.33x price · mistral/mistral-small-latest: output_price_per_mtok $0.18/Mtok to $0.60/Mtok2026-08-21
truncationqwen-69% max out · qwen/qwen3.5-122b-a10b: max_output_tokens cut 262,144 to 81,920 +1 alias2026-08-24
costdeepseek3.14x price · deepseek-v4-flash: input_price_per_mtok $0.14/Mtok to $0.44/Mtok +1 alias2026-08-21
costdeepseek3.03x price · deepseek-v4-pro: input_price_per_mtok $0.43/Mtok to $1.32/Mtok +1 alias2026-08-21
capabilityx-ailost frequency_penalty, presence_penalty, stop · x-ai/grok-4.3: capabilities lost frequency_penalty, presence_penalty, stop +3 aliases2026-08-18
capability~x-ailost frequency_penalty, presence_penalty, stop · ~x-ai/grok-latest: capabilities lost frequency_penalty, presence_penalty, stop inferred2026-08-18
capabilityanthropiclost max_completion_tokens, response_format, structured_outputs · anthropic/claude-opus-4.1: capabilities lost max_completion_tokens, response_format, structured_outputs2026-08-06
truncationdeepseek-67% max out · deepseek/deepseek-v4-flash: max_output_tokens cut 393,216 to 131,0722026-08-06
costdeepinfra3.00x price · deepinfra/Qwen/Qwen2.5-72B-Instruct: input_price_per_mtok $0.12/Mtok to $0.36/Mtok2026-08-28
truncationarcee-ai-66% max out · arcee-ai/trinity-large-thinking: max_output_tokens cut 235,929 to 80,000 inferred2026-08-30
truncationdeepseek-66% max out · deepseek/deepseek-v4-flash-0731: max_output_tokens cut 384,000 to 131,0722026-08-24
truncation~deepseek-66% max out · ~deepseek/deepseek-v4-flash-latest: max_output_tokens cut 384,000 to 131,072 inferred2026-08-08
truncationqwen-65% max out · qwen/qwen3.5-122b-a10b: max_output_tokens cut 235,929 to 81,920 +2 aliases2026-08-28
costtencent2.86x price · tencent/hy3-preview: output_price_per_mtok $0.21/Mtok to $0.60/Mtok2026-08-15
costtencent2.86x price · tencent/hy3-preview: input_price_per_mtok $0.06/Mtok to $0.18/Mtok2026-08-15
costtencent2.86x price · tencent/hy3-preview: cache_read_price_per_mtok $0.02/Mtok to $0.06/Mtok2026-08-15
costdeepseek2.76x price · deepseek/deepseek-v4-pro: cache_read_price_per_mtok $0.04/Mtok to $0.12/Mtok2026-08-19
costxai2.67x price · xai/grok-3-mini-beta: cache_read_price_per_mtok $0.07/Mtok to $0.20/Mtok +2 aliases2026-08-30
truncationnovita-62% max out · novita/moonshotai/kimi-k2-0905: max_output_tokens cut 262,144 to 100,352 +1 alias2026-08-28
truncationdeepseek-61% ctx · deepseek/deepseek-r1: context_tokens cut 163,840 to 64,0002026-08-13
truncationnovita-60% max out · novita/deepseek/deepseek-v3-0324: max_output_tokens cut 163,840 to 65,536 inferred2026-08-28
truncationdeepseek-60% max out · deepseek/deepseek-v3.2: max_output_tokens cut 163,840 to 65,536 +1 alias2026-08-09
costmistral2.50x price · mistral/mistral-small-latest: input_price_per_mtok $0.06/Mtok to $0.15/Mtok2026-08-21
costz-ai2.50x price · z-ai/glm-5.2: output_price_per_mtok $0.97/Mtok to $2.42/Mtok2026-08-17
costz-ai2.47x price · z-ai/glm-5.2: input_price_per_mtok $0.31/Mtok to $0.76/Mtok2026-08-17
truncationdeepseek-59% max out · deepseek/deepseek-v4-pro-0813: max_output_tokens cut 943,717 to 384,0002026-08-28
costqwen2.45x price · qwen/qwen3-vl-235b-a22b-thinking: input_price_per_mtok $0.40/Mtok to $0.98/Mtok +1 alias2026-08-06
costz-ai2.45x price · z-ai/glm-5.2: cache_read_price_per_mtok $0.06/Mtok to $0.14/Mtok2026-08-17
truncationnovita-59% max out · novita/qwen/qwen3-4b-fp8: max_output_tokens cut 20,000 to 8,192 inferred2026-08-28
costz-ai2.43x price · z-ai/glm-5.2: cache_read_price_per_mtok $0.09/Mtok to $0.22/Mtok2026-08-18
costz-ai2.43x price · z-ai/glm-5.2: output_price_per_mtok $1.54/Mtok to $3.74/Mtok2026-08-18
costz-ai2.43x price · z-ai/glm-5.2: input_price_per_mtok $0.49/Mtok to $1.19/Mtok2026-08-18
costz-ai2.34x price · z-ai/glm-5.2: cache_read_price_per_mtok $0.09/Mtok to $0.22/Mtok2026-08-14
costdeepinfra2.33x price · deepinfra/NousResearch/Hermes-3-Llama-3.1-70B: input_price_per_mtok $0.30/Mtok to $0.70/Mtok2026-08-28
costdeepinfra2.33x price · deepinfra/NousResearch/Hermes-3-Llama-3.1-70B: output_price_per_mtok $0.30/Mtok to $0.70/Mtok2026-08-28
costdeepseek2.28x price · deepseek/deepseek-v4-pro-0813: output_price_per_mtok $0.87/Mtok to $1.98/Mtok2026-08-16
truncationdeepseek-56% max out · deepseek/deepseek-v3.2: max_output_tokens cut 147,456 to 65,536 +2 aliases2026-08-30
costnvidia2.25x price · nvidia/nemotron-3-super-120b-a12b: output_price_per_mtok $0.40/Mtok to $0.90/Mtok +1 alias2026-08-09
truncationsao10k-55% max out · sao10k/l3-lunaris-8b: max_output_tokens cut 16,384 to 7,372 inferred2026-08-25
costmeta-llama2.22x price · meta-llama/llama-3.3-70b-instruct: output_price_per_mtok $0.32/Mtok to $0.71/Mtok2026-08-26
costdeepseek2.13x price · deepseek/deepseek-v4-flash-0731: input_price_per_mtok $0.07/Mtok to $0.14/Mtok2026-08-25
costdeepseek2.13x price · deepseek/deepseek-v4-flash-0731: output_price_per_mtok $0.13/Mtok to $0.28/Mtok2026-08-25
costdeepseek2.13x price · deepseek/deepseek-v4-flash-0731: cache_read_price_per_mtok $0.01/Mtok to $0.03/Mtok2026-08-25
costz-ai2.11x price · z-ai/glm-5.2: output_price_per_mtok $1.50/Mtok to $3.15/Mtok2026-08-18
costxai2.08x price · xai/grok-3-mini-fast-beta: input_price_per_mtok $0.60/Mtok to $1.25/Mtok +2 aliases2026-08-30
costqwen2.08x price · qwen/qwen3.6-27b: input_price_per_mtok $0.29/Mtok to $0.60/Mtok +1 alias2026-08-18
truncationz-ai-51% max out · z-ai/glm-5.2: max_output_tokens cut 262,144 to 128,0002026-08-06
costz-ai2.05x price · z-ai/glm-5.2: output_price_per_mtok $1.54/Mtok to $3.15/Mtok2026-08-13
capabilitykwaipilotlost logprobs, top_logprobs · kwaipilot/kat-coder-pro-v2.5: capabilities lost logprobs, top_logprobs +1 alias inferred2026-08-31
capabilityz-ailost logprobs, top_logprobs · z-ai/glm-4.7: capabilities lost logprobs, top_logprobs2026-08-31
capabilityxailost function_calling, tool_choice · xai/grok-4.20-multi-agent-0309: capabilities lost function_calling, tool_choice inferred2026-08-30
truncationqwen-50% max out · qwen/qwen3-235b-a22b-2507: max_output_tokens cut 32,768 to 16,384 +2 aliases2026-08-28
truncationgoogle-50% ctx · google/gemma-3-27b-it: context_tokens cut 262,144 to 131,0722026-08-28
capabilityqwenlost logit_bias, min_p · qwen/qwen3-235b-a22b-thinking-2507: capabilities lost logit_bias, min_p2026-08-24
truncationqwen-50% ctx · qwen/qwen3-235b-a22b-thinking-2507: context_tokens cut 262,144 to 131,0722026-08-24
truncationmistral-50% max out · mistral/codestral-2508: max_output_tokens cut 256,000 to 128,0002026-08-21
truncationmistral-50% ctx · mistral/codestral-2508: context_tokens cut 256,000 to 128,0002026-08-21
truncationqwen-50% max out · qwen/qwen3.6-27b: max_output_tokens cut 131,072 to 65,5362026-08-19
truncationnvidia-50% max out · nvidia/nemotron-3.5-lightning: max_output_tokens cut 262,144 to 131,0722026-08-16
truncationxai-50% max out · xai/grok-4.20-0309-reasoning: max_output_tokens cut 2,000,000 to 1,000,000 +3 aliases inferred2026-08-12
truncationxai-50% ctx · xai/grok-4.20-0309-reasoning: context_tokens cut 2,000,000 to 1,000,000 +3 aliases inferred2026-08-12
costdeepseek2.00x price · deepseek/deepseek-v4-pro-0813: cache_read_price_per_mtok $0.02/Mtok to $0.04/Mtok +5 aliases2026-08-31
costdeepseek2.00x price · deepseek/deepseek-v4-pro-0813: output_price_per_mtok $1.98/Mtok to $3.96/Mtok +5 aliases2026-08-31
costdeepseek2.00x price · deepseek/deepseek-v4-pro-0813: input_price_per_mtok $0.66/Mtok to $1.32/Mtok +5 aliases2026-08-31
costdeepseek2.00x price · deepseek/deepseek-v4-flash-0731: output_price_per_mtok $0.09/Mtok to $0.18/Mtok2026-08-30
cost~google2.00x price · ~google/gemini-flash-latest: cache_read_price_per_mtok $0.04/Mtok to $0.07/Mtok2026-08-28
costgoogle2.00x price · google/gemini-3.7-flash: output_price_per_mtok $1.88/Mtok to $3.75/Mtok2026-08-28
cost~google2.00x price · ~google/gemini-flash-latest: input_price_per_mtok $0.38/Mtok to $0.75/Mtok2026-08-28
cost~google2.00x price · ~google/gemini-flash-latest: output_price_per_mtok $1.88/Mtok to $3.75/Mtok2026-08-28
costgoogle2.00x price · google/gemini-3.7-flash: cache_read_price_per_mtok $0.04/Mtok to $0.07/Mtok2026-08-28
costgoogle2.00x price · google/gemini-3.7-flash: input_price_per_mtok $0.38/Mtok to $0.75/Mtok2026-08-28
costdeepinfra2.00x price · deepinfra/Qwen/Qwen3-14B: input_price_per_mtok $0.06/Mtok to $0.12/Mtok2026-08-28
costgemini2.00x price · gemini/gemini-2.5-pro-preview-tts: output_price_per_mtok $10.00/Mtok to $20.00/Mtok2026-08-28
costtogether_ai2.00x price · together_ai/Qwen/Qwen3.7-Max: output_price_per_mtok $3.75/Mtok to $7.50/Mtok2026-08-28
costtogether_ai2.00x price · together_ai/Qwen/Qwen3.7-Max: input_price_per_mtok $1.25/Mtok to $2.50/Mtok2026-08-28
costvertex_ai-language-models2.00x price · gemini-2.5-pro-preview-tts: output_price_per_mtok $10.00/Mtok to $20.00/Mtok2026-08-28
costgemini2.00x price · gemini/gemini-3.1-flash-image-preview: input_price_per_mtok $0.25/Mtok to $0.50/Mtok +1 alias2026-08-21
costgemini2.00x price · gemini/gemini-3.1-flash-image-preview: output_price_per_mtok $1.50/Mtok to $3.00/Mtok +1 alias2026-08-21
costgoogle2.00x price · google/gemma-4-31b-it: cache_read_price_per_mtok $0.05/Mtok to $0.10/Mtok2026-08-21
costqwen2.00x price · qwen/qwen3.6-27b: input_price_per_mtok $0.30/Mtok to $0.60/Mtok +1 alias2026-08-20
costminimax2.00x price · minimax/minimax-m3:batch: input_price_per_mtok $0.15/Mtok to $0.30/Mtok2026-08-18
costnvidia2.00x price · nvidia/nemotron-3-ultra-550b-a55b:batch: input_price_per_mtok $0.30/Mtok to $0.60/Mtok2026-08-18
costz-ai2.00x price · z-ai/glm-5.2:batch: input_price_per_mtok $0.70/Mtok to $1.40/Mtok2026-08-18
costmoonshotai2.00x price · moonshotai/kimi-k2.7-code:batch: cache_read_price_per_mtok $0.10/Mtok to $0.19/Mtok2026-08-18
costz-ai2.00x price · z-ai/glm-5.2:batch: output_price_per_mtok $2.20/Mtok to $4.40/Mtok2026-08-18
costthinkingmachines2.00x price · thinkingmachines/inkling:batch: cache_read_price_per_mtok $0.09/Mtok to $0.17/Mtok2026-08-18
costminimax2.00x price · minimax/minimax-m3:batch: output_price_per_mtok $0.60/Mtok to $1.20/Mtok2026-08-18
costmoonshotai2.00x price · moonshotai/kimi-k2.7-code:batch: output_price_per_mtok $2.00/Mtok to $4.00/Mtok2026-08-18
costminimax2.00x price · minimax/minimax-m3:batch: cache_read_price_per_mtok $0.03/Mtok to $0.06/Mtok2026-08-18
costthinkingmachines2.00x price · thinkingmachines/inkling:batch: output_price_per_mtok $2.02/Mtok to $4.05/Mtok2026-08-18
costthinkingmachines2.00x price · thinkingmachines/inkling:batch: input_price_per_mtok $0.50/Mtok to $1.00/Mtok2026-08-18
costmoonshotai2.00x price · moonshotai/kimi-k2.7-code:batch: input_price_per_mtok $0.47/Mtok to $0.95/Mtok2026-08-18
costz-ai2.00x price · z-ai/glm-5.2:batch: cache_read_price_per_mtok $0.13/Mtok to $0.26/Mtok2026-08-18
costnvidia2.00x price · nvidia/nemotron-3-ultra-550b-a55b:batch: output_price_per_mtok $1.80/Mtok to $3.60/Mtok2026-08-18
costnvidia2.00x price · nvidia/nemotron-3-ultra-550b-a55b: cache_read_price_per_mtok $0.10/Mtok to $0.20/Mtok +1 alias2026-08-18
costopenai2.00x price · openai/gpt-5.6-luna-pro: cache_read_price_per_mtok $0.01/Mtok to $0.02/Mtok +1 alias2026-08-17
costopenai2.00x price · openai/gpt-5.6-terra-pro: output_price_per_mtok $6.00/Mtok to $12.00/Mtok +1 alias2026-08-17
costopenai2.00x price · openai/gpt-5.6-terra-pro: input_price_per_mtok $1.00/Mtok to $2.00/Mtok +1 alias2026-08-17
costopenai2.00x price · openai/gpt-5.6-terra-pro: cache_read_price_per_mtok $0.10/Mtok to $0.20/Mtok +1 alias2026-08-17
costopenai2.00x price · openai/gpt-5.6-luna-pro: input_price_per_mtok $0.10/Mtok to $0.20/Mtok +1 alias2026-08-17
costopenai2.00x price · openai/gpt-5.6-luna-pro: output_price_per_mtok $0.60/Mtok to $1.20/Mtok +1 alias2026-08-17
truncationnvidia-49% ctx · nvidia/nemotron-3-ultra-550b-a55b: context_tokens cut 512,288 to 262,1442026-08-28
truncationmistralai-49% ctx · mistralai/mistral-small-3.2-24b-instruct: context_tokens cut 256,000 to 131,0722026-08-22
truncationliquid-49% ctx · liquid/lfm-2.5-2.6b:free: context_tokens cut 128,000 to 65,536 inferred2026-08-21
costz-ai1.89x price · z-ai/glm-5.2: output_price_per_mtok $1.98/Mtok to $3.74/Mtok2026-08-14
costz-ai1.89x price · z-ai/glm-5.2: input_price_per_mtok $0.63/Mtok to $1.19/Mtok2026-08-14
costqwen1.88x price · qwen/qwen3.6-27b: input_price_per_mtok $0.32/Mtok to $0.60/Mtok +2 aliases2026-08-28
costdeepseek1.85x price · deepseek/deepseek-v4-pro: cache_read_price_per_mtok $0.05/Mtok to $0.10/Mtok2026-08-12
costdeepseek1.85x price · deepseek/deepseek-v4-pro: input_price_per_mtok $0.63/Mtok to $1.17/Mtok2026-08-12
costdeepseek1.85x price · deepseek/deepseek-v4-pro: output_price_per_mtok $1.26/Mtok to $2.34/Mtok2026-08-12
costgryphe1.83x price · gryphe/mythomax-l2-13b: output_price_per_mtok $0.06/Mtok to $0.11/Mtok2026-08-04
costqwen1.83x price · qwen/qwen3-vl-235b-a22b-instruct: output_price_per_mtok $1.04/Mtok to $1.90/Mtok2026-08-17
costdeepseek1.80x price · deepseek/deepseek-v4-pro-0813: cache_read_price_per_mtok $0.02/Mtok to $0.04/Mtok2026-08-20
costqwen1.80x price · qwen/qwen3.6-27b: output_price_per_mtok $2.00/Mtok to $3.60/Mtok +1 alias2026-08-20
costdeepseek1.80x price · deepseek/deepseek-v4-pro-0813: output_price_per_mtok $1.98/Mtok to $3.56/Mtok2026-08-20
costdeepseek1.80x price · deepseek/deepseek-v4-pro-0813: input_price_per_mtok $0.66/Mtok to $1.19/Mtok2026-08-20
truncationnvidia-44% max out · nvidia/nemotron-3.5-lightning: max_output_tokens cut 235,929 to 131,0722026-08-28
costz-ai1.79x price · z-ai/glm-5.2: output_price_per_mtok $1.76/Mtok to $3.15/Mtok2026-08-12
costqwen1.79x price · qwen/qwen3.5-35b-a3b: input_price_per_mtok $0.14/Mtok to $0.25/Mtok2026-08-13
costdeepseek1.78x price · deepseek/deepseek-v4-flash-0731: cache_read_price_per_mtok $0.0090/Mtok to $0.02/Mtok2026-08-30
costmoonshotai1.75x price · moonshotai/kimi-k2.6: output_price_per_mtok $2.28/Mtok to $4.00/Mtok +4 aliases2026-08-24
costmoonshotai1.75x price · moonshotai/kimi-k2.6: input_price_per_mtok $0.54/Mtok to $0.95/Mtok +4 aliases2026-08-24
costmoonshotai1.75x price · moonshotai/kimi-k2.6: cache_read_price_per_mtok $0.09/Mtok to $0.16/Mtok +4 aliases2026-08-24
costdeepseek1.75x price · deepseek/deepseek-v4-flash-0731: input_price_per_mtok $0.08/Mtok to $0.14/Mtok +1 alias2026-08-24
cost~deepseek1.75x price · ~deepseek/deepseek-v4-flash-latest: output_price_per_mtok $0.14/Mtok to $0.25/Mtok +1 alias2026-08-14
costdeepseek1.75x price · deepseek/deepseek-v4-flash-0731: cache_read_price_per_mtok $0.02/Mtok to $0.03/Mtok +1 alias2026-08-24
cost~deepseek1.75x price · ~deepseek/deepseek-v4-flash-latest: cache_read_price_per_mtok $0.01/Mtok to $0.03/Mtok2026-08-12
costdeepinfra1.72x price · deepinfra/Qwen/Qwen3-30B-A3B: output_price_per_mtok $0.29/Mtok to $0.50/Mtok2026-08-28
costgoogle1.71x price · google/gemma-4-26b-a4b-it: input_price_per_mtok $0.07/Mtok to $0.12/Mtok2026-08-10
costqwen1.70x price · qwen/qwen3.8-27b: cache_read_price_per_mtok $0.05/Mtok to $0.09/Mtok2026-08-25
costdeepseek1.70x price · deepseek/deepseek-v4-pro: output_price_per_mtok $2.34/Mtok to $3.96/Mtok2026-08-17
costmoonshotai1.69x price · moonshotai/kimi-k2.6: output_price_per_mtok $2.36/Mtok to $4.00/Mtok +2 aliases2026-08-19
costmoonshotai1.69x price · moonshotai/kimi-k2.6: cache_read_price_per_mtok $0.09/Mtok to $0.16/Mtok +2 aliases2026-08-19
costmoonshotai1.69x price · moonshotai/kimi-k2.6: input_price_per_mtok $0.56/Mtok to $0.95/Mtok +2 aliases2026-08-19
cost~deepseek1.67x price · ~deepseek/deepseek-v4-flash-latest: input_price_per_mtok $0.03/Mtok to $0.05/Mtok2026-08-31
costgemini1.67x price · gemini-2.5-flash-native-audio-latest: input_price_per_mtok $0.30/Mtok to $0.50/Mtok +8 aliases2026-08-28
costvertex_ai-language-models1.67x price · gemini-live-2.5-flash-preview-native-audio-09-2025: input_price_per_mtok $0.30/Mtok to $0.50/Mtok2026-08-28
cost~x-ai1.67x price · ~x-ai/grok-latest: cache_read_price_per_mtok $0.30/Mtok to $0.50/Mtok2026-08-12
costqwen1.66x price · qwen/qwen3-235b-a22b-2507: input_price_per_mtok $0.09/Mtok to $0.15/Mtok2026-08-03
costmoonshotai1.64x price · moonshotai/kimi-k2.6: cache_read_price_per_mtok $0.09/Mtok to $0.15/Mtok2026-08-16
costmoonshotai1.64x price · moonshotai/kimi-k2.6: output_price_per_mtok $2.44/Mtok to $4.00/Mtok +3 aliases2026-08-13
costmoonshotai1.64x price · moonshotai/kimi-k2.6: input_price_per_mtok $0.58/Mtok to $0.95/Mtok +3 aliases2026-08-13
costmoonshotai1.64x price · moonshotai/kimi-k2.6: cache_read_price_per_mtok $0.10/Mtok to $0.16/Mtok +3 aliases2026-08-13
costnvidia1.64x price · nvidia/nemotron-3-ultra-550b-a55b: output_price_per_mtok $2.20/Mtok to $3.60/Mtok2026-08-01
costz-ai1.61x price · z-ai/glm-5.2: output_price_per_mtok $1.23/Mtok to $1.98/Mtok2026-08-14
costz-ai1.61x price · z-ai/glm-5.2: input_price_per_mtok $0.39/Mtok to $0.63/Mtok2026-08-14
costdeepinfra1.60x price · deepinfra/nvidia/NVIDIA-Nemotron-3.5-Lightning: input_price_per_mtok $0.05/Mtok to $0.08/Mtok2026-08-28
costmoonshotai1.59x price · moonshotai/kimi-k2.6: cache_read_price_per_mtok $0.09/Mtok to $0.15/Mtok2026-08-15
costdeepseek1.59x price · deepseek/deepseek-v4-flash: output_price_per_mtok $0.18/Mtok to $0.28/Mtok2026-08-07
costdeepseek1.59x price · deepseek/deepseek-v4-flash: cache_read_price_per_mtok $0.02/Mtok to $0.03/Mtok2026-08-07
costdeepseek1.59x price · deepseek/deepseek-v4-flash: input_price_per_mtok $0.09/Mtok to $0.14/Mtok2026-08-07
costdeepseek1.58x price · deepseek/deepseek-v4-flash: input_price_per_mtok $0.06/Mtok to $0.09/Mtok2026-08-25
costdeepseek1.58x price · deepseek/deepseek-v4-flash: output_price_per_mtok $0.11/Mtok to $0.18/Mtok2026-08-25
costdeepseek1.58x price · deepseek/deepseek-v4-flash: cache_read_price_per_mtok $0.01/Mtok to $0.02/Mtok2026-08-25
cost~deepseek1.58x price · ~deepseek/deepseek-v4-flash-latest: cache_read_price_per_mtok $0.02/Mtok to $0.03/Mtok2026-08-11
cost~deepseek1.58x price · ~deepseek/deepseek-v4-flash-latest: output_price_per_mtok $0.16/Mtok to $0.25/Mtok2026-08-11
costz-ai1.58x price · z-ai/glm-5.2: cache_read_price_per_mtok $0.14/Mtok to $0.22/Mtok2026-08-17
cost~deepseek1.57x price · ~deepseek/deepseek-v4-flash-latest: cache_read_price_per_mtok $0.02/Mtok to $0.03/Mtok2026-08-14
costz-ai1.57x price · z-ai/glm-5.2: input_price_per_mtok $0.76/Mtok to $1.19/Mtok2026-08-17
costdeepseek1.56x price · deepseek/deepseek-v4-flash-0731: output_price_per_mtok $0.18/Mtok to $0.28/Mtok +1 alias2026-08-24
costz-ai1.55x price · z-ai/glm-5.2: output_price_per_mtok $2.42/Mtok to $3.74/Mtok2026-08-17
costqwen1.54x price · qwen/qwen3.5-397b-a17b: output_price_per_mtok $2.34/Mtok to $3.60/Mtok +2 aliases2026-08-24
costqwen1.54x price · qwen/qwen3.5-122b-a10b: input_price_per_mtok $0.26/Mtok to $0.40/Mtok2026-08-03
costqwen1.54x price · qwen/qwen3.5-122b-a10b: output_price_per_mtok $2.08/Mtok to $3.20/Mtok2026-08-03
costdeepseek1.52x price · deepseek/deepseek-v4-pro-0813: input_price_per_mtok $0.43/Mtok to $0.66/Mtok2026-08-16
costdeepseek1.51x price · deepseek/deepseek-v4-pro: cache_read_price_per_mtok $0.04/Mtok to $0.07/Mtok2026-08-25
costdeepseek1.51x price · deepseek/deepseek-v4-pro: input_price_per_mtok $0.52/Mtok to $0.79/Mtok2026-08-25
costdeepseek1.51x price · deepseek/deepseek-v4-pro: output_price_per_mtok $1.04/Mtok to $1.58/Mtok2026-08-25
truncationundi95-33% max out · undi95/remm-slerp-l2-13b: max_output_tokens cut 6,144 to 4,0962026-08-24
costbedrock_mantle1.50x price · bedrock_mantle/openai.gpt-5.6-sol: output_price_per_mtok $22.00/Mtok to $33.00/Mtok2026-08-30
costdeepinfra1.50x price · deepinfra/Qwen/Qwen3-30B-A3B: input_price_per_mtok $0.08/Mtok to $0.12/Mtok2026-08-28
cost~deepseek1.50x price · ~deepseek/deepseek-v4-flash-latest: output_price_per_mtok $0.08/Mtok to $0.12/Mtok2026-08-25
costqwen1.50x price · qwen/qwen3.6-27b: output_price_per_mtok $2.40/Mtok to $3.60/Mtok +1 alias2026-08-18
costdeepinfra1.50x price · deepinfra/google/gemma-3-12b-it: output_price_per_mtok $0.10/Mtok to $0.15/Mtok2026-08-28
costmoonshotai1.50x price · moonshotai/kimi-k2.6: output_price_per_mtok $2.28/Mtok to $3.41/Mtok2026-08-16
costdeepseek1.48x price · deepseek/deepseek-v4-pro: input_price_per_mtok $0.43/Mtok to $0.64/Mtok2026-08-11
costdeepseek1.48x price · deepseek/deepseek-v4-pro: output_price_per_mtok $0.87/Mtok to $1.29/Mtok2026-08-11
costz-ai1.47x price · z-ai/glm-5.1: output_price_per_mtok $2.99/Mtok to $4.40/Mtok +3 aliases2026-08-14
costz-ai1.47x price · z-ai/glm-5.1: input_price_per_mtok $0.95/Mtok to $1.40/Mtok +3 aliases2026-08-14
costz-ai1.47x price · z-ai/glm-5.1: cache_read_price_per_mtok $0.18/Mtok to $0.26/Mtok +3 aliases2026-08-14
costminimax1.47x price · minimax/minimax-m2.5: input_price_per_mtok $0.15/Mtok to $0.22/Mtok2026-08-05
truncation~deepseek-32% max out · ~deepseek/deepseek-v4-flash-latest: max_output_tokens cut 384,000 to 262,144 +3 aliases inferred2026-08-19
costz-ai1.45x price · z-ai/glm-5.2: output_price_per_mtok $1.67/Mtok to $2.42/Mtok2026-08-11
costmoonshotai1.44x price · moonshotai/kimi-k2.6: output_price_per_mtok $2.36/Mtok to $3.41/Mtok2026-08-15
costdeepseek1.44x price · deepseek/deepseek-v4-flash-0731: input_price_per_mtok $0.04/Mtok to $0.07/Mtok2026-08-30
costqwen1.44x price · qwen/qwen3.5-35b-a3b: output_price_per_mtok $1.25/Mtok to $1.80/Mtok +1 alias2026-08-27
costz-ai1.43x price · z-ai/glm-5.2: input_price_per_mtok $0.53/Mtok to $0.76/Mtok2026-08-11
cost~deepseek1.43x price · ~deepseek/deepseek-v4-flash-latest: cache_read_price_per_mtok $0.0070/Mtok to $0.01/Mtok2026-08-30
costdeepinfra1.43x price · deepinfra/meta-llama/Meta-Llama-3.1-70B-Instruct-Turbo: output_price_per_mtok $0.28/Mtok to $0.40/Mtok2026-08-28
costmoonshotai1.43x price · moonshotai/kimi-k2.5: cache_read_price_per_mtok $0.07/Mtok to $0.10/Mtok +1 alias2026-08-27
cost~deepseek1.43x price · ~deepseek/deepseek-v4-flash-latest: cache_read_price_per_mtok $0.01/Mtok to $0.02/Mtok2026-08-21
truncationz-ai-30% max out · z-ai/glm-5.1: max_output_tokens cut 182,476 to 128,0002026-08-30
costz-ai1.42x price · z-ai/glm-5.2: cache_read_price_per_mtok $0.10/Mtok to $0.14/Mtok2026-08-11
cost~deepseek1.41x price · ~deepseek/deepseek-v4-flash-latest: cache_read_price_per_mtok $0.02/Mtok to $0.03/Mtok2026-08-09
cost~deepseek1.41x price · ~deepseek/deepseek-v4-flash-latest: output_price_per_mtok $0.18/Mtok to $0.25/Mtok2026-08-09
cost~deepseek1.40x price · ~deepseek/deepseek-v4-flash-latest: output_price_per_mtok $0.10/Mtok to $0.14/Mtok2026-08-30
cost~deepseek1.40x price · ~deepseek/deepseek-v4-flash-latest: cache_read_price_per_mtok $0.02/Mtok to $0.03/Mtok2026-08-23
costdeepseek1.38x price · deepseek/deepseek-v4-pro: output_price_per_mtok $1.15/Mtok to $1.58/Mtok2026-08-26
costdeepseek1.38x price · deepseek/deepseek-v4-pro: cache_read_price_per_mtok $0.05/Mtok to $0.07/Mtok2026-08-26
costdeepseek1.38x price · deepseek/deepseek-v4-pro: input_price_per_mtok $0.57/Mtok to $0.79/Mtok2026-08-26
costdeepseek1.36x price · deepseek/deepseek-v4-flash: input_price_per_mtok $0.06/Mtok to $0.08/Mtok2026-08-17
costdeepseek1.36x price · deepseek/deepseek-v4-flash: output_price_per_mtok $0.12/Mtok to $0.17/Mtok2026-08-17
costdeepseek1.36x price · deepseek/deepseek-v4-flash: cache_read_price_per_mtok $0.01/Mtok to $0.02/Mtok2026-08-17
truncationminimax-26% max out · minimax/minimax-m2.7: max_output_tokens cut 176,947 to 131,0722026-08-25
costxai1.33x price · xai/grok-3-mini-fast-beta: cache_read_price_per_mtok $0.15/Mtok to $0.20/Mtok +2 aliases2026-08-30
costdeepinfra1.33x price · deepinfra/meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8: input_price_per_mtok $0.15/Mtok to $0.20/Mtok2026-08-28
costdeepinfra1.33x price · deepinfra/meta-llama/Meta-Llama-3.1-8B-Instruct-Turbo: output_price_per_mtok $0.03/Mtok to $0.04/Mtok2026-08-28
costdeepinfra1.33x price · deepinfra/meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8: output_price_per_mtok $0.60/Mtok to $0.80/Mtok2026-08-28
cost~deepseek1.33x price · ~deepseek/deepseek-v4-flash-latest: output_price_per_mtok $0.07/Mtok to $0.10/Mtok2026-08-27
costgryphe1.33x price · gryphe/mythomax-l2-13b: input_price_per_mtok $0.06/Mtok to $0.08/Mtok2026-08-04
costmoonshotai1.33x price · moonshotai/kimi-k2.5: input_price_per_mtok $0.45/Mtok to $0.60/Mtok +1 alias2026-08-27
costmoonshotai1.33x price · moonshotai/kimi-k2.5: output_price_per_mtok $2.25/Mtok to $3.00/Mtok +1 alias2026-08-27
costxai1.33x price · xai/grok-code-fast-1-0825: output_price_per_mtok $1.50/Mtok to $2.00/Mtok +2 aliases2026-08-12
costdeepseek1.31x price · deepseek/deepseek-v4-pro: cache_read_price_per_mtok $0.03/Mtok to $0.04/Mtok2026-08-24
costdeepseek1.31x price · deepseek/deepseek-v4-pro: output_price_per_mtok $0.79/Mtok to $1.04/Mtok2026-08-24
costdeepseek1.31x price · deepseek/deepseek-v4-pro: input_price_per_mtok $0.40/Mtok to $0.52/Mtok2026-08-24
costdeepinfra1.31x price · deepinfra/Sao10K/L3.1-70B-Euryale-v2.2: input_price_per_mtok $0.65/Mtok to $0.85/Mtok2026-08-28
cost~deepseek1.31x price · ~deepseek/deepseek-v4-flash-latest: cache_read_price_per_mtok $0.02/Mtok to $0.02/Mtok2026-08-20
truncationnovita-23% max out · novita/moonshotai/kimi-k2-instruct: max_output_tokens cut 131,072 to 100,352 inferred2026-08-28
costz-ai1.30x price · z-ai/glm-5.1: output_price_per_mtok $3.04/Mtok to $3.96/Mtok2026-08-25
costz-ai1.30x price · z-ai/glm-5.1: input_price_per_mtok $0.97/Mtok to $1.26/Mtok2026-08-25
costz-ai1.30x price · z-ai/glm-5.1: cache_read_price_per_mtok $0.18/Mtok to $0.23/Mtok2026-08-25
cost~deepseek1.30x price · ~deepseek/deepseek-v4-flash-latest: input_price_per_mtok $0.06/Mtok to $0.08/Mtok2026-08-17
cost~deepseek1.30x price · ~deepseek/deepseek-v4-flash-latest: output_price_per_mtok $0.12/Mtok to $0.16/Mtok2026-08-17
cost~deepseek1.30x price · ~deepseek/deepseek-v4-flash-latest: cache_read_price_per_mtok $0.01/Mtok to $0.02/Mtok2026-08-17
costz-ai1.30x price · z-ai/glm-5.2: cache_read_price_per_mtok $0.09/Mtok to $0.12/Mtok2026-08-18
cost~deepseek1.30x price · ~deepseek/deepseek-v4-flash-latest: cache_read_price_per_mtok $0.01/Mtok to $0.01/Mtok2026-08-31
costz-ai1.30x price · z-ai/glm-5.2: cache_read_price_per_mtok $0.07/Mtok to $0.09/Mtok2026-08-14
cost~deepseek1.29x price · ~deepseek/deepseek-v4-flash-latest: output_price_per_mtok $0.14/Mtok to $0.18/Mtok2026-08-21
costqwen1.28x price · qwen/qwen3.5-397b-a17b: input_price_per_mtok $0.39/Mtok to $0.50/Mtok +2 aliases2026-08-24
costnovita1.27x price · novita/qwen/qwen3-coder-480b-a35b-instruct: input_price_per_mtok $0.30/Mtok to $0.38/Mtok2026-08-28
truncationthedrummer-20% max out · thedrummer/unslopnemo-12b: max_output_tokens cut 32,768 to 26,214 inferred2026-08-25
truncationmistralai-20% max out · mistralai/mistral-small-3.1-24b-instruct: max_output_tokens cut 128,000 to 102,4002026-08-25
truncationnvidia-13% max out · nvidia/nemotron-3-nano-30b-a3b: max_output_tokens cut 262,144 to 228,000 inferred2026-08-10
truncationazure-12% ctx · azure/eu/gpt-5.6-luna: context_tokens cut 1,050,000 to 922,000 +11 aliases2026-08-21
truncationopenai-12% ctx · gpt-5.6-luna: context_tokens cut 1,050,000 to 922,000 +3 aliases2026-08-21
truncationopenai-10% max out · openai/gpt-3.5-turbo-0613: max_output_tokens cut 4,096 to 3,685 +1 alias2026-08-25
truncationgryphe-10% max out · gryphe/mythomax-l2-13b: max_output_tokens cut 4,096 to 3,6862026-08-25
truncationdeepseek-10% max out · deepseek/deepseek-r1-distill-llama-70b: max_output_tokens cut 8,192 to 7,3722026-08-25
truncationmicrosoft-10% max out · microsoft/phi-4: max_output_tokens cut 16,384 to 14,7452026-08-25
truncationrekaai-10% max out · rekaai/reka-edge: max_output_tokens cut 16,384 to 14,745 inferred2026-08-25
truncationz-ai-10% max out · z-ai/glm-5.1: max_output_tokens cut 202,752 to 182,4762026-08-25
truncationminimax-10% max out · minimax/minimax-01: max_output_tokens cut 1,000,192 to 900,172 inferred2026-08-25
truncationstepfun-10% max out · stepfun/step-3.7-flash: max_output_tokens cut 256,000 to 230,400 inferred2026-08-25
truncationdots-studio-10% max out · dots-studio/dots-3-note-preview:free: max_output_tokens cut 512,000 to 460,800 inferred2026-08-25
truncationz-ai-10% max out · z-ai/glm-5.2:free: max_output_tokens cut 256,000 to 230,400 inferred2026-08-25
truncationibm-granite-10% max out · ibm-granite/granite-4.0-h-micro: max_output_tokens cut 131,000 to 117,9002026-08-25
truncationdeepseek-10% max out · deepseek/deepseek-chat-v3-0324: max_output_tokens cut 163,840 to 147,456 +1 alias2026-08-25
truncationmeta-llama-10% max out · meta-llama/llama-3.2-1b-instruct: max_output_tokens cut 60,000 to 54,0002026-08-25
truncationqwen-10% max out · qwen/qwen2.5-vl-72b-instruct: max_output_tokens cut 128,000 to 115,2002026-08-25
truncationdeepseek-10% max out · deepseek/deepseek-chat-v3.1: max_output_tokens cut 161,000 to 144,9002026-08-25
truncation~moonshotai-7% max out · ~moonshotai/kimi-latest: max_output_tokens cut 943,718 to 877,3572026-08-25
truncation~moonshotai-7% max out · ~moonshotai/kimi-latest: max_output_tokens cut 1,048,576 to 974,842 +1 alias2026-08-17
truncationnvidia-5% ctx · nvidia/nemotron-3.5-lightning: context_tokens cut 1,048,576 to 1,000,0002026-08-14
truncationnvidia-3% max out · nvidia/nemotron-3-nano-30b-a3b: max_output_tokens cut 235,929 to 228,000 inferred2026-08-27
truncationgoogle-1% max out · google/gemini-3.1-flash-lite-image: max_output_tokens cut 66,000 to 65,5362026-08-19
capabilitybedrock_converselost prompt_caching · global.xai.grok-4.6: capabilities lost prompt_caching +1 alias2026-08-30
capabilityqwenlost logit_bias · qwen/qwen-2.5-7b-instruct: capabilities lost logit_bias +1 alias2026-08-27
capabilitythinkingmachineslost structured_outputs · thinkingmachines/inkling-small: capabilities lost structured_outputs inferred2026-08-23
capabilitygeminilost reasoning · gemini/gemini-3.1-flash-lite-image: capabilities lost reasoning2026-08-23
capabilityvertex_ai-language-modelslost reasoning · gemini-3.1-flash-lite-image: capabilities lost reasoning; gained pdf_input, video_input +1 alias2026-08-23
capabilityminimaxlost parallel_tool_calls · minimax/minimax-m2.5: capabilities lost parallel_tool_calls2026-08-19
capabilityinclusionailost structured_outputs · inclusionai/ling-3.0-flash: capabilities lost structured_outputs inferred2026-08-19
capabilityopenrouterlost min_p · openrouter/free: capabilities lost min_p2026-08-18
capabilityopenrouterlost max_completion_tokens · openrouter/free: capabilities lost max_completion_tokens2026-08-18
capabilitygooglelost top_a · google/gemma-4-31b-it: capabilities lost top_a2026-08-15
capabilitydeepseeklost max_completion_tokens · deepseek/deepseek-r1: capabilities lost max_completion_tokens2026-08-13
capabilitykwaipilotlost frequency_penalty · kwaipilot/kat-coder-air-v2.5: capabilities lost frequency_penalty +1 alias inferred2026-08-13
capabilityopenailost max_tokens · openai/gpt-5.2-chat: capabilities lost max_tokens2026-08-10
capabilityopenailost max_completion_tokens · openai/gpt-5.3-chat: capabilities lost max_completion_tokens2026-08-08
availabilitykwaipilotkwaipilot/kat-coder-air-v2.5 is no longer listed by this source — calls to it are expected to fail2026-08-31
availabilityarcee-aiarcee-ai/virtuoso-large is no longer listed by this source — calls to it are expected to fail2026-08-29
availabilityopenaiopenai/gpt-5-mini:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.5-pro:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.6-terra:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.6-luna:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityxaixai/grok-2-1212 is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-4-turbo:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/o3-pro:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-4.1-mini:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.6-luna-pro:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.6-terra-pro:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-4.1:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-4.1-nano:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-3.5-turbo:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.4-nano:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/o3:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.4:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.1:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityxaixai/grok-vision-beta is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.6-sol-pro:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-4o-mini:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.6-sol:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/o1-pro:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityxaixai/grok-2-latest is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5-nano:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityxaixai/grok-2-vision-1212 is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.4-pro:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.2:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/o4-mini-high:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/o3-mini-high:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/o4-mini:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.5:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/o1:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5-codex:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-4o:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityxaixai/grok-beta is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityxaixai/grok-2-vision-latest is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/o3-mini:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5-pro:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.4-mini:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityopenaiopenai/gpt-5.2-pro:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilitymoonshotaimoonshotai/kimi-k2.7-code:batch is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilitythedrummerthedrummer/rocinante-12b is no longer listed by this source — calls to it are expected to fail2026-08-28
availabilityz-aiz-ai/glm-5.2:batch is no longer listed by this source — calls to it are expected to fail2026-08-26
availabilitystealthstealth/ox-alpha is no longer listed by this source — calls to it are expected to fail2026-08-26
availabilityrunwaymlrunwayml/gen4_aleph is no longer listed by this source — calls to it are expected to fail2026-08-26
availabilityrunwaymlrunwayml/gen3a_turbo is no longer listed by this source — calls to it are expected to fail2026-08-26
availabilitygooglegoogle/gemma-3n-e4b-it is no longer listed by this source — calls to it are expected to fail2026-08-26
availabilityqwenqwen/qwen-plus-2025-07-28:thinking is no longer listed by this source — calls to it are expected to fail2026-08-25
availabilityinclusionaiinclusionai/ling-2.6-1t is no longer listed by this source — calls to it are expected to fail2026-08-24
availabilityinclusionaiinclusionai/ring-2.6-1t is no longer listed by this source — calls to it are expected to fail2026-08-24
availabilitynvidianvidia/nemotron-3-nano-30b-a3b:free is no longer listed by this source — calls to it are expected to fail2026-08-24
availabilityinclusionaiinclusionai/ling-2.6-flash is no longer listed by this source — calls to it are expected to fail2026-08-24
availabilitynvidianvidia/nemotron-nano-12b-v2-vl:free is no longer listed by this source — calls to it are expected to fail2026-08-24
availabilitynvidianvidia/nemotron-nano-9b-v2:free is no longer listed by this source — calls to it are expected to fail2026-08-24
availabilityopenaiopenai/gpt-oss-20b:free is no longer listed by this source — calls to it are expected to fail2026-08-22
availabilitydeepcogitodeepcogito/cogito-v2.1-671b is no longer listed by this source — calls to it are expected to fail2026-08-22
availabilityz-aiz-ai/glm-5.2:free is no longer listed by this source — calls to it are expected to fail2026-08-18
availabilityliquidliquid/lfm-2.5-2.6b:free is no longer listed by this source — calls to it are expected to fail2026-08-18
availabilitygooglegemini-3.7-flash-video-understanding-eap is no longer listed by this source — calls to it are expected to fail2026-08-13
availabilityinclusionaiinclusionai/ling-3.0-tiny:free is no longer listed by this source — calls to it are expected to fail2026-08-13
availabilityanthropicclaude-instant-1.0 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropicclaude-instant-1.2 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropicclaude-1.0 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropicclaude-instant-1.1 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityazuregpt-4o version 2024-05-13 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropicclaude-1.1 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropicclaude-1.2 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropicclaude-1.3 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityreplicatereplicateopenai/gpt-oss-20b vanished; replicate/openai/gpt-oss-20b appeared with identical price and context — most likely a rename inferred2026-08-07
availabilityinclusionaiinclusionai/ling-3.0-flash:free vanished; inclusionai/ling-3.0-tiny:free appeared with identical price and context — most likely a rename inferred2026-08-06
availabilityallenai[one catalog only — litellm still list it] allenai/olmo-3-32b-think is no longer listed by this source — calls to it are expected to fail inferred2026-08-29
availabilityxai[one catalog only — litellm still list it] xai/grok-2 is no longer listed by this source — calls to it are expected to fail inferred2026-08-28
availabilityxai[one catalog only — litellm still list it] xai/grok-2-vision is no longer listed by this source — calls to it are expected to fail inferred2026-08-28
availabilitymistralai[one catalog only — litellm still list it] mistralai/ministral-8b is no longer listed by this source — calls to it are expected to fail inferred2026-08-27
availabilitymancer[one catalog only — litellm still list it] mancer/weaver is no longer listed by this source — calls to it are expected to fail inferred2026-08-20
availabilityai21[one catalog only — litellm still list it] ai21/jamba-large-1.7 is no longer listed by this source — calls to it are expected to fail inferred2026-08-19
availabilitygoogle[one catalog only — litellm still list it] imagen-4.0-fast-generate-001 is no longer listed by this source — calls to it are expected to fail inferred2026-08-17
availabilitygoogle[one catalog only — litellm still list it] imagen-4.0-ultra-generate-001 is no longer listed by this source — calls to it are expected to fail inferred2026-08-17
availabilitygoogle[one catalog only — litellm still list it] imagen-4.0-generate-001 is no longer listed by this source — calls to it are expected to fail inferred2026-08-17
availabilityopenai[one catalog only — litellm still list it] openai/gpt-5.3-chat is no longer listed by this source — calls to it are expected to fail inferred2026-08-10
availabilitygoogle[one catalog only — litellm still list it] gemini-robotics-er-1.5-preview is no longer listed by this source — calls to it are expected to fail inferred2026-08-10
availabilitygoogle[one catalog only — litellm still list it] gemini-3-pro-preview is no longer listed by this source — calls to it are expected to fail inferred2026-08-10
availabilitygoogle[one catalog only — litellm still list it] gemini-2.0-flash-lite is no longer listed by this source — calls to it are expected to fail inferred2026-08-10
availabilitygoogle[one catalog only — litellm still list it] gemini-2.0-flash-lite-001 is no longer listed by this source — calls to it are expected to fail inferred2026-08-10
availabilitygoogle[one catalog only — litellm still list it] gemini-2.0-flash is no longer listed by this source — calls to it are expected to fail inferred2026-08-10
availabilitygoogle[one catalog only — litellm still list it] gemini-2.0-flash-001 is no longer listed by this source — calls to it are expected to fail inferred2026-08-10
availabilityanthropic[one catalog only — docs, litellm still list it] claude-opus-4-1-20250805 is no longer listed by this source — calls to it are expected to fail inferred2026-08-05
availabilityopenai[one catalog only — litellm still list it] gpt-4o-mini-tts-2025-03-20 is no longer listed by this source — calls to it are expected to fail inferred2026-08-05
availabilityanthropic[one catalog only — docs, litellm still list it] claude-sonnet-4-20250514 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropic[one catalog only — docs, litellm still list it] claude-3-7-sonnet-20250219 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityazure[one catalog only — docs, litellm, openrouter still list it] gpt-4o-mini is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityazure[one catalog only — docs, litellm, openrouter still list it] o4-mini is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropic[one catalog only — docs, litellm still list it] claude-3-sonnet-20240229 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropic[one catalog only — docs still list it] claude-2.1 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityazure[one catalog only — docs, litellm, openrouter still list it] gpt-4o is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropic[one catalog only — docs, litellm still list it] claude-3-5-sonnet-20241022 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropic[one catalog only — docs, litellm still list it] claude-3-opus-20240229 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityazure[one catalog only — docs, litellm, openrouter still list it] gpt-4.1-nano is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityazure[one catalog only — docs, litellm, openrouter still list it] gpt-4.1 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropic[one catalog only — docs, litellm still list it] claude-3-5-sonnet-20240620 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropic[one catalog only — docs, litellm still list it] claude-3-haiku-20240307 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityazure[one catalog only — docs, litellm, openrouter still list it] gpt-4.1-mini is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropic[one catalog only — docs, litellm still list it] claude-3-5-haiku-20241022 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropic[one catalog only — docs still list it] claude-2.0 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01
availabilityanthropic[one catalog only — docs, litellm still list it] claude-opus-4-20250514 is no longer listed by this source — calls to it are expected to fail inferred2026-08-01

Announced this period

ImpactProvider What changedDetected
availabilitytogether_aitogether_ai/deepseek-ai/DeepSeek-V4-Pro: deprecation announced for 2026-08-27 (date already passed) +3 aliases2026-08-28
availabilitytogether_aitogether_ai/google/gemma-3n-E4B-it: deprecation announced for 2026-08-25 (date already passed) +1 alias2026-08-28
availabilitytogether_aitogether_ai/deepseek-ai/DeepSeek-V4-Pro: retires 2026-08-27 (1 days ago) +3 aliases2026-08-28
availabilitytogether_aitogether_ai/google/gemma-3n-E4B-it: retires 2026-08-25 (3 days ago) +1 alias2026-08-28
availabilitymoonshotaimoonshotai/kimi-k2.5: retirement announced for 2026-08-31 — in 4 days +1 alias2026-08-27
availabilitymoonshotaimoonshotai/kimi-k2.5: retires 2026-08-31 (in 7 days)2026-08-24
availabilityazure_aiazure_ai/deepseek-r1: deprecation announced for 2026-08-13 (date already passed)2026-08-21
availabilityazure_aiazure_ai/claude-opus-4-1: deprecation announced for 2026-08-05 (date already passed) inferred2026-08-21
availabilitygeminigemini/gemini-robotics-er-1.6-preview: retires 2026-08-31 (in 10 days)2026-08-21
availabilityazure_aiazure_ai/deepseek-r1: retires 2026-08-13 (8 days ago)2026-08-21
availabilityazure_aiazure_ai/MAI-Image-2e: deprecation announced for 2026-08-15 (date already passed) inferred2026-08-21
availabilitygeminigemini/gemini-robotics-er-1.6-preview: deprecation announced for 2026-08-31 — in 10 days2026-08-21
availabilityazure_aiazure_ai/MAI-Image-2e: retires 2026-08-15 (6 days ago) inferred2026-08-21
availabilityazure_aiazure_ai/claude-opus-4-1: retires 2026-08-05 (16 days ago) inferred2026-08-21
availabilityvertex_ai-anthropic_modelsvertex_ai/claude-opus-4-1: deprecation announced for 2026-08-05 (date already passed) +1 alias inferred2026-08-21
availabilityvertex_ai-anthropic_modelsvertex_ai/claude-opus-4-1: retires 2026-08-05 (16 days ago) +1 alias inferred2026-08-21
availabilitynvidianvidia/nemotron-3-nano-30b-a3b:free: retires 2026-08-24 (in 4 days) +2 aliases inferred2026-08-20
availabilitynvidianvidia/nemotron-3-nano-30b-a3b:free: retirement announced for 2026-08-24 — in 4 days +2 aliases inferred2026-08-20
availabilityinclusionaiinclusionai/ling-2.6-1t: retires 2026-08-24 (in 5 days) +2 aliases inferred2026-08-19
availabilityinclusionaiinclusionai/ling-2.6-1t: retirement announced for 2026-08-24 — in 5 days +2 aliases inferred2026-08-19
availabilitydeepseekdeepseek/deepseek-v3.1-terminus: retires 2026-08-17 (in 2 days)2026-08-15
availabilitydeepseekdeepseek/deepseek-v3.1-terminus: retirement announced for 2026-08-17 — in 2 days2026-08-15
availabilitygroqgroq/llama-3.1-8b-instant: deprecation announced for 2026-08-16 — in 3 days +1 alias inferred2026-08-13
availabilitygroqgroq/meta-llama/llama-4-scout-17b-16e-instruct: deprecation announced for 2026-07-17 (date already passed) +1 alias2026-08-13
availabilitygroqgroq/meta-llama/llama-4-scout-17b-16e-instruct: retires 2026-07-17 (27 days ago) +1 alias2026-08-13
availabilitygroqgroq/llama-3.1-8b-instant: retires 2026-08-16 (in 3 days) +1 alias inferred2026-08-13
availabilityinclusionaiinclusionai/ling-3.0-tiny:free: retirement announced for 2026-08-13 — in 1 days inferred2026-08-12
availabilityinclusionaiinclusionai/ling-3.0-tiny:free: retires 2026-08-13 (in 1 days) inferred2026-08-12
availabilitymistralmistral/devstral-2512: retires 2026-07-31 (12 days ago) +5 aliases inferred2026-08-12
availabilityanthropicclaude-opus-4-1: deprecation announced for 2026-08-05 (date already passed) inferred2026-08-12
availabilitybedrockanthropic.claude-3-sonnet-20240229-v1:0: retires 2026-07-30 (13 days ago) +8 aliases2026-08-12
availabilitygeminigemini/imagen-4.0-fast-generate-001: retires 2026-08-17 (in 5 days) +2 aliases2026-08-12
availabilitybedrockcohere.command-r-plus-v1:0: retires 2026-08-19 (in 7 days) +1 alias inferred2026-08-12
availabilitybedrockanthropic.claude-3-haiku-20240307-v1:0: deprecation announced for 2026-09-10 — in 29 days +5 aliases2026-08-12
availabilitymistralmistral/mistral-medium-2505: deprecation announced for 2026-08-31 — in 19 days +2 aliases inferred2026-08-12
availabilitybedrockanthropic.claude-3-haiku-20240307-v1:0: retires 2026-09-10 (in 29 days) +5 aliases2026-08-12
availabilitygeminigemini/imagen-4.0-fast-generate-001: deprecation announced for 2026-08-17 — in 5 days +2 aliases2026-08-12
availabilitybedrockanthropic.claude-3-sonnet-20240229-v1:0: deprecation announced for 2026-07-30 (date already passed) +8 aliases2026-08-12
availabilitymistralmistral/mistral-medium-2505: retires 2026-08-31 (in 19 days) +2 aliases inferred2026-08-12
availabilitybedrockcohere.command-r-plus-v1:0: deprecation announced for 2026-08-19 — in 7 days +1 alias inferred2026-08-12
availabilityanthropicclaude-opus-4-1-20250805: retires 2026-08-05 (in 4 days) +1 alias inferred2026-08-12
availabilitymistralmistral/devstral-2512: deprecation announced for 2026-07-31 (date already passed) +5 aliases inferred2026-08-12
availabilitygeminigemini/gemini-embedding-2-preview: retires 2026-08-10 (2 days ago)2026-08-12
availabilitygeminigemini/gemini-embedding-2-preview: deprecation announced for 2026-08-10 (date already passed)2026-08-12
availabilityopenaigpt-5.2-chat-latest: deprecation announced for 2026-08-10 (date already passed) +1 alias2026-08-12
availabilityopenaigpt-4o-mini-search-preview-2025-03-11: deprecation announced for 2026-07-23 (date already passed) +15 aliases2026-08-12
availabilityopenaigpt-4o-audio: deprecation announced for 2026-07-20 (date already passed) +8 aliases inferred2026-08-05
availabilityopenaicomputer-use-preview-2025-03-11: retires 2026-07-23 (9 days ago) +17 aliases inferred2026-08-01
availabilityopenaigpt-5.2-chat-latest: retires 2026-08-10 (in 9 days) +3 aliases2026-08-01
availabilitygeminigemini/gemini-omni-flash-preview: deprecation announced for 2026-09-30 — in 31 days2026-08-30
availabilitygeminigemini/gemini-omni-flash-preview: retires 2026-09-30 (in 31 days)2026-08-30
availabilityazureazure/gpt-4o-transcribe: retires 2026-10-15 (in 64 days) +1 alias inferred2026-08-28
availabilityazureazure/gpt-4o-transcribe: deprecation announced for 2026-10-15 — in 64 days +1 alias inferred2026-08-28
availabilityanthropicclaude-sonnet-4-5-20250929: retires 2026-09-29 (in 39 days) +1 alias2026-08-21
availabilityazureazure/eu/o3-mini-2025-01-31: retires 2026-10-01 (in 50 days) +5 aliases2026-08-21
availabilityvertex_ai-language-modelsgemini-2.5-flash-lite: retires 2026-10-20 (in 60 days) +2 aliases2026-08-21
availabilityazure_aiazure_ai/claude-haiku-4-5: retires 2026-10-19 (in 59 days) +2 aliases inferred2026-08-21
availabilityazureazure/gpt-4.1-nano-2025-04-14: retires 2026-10-14 (in 63 days) +2 aliases2026-08-21
availabilitytext-completion-openaibabbage-002: retires 2026-09-28 (in 38 days) +2 aliases2026-08-21
availabilityazureazure/o4-mini-2025-04-16: deprecation announced for 2026-10-16 — in 65 days +2 aliases2026-08-21
availabilityvertex_ai-language-modelsgemini-2.5-flash-lite: deprecation announced for 2026-10-20 — in 60 days +2 aliases2026-08-21
availabilityazureazure/gpt-4.1-nano: deprecation announced for 2026-10-14 — in 54 days2026-08-21
availabilityvertex_ai-video-modelsvertex_ai/veo-3.1-fast-generate-001: deprecation announced for 2026-11-17 — in 88 days +1 alias inferred2026-08-21
availabilityanthropicclaude-haiku-4-5-20251001: deprecation announced for 2026-10-15 — in 55 days +1 alias2026-08-21
availabilityopenaift-babbage-002: retires 2026-10-23 (in 83 days) +45 aliases2026-08-21
availabilityazureazure/o4-mini-2025-04-16: retires 2026-10-16 (in 65 days) +2 aliases2026-08-21
availabilityazureazure/gpt-image-1: retires 2026-10-23 (in 72 days) +9 aliases2026-08-21
availabilityazureazure/eu/o1-2024-12-17: retires 2026-10-21 (in 70 days) +6 aliases2026-08-21
availabilityanthropicclaude-haiku-4-5-20251001: retires 2026-10-15 (in 55 days) +1 alias2026-08-21
availabilityopenaift:gpt-3.5-turbo-0125: deprecation announced for 2026-10-23 — in 72 days +33 aliases2026-08-21
availabilityvertex_ai-anthropic_modelsvertex_ai/claude-sonnet-4-5: deprecation announced for 2026-09-29 — in 39 days +1 alias inferred2026-08-21
availabilityazureazure/eu/o3-mini-2025-01-31: deprecation announced for 2026-10-01 — in 50 days +4 aliases2026-08-21
availabilityanthropicclaude-sonnet-4-5-20250929: deprecation announced for 2026-09-29 — in 39 days +1 alias2026-08-21
availabilityvertex_ai-video-modelsvertex_ai/veo-3.1-fast-generate-001: retires 2026-11-17 (in 88 days) +1 alias inferred2026-08-21
availabilityvertex_ai-language-modelsgemini-2.5-flash-image: deprecation announced for 2026-10-02 — in 42 days +1 alias2026-08-21
availabilityazureazure/eu/o1-2024-12-17: deprecation announced for 2026-10-21 — in 70 days +4 aliases2026-08-21
availabilityvertex_ai-language-modelsgemini-2.5-flash-image: retires 2026-10-02 (in 42 days) +1 alias2026-08-21
availabilityazure_aiazure_ai/claude-haiku-4-5: deprecation announced for 2026-10-19 — in 59 days +2 aliases inferred2026-08-21
availabilityvertex_ai-anthropic_modelsvertex_ai/claude-haiku-4-5: deprecation announced for 2026-10-15 — in 55 days +1 alias inferred2026-08-21
availabilitytext-completion-openaibabbage-002: deprecation announced for 2026-09-28 — in 38 days +2 aliases2026-08-21

Who changed the most

ProviderUnannounced changes
deepseek66
openai49
z-ai44
qwen30
~deepseek23
xai21
google20
anthropic20
moonshotai20
nvidia16
deepinfra16
novita8

← July 2026  ·  September 2026 →  ·  All reports