Change feed
Every pricing change and breaking deprecation we've detected, newest first.
-
NVIDIA Nemotron models changed
Tracking since Jun 25, 2026 · 27 snapshots on file
Model nvidia/Privasis-Cleaner-4B with other license was removed from NVIDIA Nemotron.
- nvidia/Privasis-Cleaner-4B removed
- license type: other
- removal date: 2026-06-08
Full history & current snapshot →View raw diff +0 −1
- nvidia/Privasis-Cleaner-4B license:other 2026-06-08 -
Tracking since Jun 25, 2026 · 18 snapshots on file
DigitalOcean API: 3 additive changes
Full history & current snapshot →View raw diff +3 −0
+ New endpoint: POST /v2/insights/query/{region}/traces/search + New endpoint: GET /v2/insights/query/{region}/traces/{trace_id} + New endpoint: POST /v2/insights/query/{region}/spans/search
-
Google Gemini API changelog changed
Tracking since Jun 25, 2026 · 25 snapshots on file
Deep Research Agent deep-research-pro-preview-12-2025 is deprecated and will shut down on October 23, 2026; users must migrate to deep-research-preview-04-2026 or deep-research-max-preview-04-2026.
- deep-research-pro-preview-12-2025 agent deprecation with shutdown on October 23, 2026
- migrate agent parameter to deep-research-preview-04-2026 or deep-research-max-preview-04-2026
- update interactions.create requests with new agent parameter
Full history & current snapshot →View raw diff +10 −0
+ October 8, 2026 + Deep Research Agent deep-research-pro-preview-12-2025 deprecation: + The deep-research-pro-preview-12-2025 agent is deprecated and will be + shut down on **October 23, + 2026**. + Migrate your requests to one of the newer [Deep + Research](https://ai.google.dev/gemini-api/docs/deep-research) agent versions: + To migrate, update the agent parameter in your interactions.create + requests from deep-research-pro-preview-12-2025 to + deep-research-preview-04-2026 (or deep-research-max-preview-04-2026).
-
Google Gemini deprecations changed
Tracking since Jun 25, 2026 · 22 snapshots on file
Added three new deep-research model versions with April 2026 and December 2025 release dates and sunset information.
- deep-research-preview-04-2026 released April 21, 2026 with no shutdown date announced
- deep-research-max-preview-04-2026 released April 21, 2026 with no shutdown date announced
- deep-research-pro-preview-12-2025 released December 11, 2025 with October 23, 2026 shutdown date
Full history & current snapshot →View raw diff +3 −0
+ | deep-research-preview-04-2026 | April 21, 2026 | No shutdown date announced | | + | deep-research-max-preview-04-2026 | April 21, 2026 | No shutdown date announced | | + | deep-research-pro-preview-12-2025 | December 11, 2025 | October 23, 2026 | deep-research-preview-04-2026 |
-
Tracking since Jun 25, 2026 · 26 snapshots on file
Arena Elo (text, overall) updated.
Full history & current snapshot →View raw diff +34 −34
- amazon-nova-experimental-chat-26-02-10 1449 - claude-fable-5-high 1492 - claude-fable-5.1-max 1511 - claude-opus-5-high 1502 - claude-opus-5.5-high 1512 - claude-sonnet-5-high 1443 - claude-sonnet-5.5-xhigh 1467 - deepseek-v4-pro-high-20260813 1445 - ernie-5.0-preview-1203 1442 - gemini-3.1-pro-preview 1480 - gemini-3.5-flash-medium 1478 - gemini-3.7-flash-high 1487 - gemini-3.8-flash-high 1497 - gemini-4-argon-high 1533 - glm-4.7 1436 - glm-5.3-flash 1470 - gpt-5.5-high 1471 - gpt-5.6-luna-xhigh 1431 - gpt-5.6-sol-xhigh 1456 - gpt-5.6-terra-xhigh 1446 - gpt-6-astra-max 1442 - gpt-6.1-sol-max 1446 - grok-4.1-thinking 1437 - grok-4.20-beta1 1444 - grok-4.6-high 1427 - hy3 1441 - inkling 1442 - longcat-flash-chat-2602-exp 1427 - mimo-v2.6-flash 1456 - mimo-v2.6-pro 1491 - minimax-m3 1432 - muse-spark-1.2 (xHigh) 1483 - qwen3.8-max 1482 - Step 5 Preview 1448 + amazon-nova-experimental-chat-26-02-10 1448 + claude-fable-5-high 1491 + claude-fable-5.1-max 1510 + claude-opus-5-high 1503 + claude-opus-5.5-high 1515 + claude-sonnet-5-high 1442 + claude-sonnet-5.5-xhigh 1471 + deepseek-v4-pro-high-20260813 1446 + ernie-5.0-preview-1203 1443 + gemini-3.1-pro-preview 1481 + gemini-3.5-flash-medium 1477 + gemini-3.7-flash-high 1486 + gemini-3.8-flash-high 1499 + gemini-4-argon-high 1534 + glm-4.7 1435 + glm-5.3-flash 1471 + gpt-5.5-high 1472 + gpt-5.6-luna-xhigh 1430 + gpt-5.6-sol-xhigh 1457 + gpt-5.6-terra-xhigh 1447 + gpt-6-astra-max 1440 + gpt-6.1-sol-max 1447 + grok-4.1-thinking 1436 + grok-4.20-beta1 1445 + grok-4.6-high 1428 + hy3 1439
-
Anthropic model deprecations changed
Tracking since Jun 25, 2026 · 45 snapshots on file
Multiple Claude model versions were retired with replacement models specified, and temperature/top_p/top_k parameters deprecated for Claude Opus 4.7 and later models.
- claude-2.0, claude-2.1, and claude-3-sonnet-20240229 retired July 21, 2025 with replacements claude-opus-4-8 and claude-sonnet-4-6
- claude-1.0 through claude-1.3 and claude-instant versions retired November 6, 2024 with replacement claude-haiku-4-5-20251001
- temperature, top_p, and top_k parameters deprecated for Claude Opus 4.7 and later, returning 400 error when set to non-default values
Full history & current snapshot →View raw diff +0 −18
- On January 21, 2025, Anthropic notified developers using Claude 2, Claude 2.1, and Claude Sonnet 3 models of their upcoming retirements. - | July 21, 2025 | claude-2.0 | claude-opus-4-8 | - | July 21, 2025 | claude-2.1 | claude-opus-4-8 | - | July 21, 2025 | claude-3-sonnet-20240229 | claude-sonnet-4-6 | - 2024-09-04: Claude 1 and Instant models - On September 4, 2024, Anthropic notified developers using Claude 1 and Instant models of their upcoming retirements. - | November 6, 2024 | claude-1.0 | claude-haiku-4-5-20251001 | - | November 6, 2024 | claude-1.1 | claude-haiku-4-5-20251001 | - | November 6, 2024 | claude-1.2 | claude-haiku-4-5-20251001 | - | November 6, 2024 | claude-1.3 | claude-haiku-4-5-20251001 | - | November 6, 2024 | claude-instant-1.0 | claude-haiku-4-5-20251001 | - | November 6, 2024 | claude-instant-1.1 | claude-haiku-4-5-20251001 | - | November 6, 2024 | claude-instant-1.2 | claude-haiku-4-5-20251001 | - API parameter deprecations - Anthropic occasionally deprecates request parameters that no longer apply to current models. How the API treats a deprecated parameter depends on the model, as the following table shows. Most SDKs keep deprecated parameters in their request types so existing code continues to type-check. The Python SDK (v1.0 and later) removes temperature, top_p, and top_k, so passing them raises a TypeError. - | Parameter | Status | Behavior | Recommended replacement | - | temperature, top_p, top_k | Deprecated (Claude Opus 4.7 and later) | Returns a 400 error when set to a non-default value on Claude 4.7 and later models and Claude Mythos Preview. | Omit and use prompting to guide model behavior. | - For migration steps, see the migration guide.
-
xAI (Grok) rate limits changed
Tracking since Jun 25, 2026 · 17 snapshots on file
Updated code example to use OpenAI Python client library instead of xAI gRPC SDK for handling rate limit retries.
Full history & current snapshot →View raw diff +9 −11
- import grpc - from xai_sdk import Client - from xai_sdk.chat import user - client = Client(api_key=os.getenv("XAI_API_KEY")) - def request_with_backoff(prompt, max_retries=5): - chat = client.chat.create(model="grok-4.7") - chat.append(user(prompt)) - return chat.sample() - except grpc.RpcError as e: - if e.code()!= grpc.StatusCode.RESOURCE_EXHAUSTED: - raise RuntimeError("Max retries exceeded") + from openai import OpenAI, RateLimitError + client = OpenAI(base_url="https://api.x.ai/v1", api_key=os.getenv("XAI_API_KEY")) + def request_with_backoff(messages, max_retries=5): + return client.chat.completions.create( + model="grok-4.7", + messages=messages, + ) + except RateLimitError: + if attempt == max_retries - 1:
-
Tracking since Jun 25, 2026 · 18 snapshots on file
DigitalOcean API: 10 additive changes
Full history & current snapshot →View raw diff +10 −0
+ New endpoint: GET /v1/consent + New endpoint: GET /v1/consent/{agent_id} + New endpoint: PUT /v1/consent/{agent_id} + New endpoint: GET /v1/signals/agents/{agent_id}/sessions + New endpoint: GET /v1/signals/sessions/{session_id}/dialogues + New endpoint: GET /v1/signals/exports + New endpoint: POST /v1/signals/exports + New endpoint: GET /v1/signals/exports/options + New endpoint: GET /v1/signals/exports/{export_id} + New endpoint: GET /v1/signals/exports/{export_id}/download
-
Tracking since Jun 25, 2026 · 104 snapshots on file
OpenAI API: 2 additive changes
Full history & current snapshot →View raw diff +2 −0
+ New endpoint: GET /agents/environments + New endpoint: POST /agents/environments
-
Tracking since Jun 25, 2026 · 280 snapshots on file
OpenAI pricing page now displays Business and Enterprise ChatGPT plans with per-seat subscription model ($20-$100/month)
Page now shows ChatGPT Business and Enterprise plans (seat-based pricing) instead of API token pricingFull history & current snapshot →View raw diff +214 −43
- Business Pricing | OpenAI - Price - Input:$10.00 / 1M tokens Cached input:$1.00 / 1M tokens Output:$50.00 / 1M tokens - Price - Input:$2.00 / 1M tokens Cached input:$0.10 / 1M tokens Output:$10.00 / 1M tokens - Price - Input:$0.10 / 1M tokens Cached input:$0.01 / 1M tokens Output:$0.50 / 1M tokens - Price - Image:$8.00 / 1M tokens for inputs$2.00 / 1M tokens for cached inputs$30.00 / 1M tokens for outputs Text:$5.00 / 1M tokens for inputs$1.25 / 1M tokens for cached inputs - Price - Image:$8.00 / 1M tokens for inputs$2.00 / 1M tokens for cached inputs$30.00 / 1M tokens for outputs Text:$5.00 / 1M tokens for inputs$1.25 / 1M tokens for cached inputs - Price - $0.05 per minute / $0.00083 per second - Price - Audio:$32.00 / 1M tokens for inputs$0.40 / 1M tokens for cached inputs$64.00 / 1M tokens for outputs Text:$4.00 / 1M tokens for inputs$0.40 / 1M tokens for cached inputs$24.00 / 1M tokens for outputs Image:$5.00 / 1M tokens for inputs$0.50 / 1M tokens for cached inputs - Price - Audio:$10.00 / 1M tokens for inputs$0.30 / 1M tokens for cached inputs$20.00 / 1M tokens for outputs Text:$0.60 / 1M tokens for inputs$0.06 / 1M tokens for cached inputs$2.40 / 1M tokens for outputs Image:$0.80 / 1M tokens for inputs$0.08 / 1M tokens for cached inputs - Price - $0.017 per minute / $0.00028 per second - Price - $0.0045 per minute / $0.00008 per second - Price - $0.034 per minute / $0.00057 per second - Price - $10.00 / 1k calls Search content tokens are free. - Price - Now:1 GB for $0.03 / 64GB for $1.92 per container Starting March 31, 2026:1 GB for $0.03 / 64GB for $1.92 per 20-minute session per container - Provides lower costs for requests in exchange for slower response times and occasional resource unavailability. Ideal for non-production or lower priority tasks. - Enterprise offerings - Contact our sales team to learn more about Data residency(opens in a new window), Scale Tier andReserved Capacity designed for cutting-edge customers running larger workloads. - We recommend experimenting with all of these models in the Playground(opens in a new window) to explore which models provide the best price performance trade-off for your usage. - Do you offer an enterprise package or SLAs? - We offer different tiers of access to our enterprise customers that include SLAs, lower latency, and more. Please contact our sales team to learn more. - Yes, we treat Playground usage the same as regular API usage. You will be billed at the per-token input and output prices mentioned above. - How will I know how many tokens I’ve used each month? - A token is a mathematical representation of natural language. Log in to your account to view your usage tracking dashboard(opens in a new window). This dashboard will show you how many tokens you’ve used during the current and past billing cycles. - You can set a monthly budget in your billing settings(opens in a new window), after which we’ll stop serving your requests. There may be a delay in enforcing the limit, and you are responsible for any overage incurred. You can also configure an email notification threshold to receive an email alert once you cross that threshold each month. We recommend checking your usage tracking dashboard(opens in a new window) regularly to monitor your spend. - Is access to the API included in ChatGPT Plus, Business, Enterprise or Edu? - No, OpenAI APIs are billed separately from ChatGPT Plus, Business, Enterprise and Edu. ChatGPT subscription pricing can be found at openai.com/chatgpt/pricing/. - Images are converted into tokens and charged per token. Text models price image tokens at standard text token rates, while GPT Image and gpt-realtime uses a separate image token rate. Models like gpt-4.1-mini, gpt-4.1-nano, and o4-mini convert images into tokens differently. Learn more in our docs(opens in a new window). - ChatGPT Business(opens in a new window) - ChatGPT Enterprise(opens in a new window) - ChatGPT for Education(opens in a new window) + ChatGPT Business(opens in a new window) + ChatGPT Enterprise(opens in a new window) + ChatGPT for Education(opens in a new window) + Image 1 + Business + A secure workspace with company context and flexible seat types for any budget. + Get started(opens in a new window) + Standard seat + $20/ month + $20/month if billed annually. $25/month if billed monthly. + Premium seat + $100/ month + 5x more usage than standard, with no 5-hour limit + $100/month if billed annually. $125/month if billed monthly. + No training on your business data by default + Mix and match seat types + For teams of 2–200 employees. Unlimited subject to abuse guardrails. Learn more(opens in a new window)
-
Tracking since Jun 25, 2026 · 14 snapshots on file
Claude Design and Slides removed from Pro plan feature set
Pro · featuresClaude Design and Slides removed from Pro plan features listFull history & current snapshot →View raw diff +1 −1
- The Pro plan gives you everything in a Free plan with more usage and more of Claude's capabilities. That includes Claude Code, Claude Design, Slides, and Docs, along with projects to organize your chats and documents, access to more Claude models, and Claude for Microsoft 365. You can choose a monthly or annual subscription. + The Pro plan gives you everything in a Free plan with more usage and more of Claude's capabilities. That includes Claude Code, along with projects to organize your chats and documents, access to more Claude models, and Claude for Microsoft 365. You can choose a monthly or annual subscription.
-
Tracking since Jun 25, 2026 · 57 snapshots on file
Updated Rate limits page URL from /settings/limits to /usage/limits; changed 'Request rate limit increase' button label to 'Request tier increase'.
Full history & current snapshot →View raw diff +3 −3
- Limits are set at the organization level. You can see your organization's tier and current limits on the Rate limits page in the Claude Console. - Rate limits are applied separately for each model; therefore you can use different models up to their respective limits simultaneously. You can check your current rate limits and behavior on the Rate limits page in the Claude Console, or read the configured limits programmatically with the Rate Limits API. - To request higher rate limits or a higher monthly spend cap, use Request rate limit increase on the Rate limits page. Anthropic support can also raise limits; for urgent needs, contact Anthropic support. + Limits are set at the organization level. You can see your organization's tier and current limits on the Rate limits page in the Claude Console. + Rate limits are applied separately for each model; therefore you can use different models up to their respective limits simultaneously. You can check your current rate limits and behavior on the Rate limits page in the Claude Console, or read the configured limits programmatically with the Rate Limits API. + To request higher rate limits or a higher monthly spend cap, use Request tier increase on the Rate limits page. Anthropic support can also raise limits; for urgent needs, contact Anthropic support.
-
Mistral AI reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Mistral Large 4: 10 benchmark claims (via web search)
Full history & current snapshot → -
Tracking since Jun 25, 2026 · 57 snapshots on file
Removed mention of Bedrock availability for retired Claude Haiku 3.5 model in rate limits table.
Full history & current snapshot →View raw diff +1 −1
- | Claude Haiku 3.5 (retired, except on Bedrock and Google Cloud ) | 1,000 | 100,000 4 | 20,000 | + | Claude Haiku 3.5 (retired, except on Google Cloud ) | 1,000 | 100,000 4 | 20,000 |
-
Tracking since Jun 25, 2026 · 14 snapshots on file
Claude 3 model pricing now explicitly separated by context window length (≤100K vs >100K tokens)
Pricing structure documentation now includes explicit tiered rates for context length on older Claude models (Claude 3 Haiku and Sonnet)Full history & current snapshot →View raw diff +10 −1
- Ideal for complex agentic coding and enterprise work + Prompts ≤ 100K tokens + $0.01 / MTok + $0.125 / MTok + Prompts > 100K tokens + $0.05 / MTok + $0.625 / MTok + Prompts ≤ 100K tokens + Prompts > 100K tokens + Prompts ≤ 100K tokens + Prompts > 100K tokens
-
Tracking since Jun 25, 2026 · 5 snapshots on file
Model registry updated with new model additions and removals: MolmoWeb models removed, new Bolmo, Bwen, and Llama-3-Blama models added.
- Removed: allenai/MolmoWeb-8B-Native and allenai/MolmoWeb-Pretrained-4B
- Added: allenai/Bolmo-1B-Stage1, allenai/Bolmo-7B-Stage1, allenai/Bwen-8B, allenai/Bwen-8B-Stage1, allenai/Llama-3-Blama-8B, allenai/Llama-3-Blama-8B-Stage1
- All new models dated 2026-08-26 with apache-2.0 or llama3 licenses
Full history & current snapshot →View raw diff +6 −2
- allenai/MolmoWeb-8B-Native license:apache-2.0 2026-03-24 - allenai/MolmoWeb-Pretrained-4B license:apache-2.0 2026-04-08 + allenai/Bolmo-1B-Stage1 license:apache-2.0 2026-08-26 + allenai/Bolmo-7B-Stage1 license:apache-2.0 2026-08-26 + allenai/Bwen-8B license:apache-2.0 2026-08-26 + allenai/Bwen-8B-Stage1 license:apache-2.0 2026-08-26 + allenai/Llama-3-Blama-8B license:llama3 2026-08-26 + allenai/Llama-3-Blama-8B-Stage1 license:llama3 2026-08-26
-
Anthropic models model list changed
Tracking since Jun 25, 2026 · 8 snapshots on file
New Anthropic model claude-haiku-5-5 is now available.
- claude-haiku-5-5
Full history & current snapshot →View raw diff +1 −0
+ claude-haiku-5-5 -
Tracking since Jun 25, 2026 · 57 snapshots on file
Added Claude Haiku 5.5 model to Start tier rate limits (1,000 RPM / 2M ITPM / 400K OTPM); marked Sonnet 4.x as deprecated.
Full history & current snapshot →View raw diff +2 −1
- _3 Sonnet 4.x rate limit is a total limit that applies to combined traffic across Sonnet 4.6 and Sonnet 4.5. Claude Sonnet 5.5 and Claude Sonnet 5 each have a separate rate limit and are not part of this combined bucket._ + | Claude Haiku 5.5 | 1,000 | 2,000,000 | 400,000 | + _3 Sonnet 4.x rate limit is a total limit that applies to combined traffic across Sonnet 4.6 and Sonnet 4.5 (deprecated ). Claude Sonnet 5.5 and Claude Sonnet 5 each have a separate rate limit and are not part of this combined bucket._
-
Anthropic model deprecations changed
Tracking since Jun 25, 2026 · 45 snapshots on file
Added a table row indicating an Active status with a date of October 7, 2027.
- Status: Active
- Date: October 7, 2027
Full history & current snapshot →View raw diff +1 −0
+ | | Active | N/A | Not sooner than October 7, 2027 | -
NVIDIA Nemotron models changed
Tracking since Jun 25, 2026 · 27 snapshots on file
NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 model removed or no longer available as of 2026-06-03.
- nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 deleted
- license:other
- expiration date 2026-06-03
Full history & current snapshot →View raw diff +0 −1
- nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 license:other 2026-06-03 -
Tracking since Jun 25, 2026 · 18 snapshots on file
DigitalOcean API: 4 additive changes
Full history & current snapshot →View raw diff +4 −0
+ New endpoint: POST /v2/insights/query/{region}/prom/api/v1/query + New endpoint: POST /v2/insights/query/{region}/prom/api/v1/query_range + New endpoint: POST /v2/insights/query/{region}/prom/api/v1/labels + New endpoint: POST /v2/insights/query/{region}/prom/api/v1/series
-
Google Gemini API changelog changed
Tracking since Jun 25, 2026 · 25 snapshots on file
Gemini API deprecation notice changed from having a specific shutdown date (October 29, 2026) to stating no shutdown date has been announced.
- shutdown date changed from October 29, 2026 to no shutdown date announced
- deprecation link to shutdown timeline removed
Full history & current snapshot →View raw diff +1 −2
- deprecated and will be shut down on - October 29, 2026. Migrate to + deprecated (no shutdown date announced). Migrate to
-
Google Gemini deprecations changed
Tracking since Jun 25, 2026 · 22 snapshots on file
The shutdown date for gemini-3.1-flash-image model changed from October 29, 2026 to no announced shutdown date.
- gemini-3.1-flash-image shutdown date changed from October 29, 2026 to No shutdown date announced
Full history & current snapshot →View raw diff +1 −1
- | gemini-3.1-flash-image | May 28, 2026 | October 29, 2026 | gemini-nano-banana-2.1 | + | gemini-3.1-flash-image | May 28, 2026 | No shutdown date announced | gemini-nano-banana-2.1 |
-
Google Gemini API pricing changed
Tracking since Jun 25, 2026 · 22 snapshots on file
Gemini 3.0 Flash and Flash-8B image generation prices increased up to 49% for 4K output
Gemini 3.0 Flash (image output) · output_token_price$0.0756 per 4K image→$0.113 per 4K imageGemini 3.0 Flash-8B (image output) · output_token_price$0.0378 per 4K image→$0.0567 per 4K imageFull history & current snapshot →View raw diff +4 −4
- and $0.0756 per 4K image*. - and $0.0378 per 4K image*. - Output images at 4K (4096x4096px) consume 2520 tokens and are equivalent to - $0.0756 per image. + and $0.113 per 4K image*. + and $0.0567 per 4K image*. + Output images at 4K (4096x4096px) consume 3780 tokens and are equivalent to + $0.113 per image.
-
Tracking since Jun 25, 2026 · 13 snapshots on file
Usage tier system restructured from 5 tiers (Tier 1–5) to 3 named tiers (Build, Launch, Grow) with adjusted spend thresholds and new standard rate limit table added showing model-specific RPM and TPM values.
now$500was$100/mo$50$1,000/mo$250$1,000Full history & current snapshot →View raw diff +15 −7
- You can view the rate and usage limits for your organization under the limits section of your account settings. As your spend on our API goes up, we automatically graduate you to the next usage tier. This usually results in an increase in rate limits across most models. - | Tier 1 | $5 paid | $100 / month | - | Tier 2 | $50 paid | $500 / month | - | Tier 3 | $100 paid | $1,000 / month | - | Tier 4 | $250 paid | $5,000 / month | - | Tier 5 | $1,000 paid | $200,000 / month | - To view a high-level summary of rate limits per model, visit the models page. + The three paid usage tiers are Build, Launch, and Grow. Your organization’s usage tier upgrades automatically as its total credit purchases reach each threshold. Higher tiers generally provide higher rate limits across models. + | Build | $5 in total credit purchases | $500 / month | + | Launch | $100 in total credit purchases | $5,000 / month | + | Grow | $500 in total credit purchases | $200,000 / month | + Standard rate limits + To view the limits for each model at your usage tier, go to Settings > Organization > Limits and review Rate limits. To upgrade your usage tier, select Upgrade tier in the Usage Tiers section. + These Standard rate limits differ from Ultrafast rate limits. + | Tier | Model | RPM | TPM | + | --- | --- | --- | --- | + | Build | Astra, Sol, Terra | 5,000 | 1,000,000 | + | Luna | 5,000 | 2,000,000 | + | Launch | Astra, Sol, Terra | 10,000 | 4,000,000 | + | Luna | 10,000 | 10,000,000 | + | Grow | Astra, Sol, Terra | 15,000 | 40,000,000 | + | Luna | 30,000 | 180,000,000 |
-
Tracking since Jun 25, 2026 · 104 snapshots on file
OpenAI API: 4 additive changes
Full history & current snapshot →View raw diff +4 −0
+ New endpoint: POST /decisions + New endpoint: POST /vaults/{vault_id} + New optional parameter: metadata (query) + New optional parameter: metadata (query)
-
Tracking since Jun 25, 2026 · 25 snapshots on file
Cursor released remote control feature for local agents accessible via iOS app, with default availability for non-Enterprise organizations.
- Remote control available in Cursor iOS app for viewing and messaging local agents
- Available by default for everyone except Enterprise organizations
- Requires computer to stay on and online for remote control to work
Full history & current snapshot →View raw diff +26 −36
- Aug 19, 2026 · Changelog - Cloud Agents and Cursor Harness Improvements - We're continuing to improve cloud agents and the Cursor harness so always-on agents can operate as a system, building and shipping software on their own without the need for intervention at each loop. - With this release, cloud agents can automatically pick up work in response to events, hold a goal until it's met, and stay on course through long-running sessions. - Cursor can now monitor your PRs, watch a Slack thread, or run scheduled tasks. Cursor Agent subscribes to an event source (a thread or conversation) and wakes when something happens. Subscriptions are available for cloud agents only, for now. - Cloud agents automatically subscribe to PRs they create and drive them to completion, fixing CI and addressing bot comments. In Slack, ask @cursor check back in an hour and keep going until that feedback is in. - #Custom modes - Use any skill as a Custom Mode: a skill that stays pinned in the chat. Custom modes keep agents focused on a skill - you can think about it like "always on" skills. - From /, pick a skill and press ⌥⏎ (Mac) or Alt+Enter (Windows), or choose Use as Mode. - #Subagents on their own machines - Subagents can now run on their own virtual machines. Each gets an isolated copy of the project with clean context in its own cloud environment. - Have subagents test the parent agent's changes in fresh environments or swarm independent fixes without collisions. Try run a swarm of subagents to test my app for bugs, each in its own environment. - #/goal - Use /goal to give the agent a long-lived objective to work towards until it's fully complete. - Try /goal fix all flaky tests and make CI green in a new chat. Pair it with a custom mode to follow a playbook, or /loop for recurring check-ins. - #Steering improvements - You can now send a message to steer the agent while it's working without interruption. Follow-ups wait for the next tool call instead of cutting the agent off mid-action. - Type a follow-up and hit Send now, or press ⏎ twice. - Aug 19, 2026 · Changelog - Cloud Agents and Cursor Harness Improvements - We're continuing to improve cloud agents and the Cursor harness so always-on agents can operate as a system, building and shipping software on their own without the need for intervention at each loop. - With this release, cloud agents can automatically pick up work in response to events, hold a goal until it's met, and stay on course through long-running sessions. - Cursor can now monitor your PRs, watch a Slack thread, or run scheduled tasks. Cursor Agent subscribes to an event source (a thread or conversation) and wakes when something happens. Subscriptions are available for cloud agents only, for now. - Cloud agents automatically subscribe to PRs they create and drive them to completion, fixing CI and addressing bot comments. In Slack, ask @cursor check back in an hour and keep going until that feedback is in. - #Custom modes - Use any skill as a Custom Mode: a skill that stays pinned in the chat. Custom modes keep agents focused on a skill - you can think about it like "always on" skills. - From /, pick a skill and press ⌥⏎ (Mac) or Alt+Enter (Windows), or choose Use as Mode. - #Subagents on their own machines - Subagents can now run on their own virtual machines. Each gets an isolated copy of the project with clean context in its own cloud environment. - Have subagents test the parent agent's changes in fresh environments or swarm independent fixes without collisions. Try run a swarm of subagents to test my app for bugs, each in its own environment. - #/goal - Use /goal to give the agent a long-lived objective to work towards until it's fully complete. - Try /goal fix all flaky tests and make CI green in a new chat. Pair it with a custom mode to follow a playbook, or /loop for recurring check-ins. - #Steering improvements - You can now send a message to steer the agent while it's working without interruption. Follow-ups wait for the next tool call instead of cutting the agent off mid-action. - Type a follow-up and hit Send now, or press ⏎ twice. + Oct 6, 2026 · Changelog + Remote control for local agents + You can now see and reply to the local agents running on your computer from the Cursor iOS app. + Remote Control is on by default for everyone except Enterprise organizations. + Download the Cursor iOS app and sign in. Computers on your account appear in the app automatically. Tap your computer in the app and approve your pairing request in the Cursor desktop app. Your local agents appear in a list. Tap one to see what it's doing or send it a message. + #Agents keep running on your computer + Remote control doesn't move your agents anywhere. It keeps running on your computer, and the app connects to it. Your computer needs to stay on and online for remote control to work. + #Keep your computer awake + To stop your computer from sleeping while you're away, turn on "Keep this computer awake" under Remote Control in Cursor's desktop settings. Your computer needs to be plugged in with the lid open. + #For Enterprise teams + Remote control is on by default, except for Enterprise organizations. Enterprise admins can turn it on in Org settings > Security and identity > Remote control. + Remote control doesn't require cloud agents to work. + Remote control is available today in the Cursor iOS app. + Oct 6, 2026 · Changelog + Remote control for local agents + You can now see and reply to the local agents running on your computer from the Cursor iOS app. + Remote Control is on by default for everyone except Enterprise organizations. + Download the Cursor iOS app and sign in. Computers on your account appear in the app automatically. Tap your computer in the app and approve your pairing request in the Cursor desktop app. Your local agents appear in a list. Tap one to see what it's doing or send it a message. + #Agents keep running on your computer + Remote control doesn't move your agents anywhere. It keeps running on your computer, and the app connects to it. Your computer needs to stay on and online for remote control to work. + #Keep your computer awake + To stop your computer from sleeping while you're away, turn on "Keep this computer awake" under Remote Control in Cursor's desktop settings. Your computer needs to be plugged in with the lid open. + #For Enterprise teams + Remote control is on by default, except for Enterprise organizations. Enterprise admins can turn it on in Org settings > Security and identity > Remote control.
-
Google Gemini API changelog changed
Tracking since Jun 25, 2026 · 25 snapshots on file
Gemini Nano Banana 2.1 model released as GA and gemini-3.1-flash-image deprecated with October 29, 2026 shutdown date.
- Gemini Nano Banana 2.1 (gemini-nano-banana-2.1) released as high-efficiency image generation model
- gemini-3.1-flash-image model deprecated and will shut down October 29, 2026
- Supports wide and panoramic aspect ratios (1:4, 4:1, 1:8, 8:1) at 1K, 2K, and 4K resolutions
Full history & current snapshot →View raw diff +16 −0
+ October 6, 2026 + Gemini Nano Banana 2.1 generally available (GA): Released + Gemini Nano Banana 2.1 + (gemini-nano-banana-2.1), the latest high-efficiency image generation and + conversational editing model. An update to Nano Banana 2 + (gemini-3.1-flash-image ), + it maintains Flash-level speed and cost efficiency while delivering + significant improvements in visual quality, prompt adherence, multi-turn + character consistency, text rendering, and wide and panoramic aspect ratio + generation (1:4, 4:1, 1:8, 8:1) across 1K, 2K, and 4K + resolutions. + See the Image generation guide to get + Deprecation announcement: The gemini-3.1-flash-image model is + deprecated and will be shut down on + October 29, 2026. Migrate to + gemini-nano-banana-2.1.
-
Google Gemini deprecations changed
Tracking since Jun 25, 2026 · 22 snapshots on file
Multiple preview models now have shutdown dates set to November 17, 2026, and new Nano Banana models section added with updated migration paths for image and TTS models.
- gemini-3.1-flash-tts-preview, gemini-3.1-flash-live-preview, and gemini-2.5-flash-native-audio-preview-12-2025 shutdown date changed to November 17, 2026
- New gemini-nano-banana-2.1 model added with release October 6, 2026
- Image models now migrate to gemini-nano-banana-2.1 instead of previous migration targets
Full history & current snapshot →View raw diff +17 −13
- | gemini-3.1-flash-image | May 28, 2026 | No shutdown date announced | | - | gemini-3.1-flash-tts-preview | February 26, 2026 | No shutdown date announced | gemini-3.8-flash-tts or gemini-3.8-flash-lite-tts | - | gemini-3.1-flash-image-preview | February 26, 2026 | June 25, 2026 | gemini-3.1-flash-image | - | gemini-2.5-flash-image | October 2, 2025 | October 2, 2026 | gemini-3.1-flash-image-preview | - | gemini-2.5-flash-image-preview | May 7, 2025 | January 15, 2026 | gemini-2.5-flash-image | - | gemini-2.0-flash-preview-image-generation | May 7, 2025 | November 14, 2025 | gemini-2.5-flash-image | - | gemini-3.1-flash-live-preview | March 11, 2026 | No shutdown date announced | gemini-3.8-live | - | gemini-2.5-flash-native-audio-preview-12-2025 | December 12, 2025 | No shutdown date announced | gemini-3.8-live | - | gemini-2.5-flash-preview-tts | May 20, 2025 | No shutdown date announced | gemini-3.8-flash-tts or gemini-3.8-flash-lite-tts | - | gemini-2.5-pro-preview-tts | May 20, 2025 | No shutdown date announced | gemini-3.8-flash-tts or gemini-3.8-flash-lite-tts | - | imagen-4.0-generate-001 | June 24, 2025 | August 17, 2026 | gemini-3.1-flash-image | - | imagen-4.0-ultra-generate-001 | June 24, 2025 | August 17, 2026 | gemini-3.1-flash-image | - | imagen-4.0-fast-generate-001 | June 24, 2025 | August 17, 2026 | gemini-3.1-flash-image | + | gemini-3.1-flash-tts-preview | February 26, 2026 | November 17, 2026 | gemini-3.8-flash-tts or gemini-3.8-flash-lite-tts | + | gemini-3.1-flash-live-preview | March 11, 2026 | November 17, 2026 | gemini-3.8-live | + | gemini-2.5-flash-native-audio-preview-12-2025 | December 12, 2025 | November 17, 2026 | gemini-3.8-live | + | gemini-2.5-flash-preview-tts | May 20, 2025 | November 17, 2026 | gemini-3.8-flash-tts or gemini-3.8-flash-lite-tts | + | gemini-2.5-pro-preview-tts | May 20, 2025 | November 17, 2026 | gemini-3.8-flash-tts or gemini-3.8-flash-lite-tts | + Nano Banana models + | gemini-nano-banana-2.1 | October 6, 2026 | No shutdown date announced | | + | gemini-3.1-flash-lite-image | June 30, 2026 | No shutdown date announced | | + | gemini-3.1-flash-image | May 28, 2026 | October 29, 2026 | gemini-nano-banana-2.1 | + | gemini-2.5-flash-image | October 2, 2025 | March 15, 2027 | gemini-3.1-flash-lite-image | + | gemini-3.1-flash-image-preview | February 26, 2026 | June 25, 2026 | gemini-nano-banana-2.1 | + | gemini-2.5-flash-image-preview | May 7, 2025 | January 15, 2026 | gemini-nano-banana-2.1 | + | gemini-2.0-flash-preview-image-generation | May 7, 2025 | November 14, 2025 | gemini-nano-banana-2.1 | + | imagen-4.0-generate-001 | June 24, 2025 | August 17, 2026 | gemini-nano-banana-2.1 | + | imagen-4.0-ultra-generate-001 | June 24, 2025 | August 17, 2026 | gemini-nano-banana-2.1 | + | imagen-4.0-fast-generate-001 | June 24, 2025 | August 17, 2026 | gemini-nano-banana-2.1 | + | gemini-omni-flash-preview | June 30, 2026 | October 22, 2026 | gemini-omni-1.1-flash |
-
Tracking since Jun 25, 2026 · 7 snapshots on file
Free plan credit description updated to reference 'more powerful models' instead of 'most powerful models'
Free · credit_allocation$20 towards most powerful models→$20 towards more powerful modelsFull history & current snapshot →View raw diff +1 −1
- $20 towards most powerful models + $20 towards more powerful models
-
Mistral models & deprecations changed
Tracking since Jun 25, 2026 · 12 snapshots on file
Mistral Large 4 model added with version v26.10
- Mistral Large 4
- v26.10
Full history & current snapshot →View raw diff +3 −0
+ Mistral Large 4 + Mistral Large 4 + v26.10
-
Google Gemini API pricing changed
Tracking since Jun 25, 2026 · 22 snapshots on file
Added image generation capabilities to Gemini 1.5 Pro and Flash with separate image output pricing
Gemini 1.5 Pro with images added: $1.50 input, $7.50 text output / $30.00 image output per 1M tokensGemini 1.5 Flash with images added: $0.75 input, $3.75 text output / $15.00 image output per 1M tokensFull history & current snapshot →View raw diff +16 −0
+ $1.50 (text/image/video) + $7.50 (text and thinking) + Equivalent to $0.0336 per 1K image*, + $0.0504 per 2K image*, + and $0.0756 per 4K image*. + $0.75 (text, image, video) + $3.75 (text and thinking) + Equivalent to $0.0168 per 1K image*, + $0.0252 per 2K image*, + and $0.0378 per 4K image*. + Image output is priced at $30 per 1,000,000 tokens (Standard) and + $15 per 1,000,000 tokens (Batch). Output images at 1K (1024x1024px) consume 1120 + tokens and are equivalent to $0.0336 per image. Output images at 2K + (2048x2048px) consume 1680 tokens and are equivalent to $0.0504 per image. + Output images at 4K (4096x4096px) consume 2520 tokens and are equivalent to + $0.0756 per image.
-
Tracking since Jun 25, 2026 · 280 snapshots on file
ChatGPT Business and Enterprise subscription plans removed from API pricing page listing
ChatGPT Business and Enterprise subscriptions removed from API pricing pageFull history & current snapshot →View raw diff +43 −214
- ChatGPT Business(opens in a new window) - ChatGPT Enterprise(opens in a new window) - ChatGPT for Education(opens in a new window) - Image 1 - Business - A secure workspace with company context and flexible seat types for any budget. - Get started(opens in a new window) - Standard seat - $20/ month - $20/month if billed annually. $25/month if billed monthly. - Premium seat - $100/ month - 5x more usage than standard, with no 5-hour limit - $100/month if billed annually. $125/month if billed monthly. - No training on your business data by default - Mix and match seat types - For teams of 2–200 employees. Unlimited subject to abuse guardrails. Learn more(opens in a new window) - Image 2 - Enterprise - Enterprise-grade AI, security, and support for businesses operating at scale - Contact our sales team to discuss enterprise pricing.* - Enterprise-level security and controls, including SCIM, EKM, user analytics, domain verification, and role-based access controls - Advanced data privacy with custom data retention policies, encryption at rest and in transit, and no training on your business data by default. Learn more - *Credit-based(opens in a new window) pricing and token-based(opens in a new window) pricing are available for Enterprise plans. - Not sure whether Business or Enterprise fits? - Ask about seats, security, deployment, or purchasing. - Compare features across plans - Try ChatGPT Business(opens in a new window) - Enterprise - Unlimited*Plan: Business, Feature: Everyday text chats, Unlimited* - Unlimited*Plan: Enterprise, Feature: Everyday text chats, Unlimited* - Unlimited*Plan: Business, Feature: Chat history, Unlimited* - Unlimited*Plan: Enterprise, Feature: Chat history, Unlimited* - Plan: Business, Feature: Access on web, iOS, Android, Yes - Plan: Enterprise, Feature: Access on web, iOS, Android, Yes - FlexiblePlan: Business, Feature: GPT-6.1 Sol, Flexible - FlexiblePlan: Enterprise, Feature: GPT-6.1 Sol, Flexible - FlexiblePlan: Business, Feature: GPT-6 Astra, Flexible - FlexiblePlan: Enterprise, Feature: GPT-6 Astra, Flexible - FlexiblePlan: Business, Feature: GPT-6 Sol, Flexible - FlexiblePlan: Enterprise, Feature: GPT-6 Sol, Flexible - FlexiblePlan: Business, Feature: GPT-6 Luna, Flexible - FlexiblePlan: Enterprise, Feature: GPT-6 Luna, Flexible - FlexiblePlan: Business, Feature: GPT-5.6 Sol, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Sol, Flexible - GPT-5.6 Sol Pro - FlexiblePlan: Business, Feature: GPT-5.6 Sol Pro, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Sol Pro, Flexible - FlexiblePlan: Business, Feature: GPT-5.6 Terra, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Terra, Flexible - FlexiblePlan: Business, Feature: GPT-5.6 Luna, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Luna, Flexible - FlexiblePlan: Business, Feature: GPT-5 Thinking Mini, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5 Thinking Mini, Flexible - Plan: Business, Feature: Legacy models, Yes - Plan: Enterprise, Feature: Legacy models, Yes - Fast Plan: Business, Feature: Response times, Fast - Fastest Plan: Enterprise, Feature: Response times, Fastest - 54K Plan: Business, Feature: GPT Instant total context window, 54K - 128K Plan: Enterprise, Feature: GPT Instant total context window, 128K
-
Tracking since Jun 25, 2026 · 10 snapshots on file
IBM Granite model ibm-granite/granite-4.0-micro-base with Apache 2.0 license was removed as of 2025-09-16.
- Model: ibm-granite/granite-4.0-micro-base
- License: apache-2.0
- Date: 2025-09-16
Full history & current snapshot →View raw diff +0 −1
- ibm-granite/granite-4.0-micro-base license:apache-2.0 2025-09-16 -
xAI (Grok) rate limits changed
Tracking since Jun 25, 2026 · 17 snapshots on file
RPS calculation changed from RPM/60 to RPM/48 with minimum of 2 RPS floor.
Full history & current snapshot →View raw diff +1 −1
- Every xAI API team has per-model rate limits on two dimensions: requests per second (RPS) and tokens per minute (TPM). Your per-second limit is derived from your per-minute request budget (RPM / 60): you cannot spend a full minute's requests in a single second, which protects the API from sudden bursts. These limits scale with your team's tier, which is determined by cumulative spend on the API. + Every xAI API team has per-model rate limits on two dimensions: requests per second (RPS) and tokens per minute (TPM). Your per-second limit is derived from your per-minute request budget (RPM / 48, with a minimum of 2): you cannot spend a full minute's requests in a single second, which protects the API from sudden bursts. These limits scale with your team's tier, which is determined by cumulative spend on the API.
-
Tracking since Jun 25, 2026 · 57 snapshots on file
Anthropic publishes Start tier rate limits with model-specific RPM/ITPM/OTPM; cache-aware ITPM counts only uncached tokens; Build/Scale/Custom tier limits dashboard-only.
Full history & current snapshot →View raw diff +2 −0
+ Start tier + Start tier
-
Framework adapters in @supabase/server are deprecated
The Hono, H3, Elysia, and NestJS adapters in @supabase/server are deprecated and will be removed on December 1, 2026. Move to the framework bridges in the frameworks guide. What is being deprecated The four framework adapters that ship inside @supabase/server: @supabase/server/adapters/hono @supabase/server/adapters/h3 @supabase/server/adapters/elysia @supabase/server/adapters/nestjs They still work today. They will be removed from the package on December 1, 2026. No new adapters are accepted. N
Full history & current snapshot → -
Meta reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Llama 4 Maverick: 11 benchmark claims (via web search)
Full history & current snapshot → -
Tracking since Jun 25, 2026 · 8 snapshots on file
New Google Gemma model variant DiarizationLM-Gemma-4-E4B-v1 released with Apache 2.0 license.
- Model: google/DiarizationLM-Gemma-4-E4B-v1
- License: apache-2.0
- Release date: 2026-10-04
Full history & current snapshot →View raw diff +1 −0
+ google/DiarizationLM-Gemma-4-E4B-v1 license:apache-2.0 2026-10-04 -
Moonshot AI reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Kimi K3: 12 benchmark claims (via web search)
Full history & current snapshot → -
Zhipu AI (Z.ai) reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
GLM-5.3: 10 benchmark claims (via web search)
Full history & current snapshot → -
Alibaba reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Qwen3.8-Max: 12 benchmark claims (via web search)
Full history & current snapshot → -
DeepSeek reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
DeepSeek-V4.1-Flash: 11 benchmark claims (via web search)
Full history & current snapshot → -
xAI reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Grok 4.7: 10 benchmark claims (via web search)
Full history & current snapshot → -
Google reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Gemini 3.1 Pro: 9 benchmark claims (via web search)
Full history & current snapshot → -
OpenAI reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
GPT-5: 6 benchmark claims (via web search)
Full history & current snapshot → -
Anthropic reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Claude Opus 5.5: 15 benchmark claims (via web search)
Full history & current snapshot → -
Tracking since Jun 25, 2026 · 57 snapshots on file
GitHub API: 1 additive change
Full history & current snapshot →View raw diff +1 −0
+ New endpoint: POST /repos/{owner}/{repo}/pulls/{pull_number}/requested_reviewers/rerequest -
Tracking since Jun 25, 2026 · 26 snapshots on file
Arena Elo (text, overall) updated.
Full history & current snapshot →View raw diff +24 −24
- claude-opus-4-6-high 1504 - claude-opus-4-7 1484 - claude-opus-5-high 1503 - deepseek-v3.2 1425 - deepseek-v3.2-exp-thinking 1425 - deepseek-v4.1-flash-max 1461 - dola-seed-2.0-pro 1449 - gemini-3.5-flash-high 1481 - gemini-3.5-flash-lite 1434 - gemini-3.6-flash-high 1480 - gemini-3.8-flash-high 1496 - glm-5.3-max 1472 - glm-5v-turbo 1438 - gpt-6-astra-max 1441 - grok-3-preview-02-24 1426 - grok-4.20-multi-agent-beta-0309 1451 - hy3 1442 - mimo-v2.6-flash 1458 - mistral-large-3 1427 - muse-spark-1.1 1480 - muse-spark-1.2 (xHigh) 1484 - qwen3.5-max-preview 1471 - qwen3.7-max-preview 1475 - qwen3.8-max 1481 + claude-opus-4-6-high 1503 + claude-opus-4-7 1483 + claude-opus-5-high 1502 + claude-sonnet-5.5-xhigh 1467 + deepseek-v4.1-flash-max 1462 + dola-seed-2.0-pro 1448 + gemini-3.5-flash-high 1482 + gemini-3.5-flash-lite 1435 + gemini-3.6-flash-high 1479 + gemini-3.8-flash-high 1497 + glm-5.3-max 1471 + glm-5v-turbo 1437 + gpt-6-astra-max 1442 + gpt-6.1-sol-max 1446 + grok-4.20-multi-agent-beta-0309 1450 + hy3 1441 + mimo-v2.6-flash 1456 + mistral-large-3 1428 + muse-spark-1.1 1479 + muse-spark-1.2 (xHigh) 1483 + qwen3.5-max-preview 1470 + qwen3.7-max-preview 1476 + qwen3.8-max 1482 + Step 5 Preview 1448
-
Tracking since Jun 25, 2026 · 15 snapshots on file
Removed Microsoft Phi model Dayhoff-170M-UR90-46000 with MIT license.
- microsoft/Dayhoff-170M-UR90-46000
- license:mit
- 2026-01-24
Full history & current snapshot →View raw diff +0 −1
- microsoft/Dayhoff-170M-UR90-46000 license:mit 2026-01-24 -
Tracking since Jun 25, 2026 · 104 snapshots on file
OpenAI API: 1 additive change
Full history & current snapshot →View raw diff +1 −0
+ New endpoint: GET /agents/sessions/{session_id}/turns/{turn_id}/items -
Selected models in GitHub Copilot deprecated
As of today, October 2, 2026, we have deprecated the following models across all GitHub Copilot experiences (including Copilot Chat, inline edits, ask and agent modes, and code completions). Model… The post Selected models in GitHub Copilot deprecated appeared first on The GitHub Blog.
Full history & current snapshot → -
Tracking since Jun 25, 2026 · 5 snapshots on file
Model changed from MolmoWeb-4B-Native to AstaBrief_8B_SFT with updated license date from 2026-03-23 to 2026-09-10.
- Model replaced: MolmoWeb-4B-Native → AstaBrief_8B_SFT
- License date updated: 2026-03-23 → 2026-09-10
- Apache 2.0 license maintained
Full history & current snapshot →View raw diff +1 −1
- allenai/MolmoWeb-4B-Native license:apache-2.0 2026-03-23 + allenai/AstaBrief_8B_SFT license:apache-2.0 2026-09-10
-
Tracking since Jun 25, 2026 · 14 snapshots on file
Pricing documentation section removed from API page
Pricing information section removed from documentationFull history & current snapshot →View raw diff +1 −22
- The prices listed below are in units of per 1M tokens. A token, the smallest unit of text that the model recognizes, can be a word, a number, or even a punctuation mark. We will bill based on the total number of input and output tokens by the model. - DeepSeek-V4-Pro-0813 - 1M INPUT TOKENS - $0.003 - $0.022 - $0.006 - $0.044 - 1M INPUT TOKENS - $0.15 - $0.66 - $0.3 - $1.32 - 1M OUTPUT TOKENS - $0.6 - $1.98 - $1.2 - $3.96 - Concurrency Limit(3) - (3) For more details on concurrency limits, please refer to Rate Limit & Isolation. - The expense = number of tokens × price. - Product prices may vary and DeepSeek reserves the right to adjust them. We recommend topping up based on your actual usage and regularly checking this page for the most recent pricing information. - Token & Token Usage + -H "Authorization: Bearer ${DEEPSEEK_API_KEY}" \
-
Tracking since Jun 25, 2026 · 21 snapshots on file
Multiple OpenAI models deprecated with specified removal dates in 2027 and recommended migration paths.
- GPT-5.3-Codex, GPT-5.1, and GPT-5.4-Nano deprecated April 1, 2027; migrate to GPT-6-Sol or GPT-6-Luna
- Text-to-speech models tts-1, tts-1-hd, gpt-4o-mini-tts-2025-03-20, and gpt-4o-mini-tts-2025-12-15 deprecated January 6, 2027; migrate to gpt-realtime-2.1-mini
Full history & current snapshot →View raw diff +11 −0
+ 2026-10-01: GPT-5.3-Codex, GPT-5.1, GPT-5.4-Nano + The following models are deprecated and will be removed from the API on April 1, 2027, with six months’ notice. Migrate to the recommended replacements before the shutdown date. + | Apr 1, 2027 | gpt-5.3-codex | gpt-6-sol | + | Apr 1, 2027 | gpt-5.4-nano | gpt-6-luna | + | Apr 1, 2027 | gpt-5.1 | gpt-6-sol | + 2026-10-01: Text-to-speech models + The following text-to-speech models are deprecated and will be removed from the API on January 6, 2027, with at least three months’ notice. Migrate to gpt-realtime-2.1-mini before the shutdown date. See the Realtime API guide to plan your migration. + | Jan 6, 2027 | tts-1 | gpt-realtime-2.1-mini | + | Jan 6, 2027 | tts-1-hd | gpt-realtime-2.1-mini | + | Jan 6, 2027 | gpt-4o-mini-tts-2025-03-20 | gpt-realtime-2.1-mini | + | Jan 6, 2027 | gpt-4o-mini-tts-2025-12-15 | gpt-realtime-2.1-mini |
-
GPQA Diamond (Epoch) scores changed
Tracking since Jun 25, 2026 · 31 snapshots on file
GPQA Diamond accuracy (0–1) updated.
Full history & current snapshot →View raw diff +1 −0
+ grok-4.7_xhigh 0.927 -
NVIDIA Nemotron models changed
Tracking since Jun 25, 2026 · 27 snapshots on file
NVIDIA-Nemotron-3-Ultra-550B-A55B-Base-BF16 model removed from availability (was licensed as other, scheduled expiration 2026-06-03).
- nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-Base-BF16 deleted
- license type was other
- expiration date 2026-06-03
Full history & current snapshot →View raw diff +0 −1
- nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-Base-BF16 license:other 2026-06-03 -
xAI (Grok) rate limits changed
Tracking since Jun 25, 2026 · 17 snapshots on file
Added grok-imagine-video-1.5-lite video generation model with 10/20/39/79/158 RPS tiers (T0-T4).
Full history & current snapshot →View raw diff +1 −0
+ | grok-imagine-video-1.5-lite | T 0 T 1 T 2 T 3 T 4 | 10 20 39 79 158 | ————— | -
Google Gemini API changelog changed
Tracking since Jun 25, 2026 · 25 snapshots on file
Updated documentation URL for Gemini 3.1 Flash Image Preview and Gemini 2.5 Flash Native Audio model reference.
- Gemini 3.1 Flash Image Preview URL changed from gemini-3.1-flash-image-preview to gemini-3.1-flash-image
- Gemini 2.5 Flash Native Audio URL changed to gemini-2.5-flash-native-audio-preview-12-2025
Full history & current snapshot →View raw diff +3 −2
- Launched Nano Banana 2, Gemini 3.1 Flash Image Preview, a high-efficiency - Released gemini-2.5-flash-native-audio-preview-12-2025, a new native audio model for the Live API. This update improves the model's ability to handle complex workflows. To learn more, see the Live API guide and Gemini 2.5 Flash Native Audio. + Launched Nano Banana 2, [Gemini 3.1 Flash Image + Preview](https://ai.google.dev/gemini-api/docs/models/gemini-3.1-flash-image), a high-efficiency + Released gemini-2.5-flash-native-audio-preview-12-2025, a new native audio model for the Live API. This update improves the model's ability to handle complex workflows. To learn more, see the Live API guide and Gemini 2.5 Flash Native Audio.
-
xAI (Grok) API pricing changed
Tracking since Jun 25, 2026 · 12 snapshots on file
Real-time Audio pricing reduced 60% from $0.05/sec to $0.02/sec
Real-time Audio rate decreased from $0.05/sec to $0.02/secFull history & current snapshot →View raw diff +1 −1
- Starting at $0.05 / sec + Starting at $0.02 / sec
-
API deprecations - Cloudflare Radar: Top browsers and NetFlows summary endpoints
Deprecation date: October 2, 2026 End of life date: April 2, 2027 The Radar top browsers endpoints and the NetFlows summary endpoint without a dimension are deprecated and will be replaced by the corresponding summary endpoints with a {dimension} path parameter. Deprecated APIs: GET /radar/http/top/browser GET /radar/http/top/browser_family GET /radar/netflows/summary Replacements: Get HTTP summary by dimension — GET /radar/http/summary/{dimension} with the BROWSER or BROWSER_FAMILY dimension Ge
Full history & current snapshot → -
GitHub Actions: macOS 14 runner image retirement
The macOS 14 runner image will be retired on November 2, 2026. To raise awareness of the upcoming removal, jobs using macOS 14 will temporarily fail during the following scheduled… The post GitHub Actions: macOS 14 runner image retirement appeared first on The GitHub Blog.
Full history & current snapshot → -
Tracking since Jun 25, 2026 · 18 snapshots on file
DigitalOcean API: 18 additive changes
Full history & current snapshot →View raw diff +18 −0
+ New endpoint: GET /v2/insights/alert-instances + New endpoint: GET /v2/insights/alert-instances/{id} + New endpoint: GET /v2/insights/alert-rules + New endpoint: POST /v2/insights/alert-rules + New endpoint: GET /v2/insights/alert-rules/{id} + New endpoint: PUT /v2/insights/alert-rules/{id} + New endpoint: DELETE /v2/insights/alert-rules/{id} + New endpoint: GET /v2/insights/notification-channels + New endpoint: POST /v2/insights/notification-channels + New endpoint: GET /v2/insights/notification-channels/{id} + New endpoint: PUT /v2/insights/notification-channels/{id} + New endpoint: DELETE /v2/insights/notification-channels/{id} + New endpoint: GET /v2/insights/query/{region}/prom/api/v1/query + New endpoint: GET /v2/insights/query/{region}/prom/api/v1/query_range + New endpoint: GET /v2/insights/query/{region}/prom/api/v1/labels + New endpoint: GET /v2/insights/query/{region}/prom/api/v1/label/{name}/values + New endpoint: GET /v2/insights/query/{region}/prom/api/v1/series + New endpoint: POST /v2/insights/query/{region}/logs/search
-
Speed Insights deprecates First Input Delay on November 1st
is deprecating First Input Delay (FID) and will stop recording new measurements on November 1st.Speed Insights Speed Insights already uses, the Core Web Vital that replaced FID, as the responsiveness metric in your. The deprecation won’t require changes on your side, affect your score, or prevent you from viewing existing FID data.Interaction to Next Paint (INP)Real Experience Score Learn more in the.Speed Insights metrics documentation Read more
Full history & current snapshot → -
Tracking since Jun 25, 2026 · 26 snapshots on file
Arena Elo (text, overall) updated.
Full history & current snapshot →View raw diff +34 −34
- claude-fable-5-high 1491 - claude-opus-4-8 1453 - claude-opus-4-8-high 1460 - claude-opus-5-high 1504 - claude-opus-5-max 1506 - claude-opus-5.5-high 1518 - deepseek-v3.2 1424 - deepseek-v4.1-flash-max 1465 - dola-seed-2.0-pro 1448 - gemini-3.5-flash-lite 1436 - gemini-3.5-flash-medium 1476 - gemini-3.6-flash-high 1478 - gemini-3.8-flash-high 1494 - glm-4.7 1435 - glm-5.1 1462 - glm-5.3-flash 1471 - glm-5v-turbo 1437 - gpt-5.6-luna-xhigh 1432 - gpt-6-astra-max 1443 - grok-4.1-thinking 1436 - grok-4.20-beta1 1445 - grok-4.20-multi-agent-beta-0309 1450 - grok-4.5 1447 - inkling 1443 - mimo-v2.5 1427 - mimo-v2.6-flash 1459 - minimax-m3 1433 - mistral-medium-2508 1424 - muse-spark-1.1 1478 - muse-spark-1.2 (xHigh) 1486 - nvidia-nemotron-3-ultra-550b-a55b-nvfp4 1444 - qwen3.7-max-preview 1476 - qwen3.7-plus 1454 - qwen3.8-max 1480 + claude-fable-5-high 1492 + claude-opus-4-8 1454 + claude-opus-4-8-high 1461 + claude-opus-5-high 1503 + claude-opus-5-max 1507 + claude-opus-5.5-high 1512 + deepseek-v3.2 1425 + deepseek-v4.1-flash-max 1461 + dola-seed-2.0-pro 1449 + gemini-3.5-flash-lite 1434 + gemini-3.5-flash-medium 1478 + gemini-3.6-flash-high 1480 + gemini-3.8-flash-high 1496 + gemini-4-argon-high 1533 + glm-4.7 1436 + glm-5.1 1461 + glm-5.3-flash 1470 + glm-5v-turbo 1438 + gpt-5.6-luna-xhigh 1431 + gpt-6-astra-max 1441 + grok-4.1-thinking 1437 + grok-4.20-beta1 1444 + grok-4.20-multi-agent-beta-0309 1451 + grok-4.5 1448 + inkling 1442 + mimo-v2.5 1428
-
Tracking since Jun 25, 2026 · 13 snapshots on file
Tier naming and structure changed from Build/Launch/Grow to numbered Tiers 1-5; model-specific rate limits removed from public documentation.
now$100/mo$50$1,000/mo$250$1,000was$500Full history & current snapshot →View raw diff +6 −14
- The three paid usage tiers are Build, Launch, and Grow. Your organization’s usage tier upgrades automatically as its total credit purchases reach each threshold. Higher tiers generally provide higher rate limits across models. - | Build | $5 in total credit purchases | $500 / month | - | Launch | $100 in total credit purchases | $5,000 / month | - | Grow | $500 in total credit purchases | $200,000 / month | - Rate limits by usage tier - To view the limits for each model at your usage tier, go to Settings > Organization > Limits and review Rate limits. To upgrade your usage tier, select Upgrade tier in the Usage Tiers section. - | Tier | Model | RPM | TPM | - | --- | --- | --- | --- | - | Build | Astra, Sol, Terra | 5,000 | 1,000,000 | - | Build | Luna | 5,000 | 2,000,000 | - | Launch | Astra, Sol, Terra | 10,000 | 4,000,000 | - | Launch | Luna | 10,000 | 10,000,000 | - | Grow | Astra, Sol, Terra | 15,000 | 40,000,000 | - | Grow | Luna | 30,000 | 180,000,000 | + You can view the rate and usage limits for your organization under the limits section of your account settings. As your spend on our API goes up, we automatically graduate you to the next usage tier. This usually results in an increase in rate limits across most models. + | Tier 1 | $5 paid | $100 / month | + | Tier 2 | $50 paid | $500 / month | + | Tier 3 | $100 paid | $1,000 / month | + | Tier 4 | $250 paid | $5,000 / month | + | Tier 5 | $1,000 paid | $200,000 / month |
-
Tracking since Jun 25, 2026 · 104 snapshots on file
OpenAI API: 3 breaking changes
Full history & current snapshot →View raw diff +0 −3
- New required body field: type - New required body field: name - New required body field: prompt
-
Google Gemini deprecations changed
Tracking since Jun 25, 2026 · 22 snapshots on file
Added new model variant gemini-2.5-computer-use-preview-10-2025 with October 7, 2025 release date and July 28, 2026 sunset date.
- Model: gemini-2.5-computer-use-preview-10-2025
- Release date: October 7, 2025
- Sunset date: July 28, 2026
Full history & current snapshot →View raw diff +1 −0
+ | gemini-2.5-computer-use-preview-10-2025 | October 7, 2025 | July 28, 2026 | gemini-3.8-flash | -
Google Gemini API pricing changed
Tracking since Jun 25, 2026 · 22 snapshots on file
Gemini 2.5 Computer Use pricing reference updated from Gemini 3.5 Flash to Gemini 3.8 Flash
model_pricing_referenceGemini 2.5 Computer Use legacy pricing reference updated to Gemini 3.8 Flash pricingFull history & current snapshot →View raw diff +1 −2
- $2.50, prompts > 200k token - Charged as regular tokens per model pricing (e.g., standard Gemini 3.5 Flash pricing). See the Gemini 2.5 Computer Use Preview pricing table for legacy model rates. + Charged as regular tokens per model pricing (e.g., standard Gemini 3.8 Flash pricing).
-
Tracking since Jun 25, 2026 · 280 snapshots on file
New GPT-6 model variants (Sol, Astra, Luna) added to Business and Enterprise plans
Analytics dashboard availability changed for Business planFull history & current snapshot →View raw diff +19 −5
- Plan: Business, Feature: Canvas, Yes - Plan: Enterprise, Feature: Canvas, Yes - Plan: Business, Feature: Excel, PowerPoint, and Google Sheets extensions, Yes - Plan: Enterprise, Feature: Excel, PowerPoint, and Google Sheets extensions, Yes - Plan: Business, Feature: Analytics dashboard, No + FlexiblePlan: Business, Feature: GPT-6.1 Sol, Flexible + FlexiblePlan: Enterprise, Feature: GPT-6.1 Sol, Flexible + FlexiblePlan: Business, Feature: GPT-6 Astra, Flexible + FlexiblePlan: Enterprise, Feature: GPT-6 Astra, Flexible + FlexiblePlan: Business, Feature: GPT-6 Sol, Flexible + FlexiblePlan: Enterprise, Feature: GPT-6 Sol, Flexible + FlexiblePlan: Business, Feature: GPT-6 Luna, Flexible + FlexiblePlan: Enterprise, Feature: GPT-6 Luna, Flexible + Business Premium Plan: Business, Feature: Dots, Business Premium + Plan: Enterprise, Feature: Dots, Yes + Plan: Business, Feature: Excel, Word, PowerPoint, and Google Sheets extensions, Yes + Plan: Enterprise, Feature: Excel, Word, PowerPoint, and Google Sheets extensions, Yes + Plan: Business, Feature: Basic user analytics, Yes + Plan: Enterprise, Feature: Basic user analytics, Yes + Plan: Business, Feature: Granular GPT controls & group permissions, No + Plan: Enterprise, Feature: Granular GPT controls & group permissions, Yes + Plan: Business, Feature: Analytics dashboard, Yes + Plan: Business, Feature: Intune for iOS, No + Plan: Enterprise, Feature: Intune for iOS, Yes
-
GPQA Diamond (Epoch) scores changed
Tracking since Jun 25, 2026 · 31 snapshots on file
GPQA Diamond accuracy (0–1) updated.
Full history & current snapshot →View raw diff +3 −0
+ gpt-6-luna_max 0.905 + gpt-6-sol_max 0.943 + gpt-6.1-sol_max 0.954
-
Tracking since Jun 25, 2026 · 104 snapshots on file
OpenAI API: 2 breaking changes · 1 additive
Full history & current snapshot →View raw diff +1 −2
- Endpoint removed: POST /evals/{eval_id}/runs/{run_id} - New required body field: model + New endpoint: POST /evals/{eval_id}/runs/{run_id}/cancel
-
Anthropic model deprecations changed
Tracking since Jun 25, 2026 · 45 snapshots on file
Claude Sonnet 4.5 model deprecated on September 30, 2026 with retirement date of November 30, 2026; claude-sonnet-4-5-20250929 will be replaced by claude-sonnet-5-5.
- Deprecation effective September 30, 2026
- Retirement date November 30, 2026
- Replacement model: claude-sonnet-5-5
Full history & current snapshot →View raw diff +4 −1
- | | Active | N/A | Not sooner than September 29, 2026 | + | | Deprecated | September 30, 2026 | November 30, 2026 | + 2026-09-30: Claude Sonnet 4.5 model + On September 30, 2026, Anthropic notified developers using Claude Sonnet 4.5 of its upcoming retirement on the Claude API. + | November 30, 2026 | claude-sonnet-4-5-20250929 | claude-sonnet-5-5 |
-
Tracking since Jun 25, 2026 · 5 snapshots on file
Free plan modified: removed $5 included credits, now emphasizes free access to v0
Free · description$5 of included monthly credits→Free access to v0Full history & current snapshot →View raw diff +1 −1
- $5 of included monthly credits + Free access to v0
-
Tracking since Jun 25, 2026 · 57 snapshots on file
Minor grammar fix: 'SDK's' changed to 'SDK' (no functional change to rate limits).
Full history & current snapshot →View raw diff +1 −1
- The error type is rate_limit_error, the same as for a rate limit, but the response has no retry-after header. Retrying, including the SDKs' automatic retries, fails until access resumes. + The error type is rate_limit_error, the same as for a rate limit, but the response has no retry-after header. Retrying, including the SDK's automatic retries, fails until access resumes.
-
Tracking since Jun 25, 2026 · 18 snapshots on file
DigitalOcean API: 2 breaking changes · 25 additive
Full history & current snapshot →View raw diff +25 −2
- Endpoint removed: GET /v2/action-gateway/tools/{name}/definition - Endpoint removed: PATCH /v2/action-gateway/connections/{id} + New endpoint: GET /v2/action-gateway/actors/{actor_id}/limits + New endpoint: POST /v2/action-gateway/actors/{actor_id}/limits:clear + New endpoint: POST /v2/action-gateway/actors/{actor_id}/limits:set + New endpoint: GET /v2/action-gateway/mcp-servers + New endpoint: POST /v2/action-gateway/mcp-servers + New endpoint: GET /v2/action-gateway/mcp-servers/{server_ref} + New endpoint: DELETE /v2/action-gateway/mcp-servers/{server_ref} + New endpoint: PATCH /v2/action-gateway/mcp-servers/{server_ref} + New endpoint: POST /v2/action-gateway/mcp-servers/{server_ref}/resync + New endpoint: GET /v2/action-gateway/mcp-servers/{server_ref}/tools + New endpoint: PUT /v2/action-gateway/mcp-servers/{server_ref}/tools + New endpoint: GET /v2/action-gateway/output-views + New endpoint: POST /v2/action-gateway/output-views + New endpoint: POST /v2/action-gateway/output-views/preview + New endpoint: GET /v2/action-gateway/output-views/{view_id} + New endpoint: DELETE /v2/action-gateway/output-views/{view_id} + New endpoint: GET /v2/action-gateway/sessions/search + New endpoint: GET /v2/action-gateway/toolbelts/search + New endpoint: GET /v2/action-gateway/toolbelts/{name}/providers + New endpoint: GET /v2/action-gateway/toolbelts/{name}/providers/{provider}/tools + New endpoint: GET /v2/action-gateway/tools/health + New endpoint: GET /v2/action-gateway/tools/health/providers + New endpoint: GET /v2/action-gateway/tools/health/tools/{tool_slug} + New endpoint: GET /v2/action-gateway/tools/providers/search + New endpoint: GET /v2/action-gateway/tools/search
-
Tracking since Jun 25, 2026 · 104 snapshots on file
OpenAI API: 1 additive change
Full history & current snapshot →View raw diff +1 −0
+ New endpoint: GET /agents/sessions/{session_id}/traces -
Tracking since Jun 25, 2026 · 280 snapshots on file
OpenAI API pricing details removed from business pricing page; Business and Enterprise plans remain
API pricing section removed from page (OpenAI API moved to separate developer pricing page)Full history & current snapshot →View raw diff +199 −43
- Business Pricing | OpenAI - Price - Input:$10.00 / 1M tokens Cached input:$1.00 / 1M tokens Output:$50.00 / 1M tokens - Price - Input:$2.00 / 1M tokens Cached input:$0.10 / 1M tokens Output:$10.00 / 1M tokens - Price - Input:$0.10 / 1M tokens Cached input:$0.01 / 1M tokens Output:$0.50 / 1M tokens - Price - Image:$8.00 / 1M tokens for inputs$2.00 / 1M tokens for cached inputs$30.00 / 1M tokens for outputs Text:$5.00 / 1M tokens for inputs$1.25 / 1M tokens for cached inputs - Price - Image:$8.00 / 1M tokens for inputs$2.00 / 1M tokens for cached inputs$30.00 / 1M tokens for outputs Text:$5.00 / 1M tokens for inputs$1.25 / 1M tokens for cached inputs - Price - $0.05 per minute / $0.00083 per second - Price - Audio:$32.00 / 1M tokens for inputs$0.40 / 1M tokens for cached inputs$64.00 / 1M tokens for outputs Text:$4.00 / 1M tokens for inputs$0.40 / 1M tokens for cached inputs$24.00 / 1M tokens for outputs Image:$5.00 / 1M tokens for inputs$0.50 / 1M tokens for cached inputs - Price - Audio:$10.00 / 1M tokens for inputs$0.30 / 1M tokens for cached inputs$20.00 / 1M tokens for outputs Text:$0.60 / 1M tokens for inputs$0.06 / 1M tokens for cached inputs$2.40 / 1M tokens for outputs Image:$0.80 / 1M tokens for inputs$0.08 / 1M tokens for cached inputs - Price - $0.017 per minute / $0.00028 per second - Price - $0.0045 per minute / $0.00008 per second - Price - $0.034 per minute / $0.00057 per second - Price - $10.00 / 1k calls Search content tokens are free. - Price - Now:1 GB for $0.03 / 64GB for $1.92 per container Starting March 31, 2026:1 GB for $0.03 / 64GB for $1.92 per 20-minute session per container - Provides lower costs for requests in exchange for slower response times and occasional resource unavailability. Ideal for non-production or lower priority tasks. - Enterprise offerings - Contact our sales team to learn more about Data residency(opens in a new window), Scale Tier andReserved Capacity designed for cutting-edge customers running larger workloads. - We recommend experimenting with all of these models in the Playground(opens in a new window) to explore which models provide the best price performance trade-off for your usage. - Do you offer an enterprise package or SLAs? - We offer different tiers of access to our enterprise customers that include SLAs, lower latency, and more. Please contact our sales team to learn more. - Yes, we treat Playground usage the same as regular API usage. You will be billed at the per-token input and output prices mentioned above. - How will I know how many tokens I’ve used each month? - A token is a mathematical representation of natural language. Log in to your account to view your usage tracking dashboard(opens in a new window). This dashboard will show you how many tokens you’ve used during the current and past billing cycles. - You can set a monthly budget in your billing settings(opens in a new window), after which we’ll stop serving your requests. There may be a delay in enforcing the limit, and you are responsible for any overage incurred. You can also configure an email notification threshold to receive an email alert once you cross that threshold each month. We recommend checking your usage tracking dashboard(opens in a new window) regularly to monitor your spend. - Is access to the API included in ChatGPT Plus, Business, Enterprise or Edu? - No, OpenAI APIs are billed separately from ChatGPT Plus, Business, Enterprise and Edu. ChatGPT subscription pricing can be found at openai.com/chatgpt/pricing/. - Images are converted into tokens and charged per token. Text models price image tokens at standard text token rates, while GPT Image and gpt-realtime uses a separate image token rate. Models like gpt-4.1-mini, gpt-4.1-nano, and o4-mini convert images into tokens differently. Learn more in our docs(opens in a new window). - ChatGPT Business(opens in a new window) - ChatGPT Enterprise(opens in a new window) - ChatGPT for Education(opens in a new window) + Image 1 + Business + A secure workspace with company context and flexible seat types for any budget. + Get started(opens in a new window) + Standard seat + $20/ month + $20/month if billed annually. $25/month if billed monthly. + Premium seat + $100/ month + 5x more usage than standard, with no 5-hour limit + $100/month if billed annually. $125/month if billed monthly. + No training on your business data by default + Mix and match seat types + For teams of 2–200 employees. Unlimited subject to abuse guardrails. Learn more(opens in a new window) + Image 2 + Enterprise + Enterprise-grade AI, security, and support for businesses operating at scale
-
Mistral AI reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Mistral Large 3: 7 benchmark claims (via web search)
Full history & current snapshot → -
Tracking since Jun 25, 2026 · 8 snapshots on file
Removed google/t5gemma-xl-xl-prefixlm-it model with gemma license from available models as of 2025-06-19.
- Model: google/t5gemma-xl-xl-prefixlm-it
- License: gemma
- Date: 2025-06-19
Full history & current snapshot →View raw diff +0 −1
- google/t5gemma-xl-xl-prefixlm-it license:gemma 2025-06-19 -
Mistral models & deprecations changed
Tracking since Jun 25, 2026 · 12 snapshots on file
Model names and version information updated with new endpoint identifiers and expiration dates added.
- Z.ai GLM 5.2 now includes endpoint zai-glm-5-2 with dates 9/29/2026 to 10/31/2026
- Leanstral 1.5 added as new model with endpoint labs-leanstral-1-5 and dates 9/29/2026 to 9/30/2026
- OCR 4.0 now has endpoint mistral-ocr-4-0 with dates 9/29/2026 to 9/30/2026
Full history & current snapshot →View raw diff +3 −10
- Z.ai GLM 5.2 - v5.2 - OCR 4.0 - Our latest OCR service with paragraph-level bounding boxes and structural block labels. - v4.0 - Other specialist models - Copy section linkOther specialist models - Specialized models for focused domains and task-specific workloads. - Updated code agent for Lean 4 formal proof engineering and automated theorem proving. - v1.5 + Z.ai GLM 5.2 ↗5.2zai-glm-5-29/29/202610/31/2026 + Leanstral 1.5 ↗1.5labs-leanstral-1-59/29/20269/30/2026 + OCR 4.0 ↗4.0mistral-ocr-4-09/29/20269/30/2026
-
Google Gemini deprecations changed
Tracking since Jun 25, 2026 · 22 snapshots on file
Three veo-3.1 preview models now have announced shutdown dates of October 22, 2026 and specified migration path to gemini-omni-1.1-flash.
- veo-3.1-lite-generate-preview shutdown date changed from 'No shutdown date announced' to October 22, 2026
- veo-3.1-generate-preview shutdown date changed from 'No shutdown date announced' to October 22, 2026
- veo-3.1-fast-generate-preview shutdown date changed from 'No shutdown date announced' to October 22, 2026
Full history & current snapshot →View raw diff +3 −5
- | veo-3.1-lite-generate-preview | March 31, 2026 | No shutdown date announced | | - | veo-3.1-generate-preview | October 15, 2025 | No shutdown date announced | | - | veo-3.1-fast-generate-preview | October 15, 2025 | No shutdown date announced | | - | Deprecated models |||| - | gemini-omni-flash-preview | June 30, 2026 | September 30, 2026 | gemini-omni-1.1-flash | + | veo-3.1-lite-generate-preview | March 31, 2026 | October 22, 2026 | gemini-omni-1.1-flash | + | veo-3.1-generate-preview | October 15, 2025 | October 22, 2026 | gemini-omni-1.1-flash | + | veo-3.1-fast-generate-preview | October 15, 2025 | October 22, 2026 | gemini-omni-1.1-flash |
-
Tracking since Jun 25, 2026 · 280 snapshots on file
ChatGPT Business/Enterprise subscription plan tiers removed from API pricing documentation
Removed ChatGPT Business and Enterprise subscription plan details from API pricing pageFull history & current snapshot →View raw diff +43 −199
- ChatGPT Business(opens in a new window) - ChatGPT Enterprise(opens in a new window) - ChatGPT for Education(opens in a new window) - Image 1 - Business - A secure workspace with company context and flexible seat types for any budget. - Get started(opens in a new window) - Standard seat - $20/ month - $20/month if billed annually. $25/month if billed monthly. - Premium seat - $100/ month - 5x more usage than standard, with no 5-hour limit - $100/month if billed annually. $125/month if billed monthly. - No training on your business data by default - Mix and match seat types - For teams of 2–200 employees. Unlimited subject to abuse guardrails. Learn more(opens in a new window) - Image 2 - Enterprise - Enterprise-grade AI, security, and support for businesses operating at scale - Contact our sales team to discuss enterprise pricing.* - Enterprise-level security and controls, including SCIM, EKM, user analytics, domain verification, and role-based access controls - Advanced data privacy with custom data retention policies, encryption at rest and in transit, and no training on your business data by default. Learn more - *Credit-based(opens in a new window) pricing and token-based(opens in a new window) pricing are available for Enterprise plans. - Looking for personal plans? - Compare features across plans - Try ChatGPT Business(opens in a new window) - Enterprise - Unlimited*Plan: Business, Feature: Everyday text chats, Unlimited* - Unlimited*Plan: Enterprise, Feature: Everyday text chats, Unlimited* - Unlimited*Plan: Business, Feature: Chat history, Unlimited* - Unlimited*Plan: Enterprise, Feature: Chat history, Unlimited* - Plan: Business, Feature: Access on web, iOS, Android, Yes - Plan: Enterprise, Feature: Access on web, iOS, Android, Yes - FlexiblePlan: Business, Feature: GPT-5.6 Sol, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Sol, Flexible - GPT-5.6 Sol Pro - FlexiblePlan: Business, Feature: GPT-5.6 Sol Pro, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Sol Pro, Flexible - FlexiblePlan: Business, Feature: GPT-5.6 Terra, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Terra, Flexible - FlexiblePlan: Business, Feature: GPT-5.6 Luna, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Luna, Flexible - FlexiblePlan: Business, Feature: GPT-5 Thinking Mini, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5 Thinking Mini, Flexible - Plan: Business, Feature: Legacy models, Yes - Plan: Enterprise, Feature: Legacy models, Yes - Fast Plan: Business, Feature: Response times, Fast - Fastest Plan: Enterprise, Feature: Response times, Fastest - 54K Plan: Business, Feature: GPT Instant total context window, 54K - 128K Plan: Enterprise, Feature: GPT Instant total context window, 128K - ~40 pages Plan: Business, Feature: GPT Instant input maximum***, ~40 pages - ~250 pages Plan: Enterprise, Feature: GPT Instant input maximum***, ~250 pages - 256K Plan: Business, Feature: GPT Reasoning total context window, 256K - 256K Plan: Enterprise, Feature: GPT Reasoning total context window, 256K - ~320 pages Plan: Business, Feature: GPT Reasoning input maximum***, ~320 pages - ~320 pages Plan: Enterprise, Feature: GPT Reasoning input maximum***, ~320 pages - Plan: Business, Feature: Regular quality & speed updates, Yes - Plan: Enterprise, Feature: Regular quality & speed updates, Yes - Desktop, web, and mobile Plan: Business, Feature: ChatGPT Work, Desktop, web, and mobile
-
GPQA Diamond (Epoch) scores changed
Tracking since Jun 25, 2026 · 31 snapshots on file
GPQA Diamond accuracy (0–1) updated.
Full history & current snapshot →View raw diff +1 −0
+ claude-sonnet-5-5_max 0.956 -
Tracking since Jun 25, 2026 · 280 snapshots on file
ChatGPT Business/Enterprise pricing consolidated on main page; API pricing section removed from this page
Removed API pricing section with model rates; page now focuses only on ChatGPT Business/Enterprise seat-based pricingFull history & current snapshot →View raw diff +199 −43
- Business Pricing | OpenAI - Price - Input:$10.00 / 1M tokens Cached input:$1.00 / 1M tokens Output:$50.00 / 1M tokens - Price - Input:$2.00 / 1M tokens Cached input:$0.10 / 1M tokens Output:$10.00 / 1M tokens - Price - Input:$0.10 / 1M tokens Cached input:$0.01 / 1M tokens Output:$0.50 / 1M tokens - Price - Image:$8.00 / 1M tokens for inputs$2.00 / 1M tokens for cached inputs$30.00 / 1M tokens for outputs Text:$5.00 / 1M tokens for inputs$1.25 / 1M tokens for cached inputs - Price - Image:$8.00 / 1M tokens for inputs$2.00 / 1M tokens for cached inputs$30.00 / 1M tokens for outputs Text:$5.00 / 1M tokens for inputs$1.25 / 1M tokens for cached inputs - Price - $0.05 per minute / $0.00083 per second - Price - Audio:$32.00 / 1M tokens for inputs$0.40 / 1M tokens for cached inputs$64.00 / 1M tokens for outputs Text:$4.00 / 1M tokens for inputs$0.40 / 1M tokens for cached inputs$24.00 / 1M tokens for outputs Image:$5.00 / 1M tokens for inputs$0.50 / 1M tokens for cached inputs - Price - Audio:$10.00 / 1M tokens for inputs$0.30 / 1M tokens for cached inputs$20.00 / 1M tokens for outputs Text:$0.60 / 1M tokens for inputs$0.06 / 1M tokens for cached inputs$2.40 / 1M tokens for outputs Image:$0.80 / 1M tokens for inputs$0.08 / 1M tokens for cached inputs - Price - $0.017 per minute / $0.00028 per second - Price - $0.0045 per minute / $0.00008 per second - Price - $0.034 per minute / $0.00057 per second - Price - $10.00 / 1k calls Search content tokens are free. - Price - Now:1 GB for $0.03 / 64GB for $1.92 per container Starting March 31, 2026:1 GB for $0.03 / 64GB for $1.92 per 20-minute session per container - Provides lower costs for requests in exchange for slower response times and occasional resource unavailability. Ideal for non-production or lower priority tasks. - Enterprise offerings - Contact our sales team to learn more about Data residency(opens in a new window), Scale Tier andReserved Capacity designed for cutting-edge customers running larger workloads. - We recommend experimenting with all of these models in the Playground(opens in a new window) to explore which models provide the best price performance trade-off for your usage. - Do you offer an enterprise package or SLAs? - We offer different tiers of access to our enterprise customers that include SLAs, lower latency, and more. Please contact our sales team to learn more. - Yes, we treat Playground usage the same as regular API usage. You will be billed at the per-token input and output prices mentioned above. - How will I know how many tokens I’ve used each month? - A token is a mathematical representation of natural language. Log in to your account to view your usage tracking dashboard(opens in a new window). This dashboard will show you how many tokens you’ve used during the current and past billing cycles. - You can set a monthly budget in your billing settings(opens in a new window), after which we’ll stop serving your requests. There may be a delay in enforcing the limit, and you are responsible for any overage incurred. You can also configure an email notification threshold to receive an email alert once you cross that threshold each month. We recommend checking your usage tracking dashboard(opens in a new window) regularly to monitor your spend. - Is access to the API included in ChatGPT Plus, Business, Enterprise or Edu? - No, OpenAI APIs are billed separately from ChatGPT Plus, Business, Enterprise and Edu. ChatGPT subscription pricing can be found at openai.com/chatgpt/pricing/. - Images are converted into tokens and charged per token. Text models price image tokens at standard text token rates, while GPT Image and gpt-realtime uses a separate image token rate. Models like gpt-4.1-mini, gpt-4.1-nano, and o4-mini convert images into tokens differently. Learn more in our docs(opens in a new window). - ChatGPT Business(opens in a new window) - ChatGPT Enterprise(opens in a new window) - ChatGPT for Education(opens in a new window) + ChatGPT Business(opens in a new window) + ChatGPT Enterprise(opens in a new window) + ChatGPT for Education(opens in a new window) + Image 1 + Business + A secure workspace with company context and flexible seat types for any budget. + Get started(opens in a new window) + Standard seat + $20/ month + $20/month if billed annually. $25/month if billed monthly. + Premium seat + $100/ month + 5x more usage than standard, with no 5-hour limit + $100/month if billed annually. $125/month if billed monthly. + No training on your business data by default + Mix and match seat types + For teams of 2–200 employees. Unlimited subject to abuse guardrails. Learn more(opens in a new window)
-
Tracking since Jun 25, 2026 · 13 snapshots on file
OpenAI restructured usage tiers from five (Tier 1-5) to three (Build, Launch, Grow) with adjusted credit thresholds and published per-model RPM/TPM limits.
now$500was$100/mo$50$1,000/mo$250$1,000Full history & current snapshot →View raw diff +21 −7
- You can view the rate and usage limits for your organization under the limits section of your account settings. As your spend on our API goes up, we automatically graduate you to the next usage tier. This usually results in an increase in rate limits across most models. - | Tier 1 | $5 paid | $100 / month | - | Tier 2 | $50 paid | $500 / month | - | Tier 3 | $100 paid | $1,000 / month | - | Tier 4 | $250 paid | $5,000 / month | - | Tier 5 | $1,000 paid | $200,000 / month | - If your use case does not require immediate responses, you can use the Batch API to more easily submit and execute large collections of requests without impacting your synchronous request rate limits. + The three paid usage tiers are Build, Launch, and Grow. Your organization’s usage tier upgrades automatically as its total credit purchases reach each threshold. Higher tiers generally provide higher rate limits across models. + | Build | $5 in total credit purchases | $500 / month | + | Launch | $100 in total credit purchases | $5,000 / month | + | Grow | $500 in total credit purchases | $200,000 / month | + Rate limits by usage tier + To view the limits for each model at your usage tier, go to Settings > Organization > Limits and review Rate limits. To upgrade your usage tier, select Upgrade tier in the Usage Tiers section. + | Tier | Model | RPM | TPM | + | --- | --- | --- | --- | + | Build | Astra, Sol, Terra | 5,000 | 1,000,000 | + | Build | Luna | 5,000 | 2,000,000 | + | Launch | Astra, Sol, Terra | 10,000 | 4,000,000 | + | Launch | Luna | 10,000 | 10,000,000 | + | Grow | Astra, Sol, Terra | 15,000 | 40,000,000 | + | Grow | Luna | 30,000 | 180,000,000 | + Spend limits + Consider setting spend limits for your organization or projects to control monthly API spend. These controls are separate from the monthly usage limits above. + | Control | What happens at the configured amount | Use it when you want to | + | Spend alert | Sends a notification; API traffic continues | Track spend without interrupting traffic | + | Hard spend limit | Affected API requests return a 429 error | Enforce a monthly organization or project cap | + The legacy Completions examples below use gpt-3.5-turbo-instruct, which has a scheduled shutdown date of September 28, 2026. After that date, retain the retry pattern but migrate the request to Responses or Chat Completions with gpt-5.6-terra; changing the model ID in a Completions request is not sufficient. + If your use case does not require immediate responses, you can use the Batch API to submit and execute large collections of requests without impacting your synchronous request rate limits.
-
Tracking since Jun 25, 2026 · 21 snapshots on file
Deprecation guidance for gpt-5.4-cyber changed from requiring migration to specific model gpt-5.6-cyber to recommending migration to the most capable cyber model available.
- gpt-5.4-cyber removal date remains October 1, 2026
- Migration target changed from gpt-5.6-cyber to the most capable cyber model available to you
Full history & current snapshot →View raw diff +2 −2
- The gpt-5.4-cyber model is deprecated and will be removed from the API on October 1, 2026. Migrate to gpt-5.6-cyber before the shutdown date. - | Oct 1, 2026 | gpt-5.4-cyber | gpt-5.6-cyber | + The gpt-5.4-cyber model is deprecated and will be removed from the API on October 1, 2026. Migrate to the most capable cyber model available to you before the shutdown date. + | Oct 1, 2026 | gpt-5.4-cyber | The most capable cyber model available to you. |
-
Tracking since Jun 25, 2026 · 280 snapshots on file
GPT-4 Turbo cached input price reduced by 50% (from $0.20 to $0.10 / 1M tokens)
rate_limitCached input price for GPT-4 Turbo (and unnamed model) reduced from $0.20 / 1M tokens to $0.10 / 1M tokensFull history & current snapshot →View raw diff +4 −4
- Input:$2.00 / 1M tokens Cached input:$0.20 / 1M tokens Output:$10.00 / 1M tokens - ChatGPT Business(opens in a new window) - ChatGPT Enterprise(opens in a new window) - ChatGPT for Education(opens in a new window) + Input:$2.00 / 1M tokens Cached input:$0.10 / 1M tokens Output:$10.00 / 1M tokens + ChatGPT Business(opens in a new window) + ChatGPT Enterprise(opens in a new window) + ChatGPT for Education(opens in a new window)
-
Windsurf (now Devin) pricing changed
Tracking since Jun 25, 2026 · 6 snapshots on file
SWE-2 Free access extended by 6 days to October 16, 2026 for Pro and Max plans
Pro · benefit_expiryOctober 10, 2026→October 16, 2026Full history & current snapshot →View raw diff +2 −2
- Free use of SWE-2 Free in Devin Desktop and CLI through October 10, 2026 - Free SWE-2Free in Devin Desktop and CLI through October 10, 2026 + Free use of SWE-2 Free in Devin Desktop and CLI through October 16, 2026 + Free SWE-2Free in Devin Desktop and CLI through October 16, 2026
-
GPQA Diamond (Epoch) scores changed
Tracking since Jun 25, 2026 · 31 snapshots on file
GPQA Diamond accuracy (0–1) updated.
Full history & current snapshot →View raw diff +1 −0
+ claude-opus-5-5_max 0.906 -
Tracking since Jun 25, 2026 · 57 snapshots on file
Claude Sonnet 5.5 added as separate model tier with 2M ITPM; Sonnet 4.x footnote updated to clarify 5.5 and 5 have independent limits
Full history & current snapshot →View raw diff +2 −1
- _3 Sonnet 4.x rate limit is a total limit that applies to combined traffic across Sonnet 4.6 and Sonnet 4.5. Claude Sonnet 5 has a separate rate limit and is not part of this combined bucket._ + | Claude Sonnet 5.5 | 1,000 | 2,000,000 | 400,000 | + _3 Sonnet 4.x rate limit is a total limit that applies to combined traffic across Sonnet 4.6 and Sonnet 4.5. Claude Sonnet 5.5 and Claude Sonnet 5 each have a separate rate limit and are not part of this combined bucket._
-
Tracking since Jun 25, 2026 · 280 snapshots on file
OpenAI pricing page now displays ChatGPT Business and Enterprise subscription plans instead of API token rates
Pricing page shifted focus from API token rates to ChatGPT Business/Enterprise seat-based subscription plansFull history & current snapshot →View raw diff +199 −42
- Price - Input:$10.00 / 1M tokens Cached input:$1.00 / 1M tokens Output:$50.00 / 1M tokens - Price - Input:$2.00 / 1M tokens Cached input:$0.20 / 1M tokens Output:$10.00 / 1M tokens - Price - Input:$0.10 / 1M tokens Cached input:$0.01 / 1M tokens Output:$0.50 / 1M tokens - Price - Image:$8.00 / 1M tokens for inputs$2.00 / 1M tokens for cached inputs$30.00 / 1M tokens for outputs Text:$5.00 / 1M tokens for inputs$1.25 / 1M tokens for cached inputs - Price - Image:$8.00 / 1M tokens for inputs$2.00 / 1M tokens for cached inputs$30.00 / 1M tokens for outputs Text:$5.00 / 1M tokens for inputs$1.25 / 1M tokens for cached inputs - Price - $0.05 per minute / $0.00083 per second - Price - Audio:$32.00 / 1M tokens for inputs$0.40 / 1M tokens for cached inputs$64.00 / 1M tokens for outputs Text:$4.00 / 1M tokens for inputs$0.40 / 1M tokens for cached inputs$24.00 / 1M tokens for outputs Image:$5.00 / 1M tokens for inputs$0.50 / 1M tokens for cached inputs - Price - Audio:$10.00 / 1M tokens for inputs$0.30 / 1M tokens for cached inputs$20.00 / 1M tokens for outputs Text:$0.60 / 1M tokens for inputs$0.06 / 1M tokens for cached inputs$2.40 / 1M tokens for outputs Image:$0.80 / 1M tokens for inputs$0.08 / 1M tokens for cached inputs - Price - $0.017 per minute / $0.00028 per second - Price - $0.0045 per minute / $0.00008 per second - Price - $0.034 per minute / $0.00057 per second - Price - $10.00 / 1k calls Search content tokens are free. - Price - Now:1 GB for $0.03 / 64GB for $1.92 per container Starting March 31, 2026:1 GB for $0.03 / 64GB for $1.92 per 20-minute session per container - Provides lower costs for requests in exchange for slower response times and occasional resource unavailability. Ideal for non-production or lower priority tasks. - Enterprise offerings - Contact our sales team to learn more about Data residency(opens in a new window), Scale Tier andReserved Capacity designed for cutting-edge customers running larger workloads. - We recommend experimenting with all of these models in the Playground(opens in a new window) to explore which models provide the best price performance trade-off for your usage. - Do you offer an enterprise package or SLAs? - We offer different tiers of access to our enterprise customers that include SLAs, lower latency, and more. Please contact our sales team to learn more. - Yes, we treat Playground usage the same as regular API usage. You will be billed at the per-token input and output prices mentioned above. - How will I know how many tokens I’ve used each month? - A token is a mathematical representation of natural language. Log in to your account to view your usage tracking dashboard(opens in a new window). This dashboard will show you how many tokens you’ve used during the current and past billing cycles. - You can set a monthly budget in your billing settings(opens in a new window), after which we’ll stop serving your requests. There may be a delay in enforcing the limit, and you are responsible for any overage incurred. You can also configure an email notification threshold to receive an email alert once you cross that threshold each month. We recommend checking your usage tracking dashboard(opens in a new window) regularly to monitor your spend. - Is access to the API included in ChatGPT Plus, Business, Enterprise or Edu? - No, OpenAI APIs are billed separately from ChatGPT Plus, Business, Enterprise and Edu. ChatGPT subscription pricing can be found at openai.com/chatgpt/pricing/. - Images are converted into tokens and charged per token. Text models price image tokens at standard text token rates, while GPT Image and gpt-realtime uses a separate image token rate. Models like gpt-4.1-mini, gpt-4.1-nano, and o4-mini convert images into tokens differently. Learn more in our docs(opens in a new window). - ChatGPT Business(opens in a new window) - ChatGPT Enterprise(opens in a new window) - ChatGPT for Education(opens in a new window) + Image 1 + Business + A secure workspace with company context and flexible seat types for any budget. + Get started(opens in a new window) + Standard seat + $20/ month + $20/month if billed annually. $25/month if billed monthly. + Premium seat + $100/ month + 5x more usage than standard, with no 5-hour limit + $100/month if billed annually. $125/month if billed monthly. + No training on your business data by default + Mix and match seat types + For teams of 2–200 employees. Unlimited subject to abuse guardrails. Learn more(opens in a new window) + Image 2 + Enterprise + Enterprise-grade AI, security, and support for businesses operating at scale + Contact our sales team to discuss enterprise pricing.*
-
Tracking since Jun 25, 2026 · 57 snapshots on file
GitHub API: 1 additive change
Full history & current snapshot →View raw diff +1 −0
+ New endpoint: GET /repos/{owner}/{repo}/actions/jobs/{job_id}/steps/{step_number}/logs -
Anthropic models model list changed
Tracking since Jun 25, 2026 · 8 snapshots on file
Claude Sonnet 5.5 model is now available.
- claude-sonnet-5-5
Full history & current snapshot →View raw diff +1 −0
+ claude-sonnet-5-5 -
Tracking since Jun 25, 2026 · 18 snapshots on file
DigitalOcean API: 1 additive change
Full history & current snapshot →View raw diff +1 −0
+ New endpoint: POST /v1/systemone -
Tracking since Jun 25, 2026 · 57 snapshots on file
GitHub API: 6 additive changes
Full history & current snapshot →View raw diff +6 −0
+ New endpoint: GET /orgs/{org}/properties/installations + New endpoint: POST /orgs/{org}/properties/installations + New endpoint: GET /orgs/{org}/properties/installations/schema + New endpoint: PATCH /orgs/{org}/properties/installations/values + New endpoint: DELETE /orgs/{org}/properties/installations/values/{property_name} + New endpoint: PATCH /orgs/{org}/properties/installations/values/{property_name}
-
Anthropic model deprecations changed
Tracking since Jun 25, 2026 · 45 snapshots on file
Added Active status with availability date of September 28, 2027.
- Active status
- September 28, 2027
Full history & current snapshot →View raw diff +1 −0
+ | | Active | N/A | Not sooner than September 28, 2027 | -
Tracking since Jun 25, 2026 · 21 snapshots on file
Image model deprecation migration paths updated to point to gpt-image-2.5-sunburst or gpt-image-2.5-flare instead of gpt-image-2.
- gpt-image-1-mini, gpt-image-1.5, and chatgpt-image-latest now migrate to gpt-image-2.5-sunburst or gpt-image-2.5-flare on Dec 1, 2026
- gpt-image-1 now migrates to gpt-image-2.5-sunburst or gpt-image-2.5-flare on October 23, 2026
- Agent Builder migration guidance changed from evaluate Agents API option to continue with Agents SDK or ChatGPT Workspace Agents
Full history & current snapshot →View raw diff +5 −5
- See Migrate from Agent Builder to evaluate the Agents API, ChatGPT Workspace Agents, or an existing Agents SDK integration. - | Dec 1, 2026 | gpt-image-1-mini | gpt-image-2 | - | Dec 1, 2026 | gpt-image-1.5 | gpt-image-2 | - | Dec 1, 2026 | chatgpt-image-latest | gpt-image-2 | - | October 23, 2026 | gpt-image-1 | gpt-image-2 | + See Migrate from Agent Builder to continue with the Agents SDK or ChatGPT Workspace Agents. + | Dec 1, 2026 | gpt-image-1-mini | gpt-image-2.5-sunburst or gpt-image-2.5-flare | + | Dec 1, 2026 | gpt-image-1.5 | gpt-image-2.5-sunburst or gpt-image-2.5-flare | + | Dec 1, 2026 | chatgpt-image-latest | gpt-image-2.5-sunburst or gpt-image-2.5-flare | + | October 23, 2026 | gpt-image-1 | gpt-image-2.5-sunburst or gpt-image-2.5-flare |
-
Meta reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Llama 4 Maverick: 10 benchmark claims (via web search)
Full history & current snapshot → -
Moonshot AI reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Kimi K3: 12 benchmark claims (via web search)
Full history & current snapshot → -
Zhipu AI (Z.ai) reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
GLM-5.2: 9 benchmark claims (via web search)
Full history & current snapshot → -
Alibaba reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Qwen3.8-Max: 10 benchmark claims (via web search)
Full history & current snapshot → -
DeepSeek reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
DeepSeek-V4-Pro-0813: 12 benchmark claims (via web search)
Full history & current snapshot → -
xAI reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Grok 4.6: 10 benchmark claims (via web search)
Full history & current snapshot → -
Google reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Gemini 3.1 Pro: 10 benchmark claims (via web search)
Full history & current snapshot → -
OpenAI reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
GPT-6 Astra: 10 benchmark claims (via web search)
Full history & current snapshot → -
Anthropic reported benchmarks updated
Tracking since Jun 25, 2026 · 15 snapshots on file
Claude Opus 5.5: 4 benchmark claims (via web search)
Full history & current snapshot → -
Tracking since Jun 25, 2026 · 280 snapshots on file
Batch API file storage billing changes from per-container to per-session billing effective March 31, 2026
Batch API file storage pricing changes from per-container to per-session model effective March 31, 2026Full history & current snapshot →View raw diff +43 −199
- Image 1 - Business - A secure workspace with company context and flexible seat types for any budget. - Get started(opens in a new window) - Standard seat - $20/ month - $20/month if billed annually. $25/month if billed monthly. - Premium seat - $100/ month - 5x more usage than standard, with no 5-hour limit - $100/month if billed annually. $125/month if billed monthly. - No training on your business data by default - Mix and match seat types - For teams of 2–200 employees. Unlimited subject to abuse guardrails. Learn more(opens in a new window) - Image 2 - Enterprise - Enterprise-grade AI, security, and support for businesses operating at scale - Contact our sales team to discuss enterprise pricing.* - Enterprise-level security and controls, including SCIM, EKM, user analytics, domain verification, and role-based access controls - Advanced data privacy with custom data retention policies, encryption at rest and in transit, and no training on your business data by default. Learn more - *Credit-based(opens in a new window) pricing and token-based(opens in a new window) pricing are available for Enterprise plans. - Looking for personal plans? - Compare features across plans - Try ChatGPT Business(opens in a new window) - Enterprise - Unlimited*Plan: Business, Feature: Everyday text chats, Unlimited* - Unlimited*Plan: Enterprise, Feature: Everyday text chats, Unlimited* - Unlimited*Plan: Business, Feature: Chat history, Unlimited* - Unlimited*Plan: Enterprise, Feature: Chat history, Unlimited* - Plan: Business, Feature: Access on web, iOS, Android, Yes - Plan: Enterprise, Feature: Access on web, iOS, Android, Yes - FlexiblePlan: Business, Feature: GPT-5.6 Sol, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Sol, Flexible - GPT-5.6 Sol Pro - FlexiblePlan: Business, Feature: GPT-5.6 Sol Pro, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Sol Pro, Flexible - FlexiblePlan: Business, Feature: GPT-5.6 Terra, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Terra, Flexible - FlexiblePlan: Business, Feature: GPT-5.6 Luna, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5.6 Luna, Flexible - FlexiblePlan: Business, Feature: GPT-5 Thinking Mini, Flexible - FlexiblePlan: Enterprise, Feature: GPT-5 Thinking Mini, Flexible - Plan: Business, Feature: Legacy models, Yes - Plan: Enterprise, Feature: Legacy models, Yes - Fast Plan: Business, Feature: Response times, Fast - Fastest Plan: Enterprise, Feature: Response times, Fastest - 54K Plan: Business, Feature: GPT Instant total context window, 54K - 128K Plan: Enterprise, Feature: GPT Instant total context window, 128K - ~40 pages Plan: Business, Feature: GPT Instant input maximum***, ~40 pages - ~250 pages Plan: Enterprise, Feature: GPT Instant input maximum***, ~250 pages - 256K Plan: Business, Feature: GPT Reasoning total context window, 256K - 256K Plan: Enterprise, Feature: GPT Reasoning total context window, 256K - ~320 pages Plan: Business, Feature: GPT Reasoning input maximum***, ~320 pages - ~320 pages Plan: Enterprise, Feature: GPT Reasoning input maximum***, ~320 pages - Plan: Business, Feature: Regular quality & speed updates, Yes - Plan: Enterprise, Feature: Regular quality & speed updates, Yes - Desktop, web, and mobile Plan: Business, Feature: ChatGPT Work, Desktop, web, and mobile - Desktop, web, and mobile Plan: Enterprise, Feature: ChatGPT Work, Desktop, web, and mobile - Plan: Business, Feature: Codex, Yes - Plan: Enterprise, Feature: Codex, Yes
-
Tracking since Jun 25, 2026 · 26 snapshots on file
Arena Elo (text, overall) updated.
Full history & current snapshot →View raw diff +58 −58
- amazon-nova-experimental-chat-26-02-10 1448 - claude-fable-5 1493 - claude-fable-5.1-max 1508 - claude-opus-4-5-20251101 1450 - claude-opus-4-5-20251101-high-32k 1447 - claude-opus-4-6-high 1503 - claude-opus-4-7 1483 - claude-opus-4-8-high 1461 - claude-opus-5-high 1505 - claude-opus-5-max 1505 - claude-sonnet-4-5-20250929 1438 - claude-sonnet-4-5-20250929-high-32k 1434 - claude-sonnet-5-high 1442 - deepseek-v3.2 1425 - deepseek-v3.2-exp 1423 - deepseek-v4-flash-high-preview 1424 - deepseek-v4-pro-high-20260813 1444 - ernie-5.0-0110 1444 - ernie-5.0-preview-1203 1443 - gemini-3-flash 1467 - gemini-3.5-flash-high 1482 - gemini-3.6-flash-high 1476 - gemini-3.7-flash-high 1490 - gemini-3.8-flash-high 1495 - gemma-4-31b 1442 - glm-4.7 1436 - glm-5.2-max 1467 - glm-5.3-flash 1472 - glm-5.3-max 1475 - glm-5v-turbo 1436 - gpt-5.1 1423 - gpt-5.1-high 1442 - gpt-5.4 1453 - gpt-5.4-high 1470 - gpt-5.5 1466 - gpt-5.6-luna-xhigh 1430 - gpt-5.6-sol-xhigh 1455 - gpt-6-astra-max 1444 - grok-3-preview-02-24 1425 - grok-4.1-thinking 1437 - grok-4.20-beta1 1444 - grok-4.5 1450 - grok-4.6-high 1430 - hy3 1441 - inkling 1440 - kimi-k2.5-thinking 1446 - kimi-k3-max 1472 - longcat-flash-chat 1423 - longcat-flash-chat-2602-exp 1426 - mistral-medium-2508 1425 - muse-spark 1473 - muse-spark-1.1 1480 - muse-spark-1.2 (xHigh) 1489 - nvidia-nemotron-3-ultra-550b-a55b-nvfp4 1445 - qwen3.6-max-preview 1446 - qwen3.7-max-preview 1474 - qwen3.8-27b 1439 - qwen3.8-max 1481 + amazon-nova-experimental-chat-26-02-10 1449 + claude-fable-5-high 1491
-
Tracking since Jun 25, 2026 · 280 snapshots on file
All API pricing information was removed from the page and replaced with ChatGPT Business and Enterprise plan marketing content.
now$20/mo$25/mo$100/mo$125/mo2userswas$10.00$1.00$50.00$2.00$0.20$0.10$0.01$0.50Full history & current snapshot →View raw diff +199 −43
- Business Pricing | OpenAI - Price - Input:$10.00 / 1M tokens Cached input:$1.00 / 1M tokens Output:$50.00 / 1M tokens - Price - Input:$2.00 / 1M tokens Cached input:$0.20 / 1M tokens Output:$10.00 / 1M tokens - Price - Input:$0.10 / 1M tokens Cached input:$0.01 / 1M tokens Output:$0.50 / 1M tokens - Price - Image:$8.00 / 1M tokens for inputs$2.00 / 1M tokens for cached inputs$30.00 / 1M tokens for outputs Text:$5.00 / 1M tokens for inputs$1.25 / 1M tokens for cached inputs - Price - Image:$8.00 / 1M tokens for inputs$2.00 / 1M tokens for cached inputs$30.00 / 1M tokens for outputs Text:$5.00 / 1M tokens for inputs$1.25 / 1M tokens for cached inputs - Price - $0.05 per minute / $0.00083 per second - Price - Audio:$32.00 / 1M tokens for inputs$0.40 / 1M tokens for cached inputs$64.00 / 1M tokens for outputs Text:$4.00 / 1M tokens for inputs$0.40 / 1M tokens for cached inputs$24.00 / 1M tokens for outputs Image:$5.00 / 1M tokens for inputs$0.50 / 1M tokens for cached inputs - Price - Audio:$10.00 / 1M tokens for inputs$0.30 / 1M tokens for cached inputs$20.00 / 1M tokens for outputs Text:$0.60 / 1M tokens for inputs$0.06 / 1M tokens for cached inputs$2.40 / 1M tokens for outputs Image:$0.80 / 1M tokens for inputs$0.08 / 1M tokens for cached inputs - Price - $0.017 per minute / $0.00028 per second - Price - $0.0045 per minute / $0.00008 per second - Price - $0.034 per minute / $0.00057 per second - Price - $10.00 / 1k calls Search content tokens are free. - Price - Now:1 GB for $0.03 / 64GB for $1.92 per container Starting March 31, 2026:1 GB for $0.03 / 64GB for $1.92 per 20-minute session per container - Provides lower costs for requests in exchange for slower response times and occasional resource unavailability. Ideal for non-production or lower priority tasks. - Enterprise offerings - Contact our sales team to learn more about Data residency(opens in a new window), Scale Tier andReserved Capacity designed for cutting-edge customers running larger workloads. - We recommend experimenting with all of these models in the Playground(opens in a new window) to explore which models provide the best price performance trade-off for your usage. - Do you offer an enterprise package or SLAs? - We offer different tiers of access to our enterprise customers that include SLAs, lower latency, and more. Please contact our sales team to learn more. - Yes, we treat Playground usage the same as regular API usage. You will be billed at the per-token input and output prices mentioned above. - How will I know how many tokens I’ve used each month? - A token is a mathematical representation of natural language. Log in to your account to view your usage tracking dashboard(opens in a new window). This dashboard will show you how many tokens you’ve used during the current and past billing cycles. - You can set a monthly budget in your billing settings(opens in a new window), after which we’ll stop serving your requests. There may be a delay in enforcing the limit, and you are responsible for any overage incurred. You can also configure an email notification threshold to receive an email alert once you cross that threshold each month. We recommend checking your usage tracking dashboard(opens in a new window) regularly to monitor your spend. - Is access to the API included in ChatGPT Plus, Business, Enterprise or Edu? - No, OpenAI APIs are billed separately from ChatGPT Plus, Business, Enterprise and Edu. ChatGPT subscription pricing can be found at openai.com/chatgpt/pricing/. - Images are converted into tokens and charged per token. Text models price image tokens at standard text token rates, while GPT Image and gpt-realtime uses a separate image token rate. Models like gpt-4.1-mini, gpt-4.1-nano, and o4-mini convert images into tokens differently. Learn more in our docs(opens in a new window). - ChatGPT Business(opens in a new window) - ChatGPT Enterprise(opens in a new window) - ChatGPT for Education(opens in a new window) + ChatGPT Business(opens in a new window) + ChatGPT Enterprise(opens in a new window) + ChatGPT for Education(opens in a new window) + Image 1 + Business + A secure workspace with company context and flexible seat types for any budget. + Get started(opens in a new window) + Standard seat + $20/ month + $20/month if billed annually. $25/month if billed monthly. + Premium seat + $100/ month + 5x more usage than standard, with no 5-hour limit + $100/month if billed annually. $125/month if billed monthly. + No training on your business data by default + Mix and match seat types + For teams of 2–200 employees. Unlimited subject to abuse guardrails. Learn more(opens in a new window)
-
NVIDIA Nemotron models changed
Tracking since Jun 25, 2026 · 27 snapshots on file
Added new model NV-Reason-CT with openmdw-1.1 license expiring 2026-09-08.
- NV-Reason-CT
- openmdw-1.1 license
- expiry: 2026-09-08
Full history & current snapshot →View raw diff +1 −0
+ nvidia/NV-Reason-CT license:openmdw-1.1 2026-09-08 -
Google Gemini deprecations changed
Tracking since Jun 25, 2026 · 22 snapshots on file
Deprecated gemini-3.1-flash-tts-preview model removed from table; migration path updated to gemini-3.8-flash-tts or gemini-3.8-flash-lite-tts for preview TTS models.
- gemini-3.1-flash-tts-preview row deleted
- gemini-2.5-flash-preview-tts migration target changed from gemini-3.1-flash-tts-preview to gemini-3.8-flash-tts or gemini-3.8-flash-lite-tts
- gemini-2.5-pro-preview-tts migration target changed from gemini-3.1-flash-tts-preview to gemini-3.8-flash-tts or gemini-3.8-flash-lite-tts
Full history & current snapshot →View raw diff +2 −3
- | gemini-3.1-flash-tts-preview | April 13, 2026 | No shutdown date announced | | - | gemini-2.5-flash-preview-tts | May 20, 2025 | No shutdown date announced | gemini-3.1-flash-tts-preview | - | gemini-2.5-pro-preview-tts | May 20, 2025 | No shutdown date announced | gemini-3.1-flash-tts-preview | + | gemini-2.5-flash-preview-tts | May 20, 2025 | No shutdown date announced | gemini-3.8-flash-tts or gemini-3.8-flash-lite-tts | + | gemini-2.5-pro-preview-tts | May 20, 2025 | No shutdown date announced | gemini-3.8-flash-tts or gemini-3.8-flash-lite-tts |
-
Tracking since Jun 25, 2026 · 57 snapshots on file
GitHub API: 3 additive changes
Full history & current snapshot →View raw diff +3 −0
+ New endpoint: GET /repos/{owner}/{repo}/issues/{issue_number}/relates_to + New endpoint: POST /repos/{owner}/{repo}/issues/{issue_number}/relates_to + New endpoint: DELETE /repos/{owner}/{repo}/issues/{issue_number}/relates_to/{issue_id}
-
Tracking since Jun 25, 2026 · 21 snapshots on file
Migration guidance text expanded to include evaluating Agents API as an option alongside existing alternatives.
- Changed from 'continue with' to 'evaluate the Agents API, ChatGPT Workspace Agents, or an existing Agents SDK integration'
Full history & current snapshot →View raw diff +1 −1
- See Migrate from Agent Builder to continue with the Agents SDK or ChatGPT Workspace Agents. + See Migrate from Agent Builder to evaluate the Agents API, ChatGPT Workspace Agents, or an existing Agents SDK integration.
-
Google Gemini API pricing changed
Tracking since Jun 25, 2026 · 22 snapshots on file
Audio output pricing specified with per-10-second equivalents and January 2027 rate increases across Gemini 2.0 audio models
Audio pricing tiers added with per-10-second equivalents for multiple Gemini 2.0 audio models, with increases scheduled for January 1, 2027Full history & current snapshot →View raw diff +16 −0
+ Equivalent to $0.00225 per 10s audio* through December 31, 2026. + Equivalent to $0.0045 per 10s audio* starting January 1, 2027. + Equivalent to $0.001125 per 10s audio* through December 31, 2026. + Equivalent to $0.00225 per 10s audio* starting January 1, 2027. + Equivalent to $0.001125 per 10s audio* through December 31, 2026. + Equivalent to $0.00225 per 10s audio* starting January 1, 2027. + Equivalent to $0.00405 per 10s audio* through December 31, 2026. + Equivalent to $0.0081 per 10s audio* starting January 1, 2027. + Equivalent to $0.0015 per 10s audio* through December 31, 2026. + Equivalent to $0.003 per 10s audio* starting January 1, 2027. + Equivalent to $0.00075 per 10s audio* through December 31, 2026. + Equivalent to $0.0015 per 10s audio* starting January 1, 2027. + Equivalent to $0.00075 per 10s audio* through December 31, 2026. + Equivalent to $0.0015 per 10s audio* starting January 1, 2027. + Equivalent to $0.0027 per 10s audio* through December 31, 2026. + Equivalent to $0.0054 per 10s audio* starting January 1, 2027.
-
GitHub Copilot pricing changed
Tracking since Jun 25, 2026 · 8 snapshots on file
GitHub Copilot Enterprise plan description updated to remove custom model fine-tuning claim
Enterprise plan feature clarification: removed reference to 'fine-tuned custom, private models for inline suggestions', now only mentions general customization featuresFull history & current snapshot →View raw diff +2 −2
- Organizations can choose between GitHub Copilot Business and GitHub Copilot Enterprise. GitHub Copilot Business primarily features GitHub Copilot in the coding environment - that is the IDE, CLI and GitHub Mobile. GitHub Copilot Enterprise includes everything in GitHub Copilot Business. It also adds an additional layer of customization for organizations and integrates into GitHub.com as a chat interface to allow developers to converse with GitHub Copilot throughout the platform. GitHub Copilot Enterprise can index an organization’s codebase for a deeper understanding of the customer’s knowledge for more tailored suggestions and will offer customers access to fine-tuned custom, private models for inline suggestions. - Organizations can choose between GitHub Copilot Business and GitHub Copilot Enterprise. GitHub Copilot Business primarily features GitHub Copilot in the coding environment - that is the IDE, CLI and GitHub Mobile. GitHub Copilot Enterprise includes everything in GitHub Copilot Business. It also adds an additional layer of customization for organizations and integrates into GitHub.com as a chat interface to allow developers to converse with GitHub Copilot throughout the platform. GitHub Copilot Enterprise can index an organization’s codebase for a deeper understanding of the customer’s knowledge for more tailored suggestions and will offer customers access to fine-tuned custom, private models for inline suggestions. + Organizations can choose between GitHub Copilot Business and GitHub Copilot Enterprise. GitHub Copilot Business primarily features GitHub Copilot in the coding environment - that is the IDE, CLI and GitHub Mobile. GitHub Copilot Enterprise includes everything in GitHub Copilot Business. It also adds an additional layer of customization for organizations and integrates into GitHub.com as a chat interface to allow developers to converse with GitHub Copilot throughout the platform. GitHub Copilot Enterprise can index an organization's codebase for a deeper understanding of the customer's knowledge for more tailored suggestions. For more details, refer to the documentation. + Organizations can choose between GitHub Copilot Business and GitHub Copilot Enterprise. GitHub Copilot Business primarily features GitHub Copilot in the coding environment - that is the IDE, CLI and GitHub Mobile. GitHub Copilot Enterprise includes everything in GitHub Copilot Business. It also adds an additional layer of customization for organizations and integrates into GitHub.com as a chat interface to allow developers to converse with GitHub Copilot throughout the platform. GitHub Copilot Enterprise can index an organization's codebase for a deeper understanding of the customer's knowledge for more tailored suggestions. For more details, refer to the documentation.
-
Tracking since Jun 25, 2026 · 4 snapshots on file
Removed Quick Start section; concurrency limits remain 2500 (deepseek-flash) and 500 (deepseek-v4-pro) per account with optional user_id isolation.
Full history & current snapshot →View raw diff +0 −1
- Quick Start -
Windsurf (now Devin) pricing changed
Tracking since Jun 25, 2026 · 6 snapshots on file
Team plan user limit changed from unlimited to 200 users maximum
Team · user_limitUnlimited team members / Unlimited flex seats→Up to 200 usersFull history & current snapshot →View raw diff +1 −2
- Unlimited team members - Unlimited flex seats+$40/month per full userEach full user includes their own generous quota and access to Devin Desktop + Up to 200 users$40/month per full userEach full user includes their own generous quota and access to Devin Desktop
-
Tracking since Jun 25, 2026 · 6 snapshots on file
Stripe API: 18 additive changes
Full history & current snapshot →View raw diff +18 −0
+ New endpoint: GET /v1/apps/installs + New endpoint: POST /v1/apps/installs + New endpoint: GET /v1/apps/installs/{id} + New endpoint: POST /v1/apps/installs/{id} + New endpoint: POST /v1/apps/installs/{id}/uninstall + New endpoint: GET /v1/product_catalog/trial_offers + New endpoint: POST /v1/product_catalog/trial_offers + New endpoint: GET /v1/product_catalog/trial_offers/{id} + New endpoint: POST /v1/product_catalog/trial_offers/{id} + New endpoint: POST /v1/subscriptions/{subscription}/pause + New endpoint: GET /v1/tax/locations + New endpoint: POST /v1/tax/locations + New endpoint: GET /v1/tax/locations/{location} + New endpoint: GET /v1/three_d_secure/authentications + New endpoint: POST /v1/three_d_secure/authentications + New endpoint: GET /v1/three_d_secure/authentications/{authentication} + New endpoint: POST /v1/three_d_secure/authentications/{authentication}/cancel + New endpoint: POST /v1/three_d_secure/authentications/{authentication}/submit
-
Tracking since Jun 25, 2026 · 25 snapshots on file
Cursor launched Rollouts and Security Review bots available today on Teams and Enterprise plans for monitoring deployments and detecting exploitable bugs in pull requests.
- Rollouts monitors pull request deployments per environment reporting verified healthy, regression detected, or inconclusive status
- Security Review posts one review comment reporting exploitable bugs on every pull request
- Both features available today on Teams and Enterprise plans
Full history & current snapshot →View raw diff +52 −48
- Aug 17, 2026 · Changelog - Origin Code Hosting - Cursor can now host your code. - Origin begins rolling out today in early beta on all paid plans. We're starting with the essentials, designed for agent scale: repos, pull requests, code browsing, and GitHub sync. Agent-native features ship soon. - #Origin Repos - The new Codebase tab is home for Origin repos. - Click +New to create a new repo and name it. Once you do, a page shows you how to install the CLI, with commands for how to clone a repo or push a local project. Push, and your code is hosted on Origin. - For Origin-hosted repos, Origin is the source of truth. Pushes land on Origin, and GitHub is not in the path. - Name your codebase when you create your first repo. That name becomes part of every repo's URL: cursor.com/codebase/acme-corp. - #Bring your GitHub repos - Your GitHub repos can sit alongside the ones Cursor hosts. Connect GitHub to Cursor, pick your org, and you'll see the repos you can sync. Select one and Cursor pulls it in. You choose what gets synced and can disconnect a repo at any time. Anyone with read or write access to a synced repo can view it in Cursor too. - Synced repos update in real time. Browse, search, and pull from the copy in Origin. For synced repos, GitHub stays the source of truth: pushes keep going to GitHub, and Origin mirrors the result. Icons next to each repo name tell you which ones Cursor hosts and which came from GitHub. - #Pull requests - Every repo has pull requests. Open one to see the timeline, commits, checks, and files changed. Review the diff, leave comments, and merge. - Pull requests on synced repos sync both ways: comment in Cursor and it posts to GitHub, react or reply on GitHub and it shows up in Cursor within seconds. Got a review assigned to you on GitHub? Review and merge it from Cursor. - #Agents in every repo - Your code, PRs, and agents are now in the same place. Ask Cursor questions about code you're browsing. It can answer, make changes, update PRs, or push a branch. - #App extensions for Cursor repos - We're building an app ecosystem so your whole stack works seamlessly with Origin. Integrations with Vercel, Depot, and Buildkite are already available, with more coming soon. - Connect Vercel from a repo's Apps tab and every PR gets a preview deployment where you can test and make comments. Merge, and it ships to production. For CI, connect Depot or Buildkite. Both run your existing GitHub Actions workflows and Buildkite also runs its native pipelines. - #Settings - Every repo has settings. Check sync status for GitHub repos, manage who has access, and see which apps are connected. - Origin is rolling out in early beta to all paid plan users starting today, except enterprise orgs whose admins opt out. Name your codebase and create your first repo. - Learn more in our docs or get started today. - Aug 17, 2026 · Changelog - Origin Code Hosting - Cursor can now host your code. - Origin begins rolling out today in early beta on all paid plans. We're starting with the essentials, designed for agent scale: repos, pull requests, code browsing, and GitHub sync. Agent-native features ship soon. - #Origin Repos - The new Codebase tab is home for Origin repos. - Click +New to create a new repo and name it. Once you do, a page shows you how to install the CLI, with commands for how to clone a repo or push a local project. Push, and your code is hosted on Origin. - For Origin-hosted repos, Origin is the source of truth. Pushes land on Origin, and GitHub is not in the path. - Name your codebase when you create your first repo. That name becomes part of every repo's URL: cursor.com/codebase/acme-corp. - #Bring your GitHub repos - Your GitHub repos can sit alongside the ones Cursor hosts. Connect GitHub to Cursor, pick your org, and you'll see the repos you can sync. Select one and Cursor pulls it in. You choose what gets synced and can disconnect a repo at any time. Anyone with read or write access to a synced repo can view it in Cursor too. - Synced repos update in real time. Browse, search, and pull from the copy in Origin. For synced repos, GitHub stays the source of truth: pushes keep going to GitHub, and Origin mirrors the result. Icons next to each repo name tell you which ones Cursor hosts and which came from GitHub. - #Pull requests - Every repo has pull requests. Open one to see the timeline, commits, checks, and files changed. Review the diff, leave comments, and merge. - Pull requests on synced repos sync both ways: comment in Cursor and it posts to GitHub, react or reply on GitHub and it shows up in Cursor within seconds. Got a review assigned to you on GitHub? Review and merge it from Cursor. - #Agents in every repo - Your code, PRs, and agents are now in the same place. Ask Cursor questions about code you're browsing. It can answer, make changes, update PRs, or push a branch. - #App extensions for Cursor repos - We're building an app ecosystem so your whole stack works seamlessly with Origin. Integrations with Vercel, Depot, and Buildkite are already available, with more coming soon. - Connect Vercel from a repo's Apps tab and every PR gets a preview deployment where you can test and make comments. Merge, and it ships to production. For CI, connect Depot or Buildkite. Both run your existing GitHub Actions workflows and Buildkite also runs its native pipelines. - #Settings - Every repo has settings. Check sync status for GitHub repos, manage who has access, and see which apps are connected. - Origin is rolling out in early beta to all paid plan users starting today, except enterprise orgs whose admins opt out. Name your codebase and create your first repo. - Learn more in our docs or get started today. + Sep 23, 2026 · Changelog + Rollouts and Security Review + Today we're launching two Cursor bots for the last mile of shipping code. Rollouts watches every change as it deploys and reports its health per environment. Security Review reports exploitable bugs on every pull request. + Both are available today on Teams and Enterprise plans. + #Rollouts + Rollouts attaches a monitor to every pull request and watches the change as it deploys, reporting change health per environment: verified healthy, regression detected, or inconclusive. It's the Cursor version of Firetiger Change Monitors, rebuilt with the Bot Development Kit. + Enable it from the dashboard and connect source control, your deploy system, and your telemetry provider. Rollouts starts watching on the next pull request. + #Monitoring plans + When a pull request opens, Rollouts reads the diff and the systems it touches, then writes a monitoring plan as a PR comment. The plan lists the risks it identified, the effect the change is meant to have, the signals it will check, and any gaps in instrumentation that would make the change hard to verify. Edit the plan in the PR and Rollouts uses your version. + #Deploy tracking + Rollouts wakes on deploy events for the change's commit and runs the plan against your logs, metrics, and traces. It tracks each environment separately, so a change can be verified in staging and still flagged in production. Rollouts checks the change's intended effect alongside error and latency signals, and reports back on the PR when it reaches a verdict. + #Regressions
No changes match these filters.