AI API pricing and model changes

BriefPanel reads the official pricing pages and changelogs of the major AI API providers once a day and records what changed: a new model, a deprecation, a price that moved. Every item below comes straight from the provider's own page and shows when that page was last checked.

9 official sourcesno changes in the last 7 days

Start free

No credit card required · Set up in 10 minutes

60 items · 9 sources

Wed, Sep 30

52 items
  • New

    Deprecation announcement: Veo models

    Veo models (`veo-2.0-generate-001`, `veo-3.0-generate-001`, `veo-3.0-fast-generate-001`) will be shut down on June 30, 2026. Update to Veo 3.1 preview or 3.1 GA models via the Gemini Enterprise Agent Platform.

  • New

    Deprecation announcement: Imagen 4 and Gemini 3 Image models

    Imagen 4 and Gemini 3 Image models (`imagen-4.0-generate-001`, `imagen-4.0-ultra-generate-001`, `imagen-4.0-fast-generate-001`) will be shut down on August 17, 2026. Migrate to newer stable or preview endpoints.

  • New

    Gemini 3.8 Flash generally available (GA)

    Released `gemini-3.8-flash`, the most intelligent Flash model engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows.

  • New

    Video-to-image generation support

    You can now pass a video file (direct upload or YouTube URL) as multimodal context to generate thumbnails, posters, or infographics. Supported exclusively on the `gemini-3.1-flash-image` model.

  • New

    Gemini 3.6 Flash and Gemini 3.5 Flash-Lite generally available (GA)

    Released stable versions of Gemini 3.6 Flash (`gemini-3.6-flash`) with improved token efficiency and reduced output verbosity, and Gemini 3.5 Flash-Lite (`gemini-3.5-flash-lite`) as a low-latency, cost-effective subagent option. Also deprecated sampling parameters `temperature`,…

  • New

    Gemini Omni Flash in public preview

    Released `gemini-omni-flash-preview`, a high-performance multimodal model for high-speed video generation and conversational video editing via the Interactions API. Supports 3–10 second video generation from text or still images and conversational refinement.

  • New

    Deprecation announcement: GMP Contextual View tool

    The experimental GMP Contextual View tool (for Grounding with Google Maps outputs) will be shut down on June 15, 2026.

  • New

    Gemini 3.5 Flash

    Released `gemini-3.5-flash`, the GA version of the most intelligent model for sustained frontier performance on agentic and coding tasks. Now the model behind `gemini-flash-latest`.

  • New

    Gemini 3.6 Flash

    Input price: Free of charge (Free Tier); $0.75 through December 31, 2026, $1.50 starting January 1, 2027 (Paid Tier). Output price: Free of charge (Free Tier); $3.75 through December 31, 2026, $7.50 starting January 1, 2027 (Paid Tier). Context caching price: Free of charge (Fre…

  • New

    Gemini 3.8 Flash

    Input price: Free of charge (Free Tier); $0.75 through December 31, 2026, $1.50 starting January 1, 2027 (Paid Tier). Output price: Free of charge (Free Tier); $3.75 through December 31, 2026, $7.50 starting January 1, 2027 (Paid Tier). Context caching price: Free of charge (Fre…

  • New

    Gemini 3.7 Flash

    Input price: Free of charge (Free Tier); $0.75 through December 31, 2026, $1.50 starting January 1, 2027 (Paid Tier). Output price: Free of charge (Free Tier); $3.75 through December 31, 2026, $7.50 starting January 1, 2027 (Paid Tier). Context caching price: Free of charge (Fre…

  • New

    Enterprise tier

    For enterprise deployments, powered by Gemini Enterprise Agent Platform. All features in Paid, plus optional access to: Dedicated support channels. Advanced security & compliance. Provisioned throughput. Volume-based discounts (based on usage). ML ops, model garden and more.

  • New

    Free tier

    For developers and small projects getting started with the Gemini API. Limited access to certain models. Free input & output tokens. Google AI Studio access. Content used to improve our products.

  • New

    Paid tier

    For production applications that require higher volumes and advanced features. Higher rate limits for production deployments. Access to Context caching. Batch API (50% cost reduction). Access to Google's most advanced models. Content not used to improve our products.

  • New

    Gemini 3.5 Flash

    Input price: Free of charge (Free Tier); $1.50 (Paid Tier). Output price: Free of charge (Free Tier); $9.00 (Paid Tier). Context caching price: Free of charge (Free Tier); $0.15; $1.00 / 1,000,000 tokens per hour (storage price) (Paid Tier). Grounding with Google Search*: Not av…

  • New

    We released Mistral Large 3 ( mistral-large-2512 ) and Ministral 3 ( ministral-3b-2512 , ministral-8b-2512 and ministral-14b-2512 ). MODEL RELEASED

    We released Mistral Large 3 (mistral-large-2512) and Ministral 3 (ministral-3b-2512, ministral-8b-2512 and ministral-14b-2512).

  • New

    We released Voxtral Mini Transcribe 2 ( voxtral-mini-2602 ) and Voxtral Mini Transcribe Realtime ( voxtral-mini-transcribe-realtime-2602 ). MODEL RELEASED

    We released Voxtral Mini Transcribe 2 (voxtral-mini-2602) and Voxtral Mini Transcribe Realtime (voxtral-mini-transcribe-realtime-2602). Introducing context biasing in our Audio Transcriptions API, Introducing diarize in our Audio Transcriptions API.

  • New

    Mistral Small 4 and Leanstral released

    We released Mistral Small 4 (mistral-small-2603), a hybrid model unifying instruct, reasoning, and coding in a single multimodal model with a 256k context window. We released Leanstral (labs-leanstral-2603), our first open-source code agent designed for Lean 4 formal proof engin…

  • New

    computer use

    Agents can complete tasks in an OpenAI-hosted browser, with website access approvals and sign-in handled by your application.

  • New

    We released inline batching allowing the creation of batch jobs without file uploading. API UPDATED

    We released inline batching allowing the creation of batch jobs without file uploading.

How this radar works

Straight from the official source

Every item is read from the official page and links back to it. Summaries are automatic; when in doubt, the source is what counts.

Automatic checks

Every page is checked automatically and often. The time of the last check shows on each source.

What changed, in one line

When an item that was already published changes, it shows as Updated, with the difference summarized.

The official page prevails. This summary may be out of date or incomplete: always check the source before acting.

Track the pages that matter to you

BriefPanel reads the page you pick, as often as you choose, and shows the before and after of every field that changed.

Start free

No credit card required · Set up in 10 minutes

AI API pricing and model changes