Release notes
Official updates from Anthropic, OpenAI, Google, xAI, Mistral, DeepSeek and Meta, newest first. Open any update for a short AI summary and the full notes.
Anthropic · Claude API
Google · Gemini API·Apr 9, 2025
Released veo-2.0-generate-001, a generally available (GA) text- and image-to-video model, capable of generating detailed and artistically nuanced videos. To learn more, see the Veo docs. Released gemini-2.0-flash-live-001, a public preview version of the Live API model with billing enabled. Enhanced Session Management and Reliability Session Resumption: Keep sessions alive across temporary network disruptions. The API now supports server-side session state storage (for up to 24 hours) and provides handles (session_resumption) to reconnect and resume where you left off. Longer Sessions via Context Compression: Enable extended interactions beyond previous time limits. Configure context window compression with a sliding window mechanism to automatically manage context length, preventing abrupt terminations due to context limits. Graceful Disconnect Notification: Receive a GoAway server message indicating when a connection is about to close, allowing for graceful handling before termination. More Control over Interaction Dynamics Configurable Voice Activity Detection (VAD): Choose sensitivity levels or disable automatic VAD entirely and use new client events (activityStart, activityEnd) for manual turn control. Configurable Interruption Handling: Decide whether user input should interrupt the model's response. Configurable Turn Coverage: Choose whether the API processes all audio and video input continuously or only captures it when the end-user is detected speaking. Configurable Media Resolution: Optimize for quality or token usage by selecting the resolution for input media. Richer Output and Features Expanded Voice & Language Options: Choose from two new voices and 30 new languages for audio output. The output language is now configurable within speechConfig. Text Streaming: Receive text responses incrementally as they are generated, enabling faster display to the user. Token Usage Reporting: Gain insights into usage with detailed token counts provided in the usageMetadata field of server messages, broken down by modality and prompt or response phases.
Google · Gemini API·Apr 4, 2025
Released gemini-2.5-pro-preview-03-25, a public preview Gemini 2.5 Pro version with billing enabled. You can continue to use gemini-2.5-pro-exp-03-25 on the free tier.
xAI · New model·Apr 1, 2025
Our latest flagship Grok 3 models are now generally available via the API. For more info, see models.
Anthropic · Claude API·Mar 31, 2025
We've moved our Go SDK from alpha to beta.
Google · Gemini app·Mar 29, 2025
What: Today we’re expanding access to our most intelligent AI model, 2.5 Pro (experimental), to all Gemini users. This state-of-the-art model has thinking capabilities natively built in, with exceptional performance in coding, math, image understanding and more. Canvas, our new interactive space that makes it easy to create, refine, and share your work, is now available to try with 2.5 Pro (experimental). With 2.5 Pro’s improved coding capabilities, coupled with Canvas, you can quickly create compelling web apps or generate code for immediate use. Gemini users will be able to try 2.5 Pro (experimental) with rate limits, and Gemini Advanced users will continue to have expanded access and a significantly larger context window. Being an experimental model, it can have unexpected behaviors and may make mistakes. Why: We want to bring the best model in the world to all Gemini users. Your feedback helps us improve these models over time and learning from experimental launches informs how we release models more widely.
Google · Gemini API·Mar 25, 2025
Released gemini-2.5-pro-exp-03-25, a public experimental Gemini model with thinking mode always on by default. To learn more, see Gemini 2.5 Pro Experimental.
Google · Gemini app·Mar 25, 2025
What: Today we’re introducing Gemini 2.5, our most intelligent AI model. Our first 2.5 release is a chat optimized version of Gemini-2.5-Pro-Exp-03-25, which is state-of-the-art on a wide range of benchmarks and debuts at #1 on LMArena by a significant margin. This model also has thinking capabilities natively built in, with improved performance in complex tasks like coding, math, and image understanding. 2.5 Pro (experimental) is now rolling out to the Gemini web and mobile app and is available to qualifying Google Workspace business and education plans. This experimental model is meant to be an early preview and can have unexpected behaviors and may make mistakes. Why: We believe in rapid iteration and bringing the best of Gemini to the world, and we want to give Gemini Advanced subscribers priority access to our latest AI innovations. Your feedback helps us improve these models over time and learning from experimental launches informs how we release models more widely.
DeepSeek · New model·Mar 24, 2025
deepseek-chat Model Upgraded to DeepSeek-V3-0324: Enhanced Reasoning Capabilities
Google · Gemini app·Mar 18, 2025
What: Starting today, you can collaborate with Gemini 2.0 Flash to write documents and code in Canvas, a new interactive space that makes refining & sharing your work really easy. Co-create documents: Generate a first draft, then rapidly refine and ask Gemini for feedback on your edits. Update specific sections or the whole draft, and use the quick editor tools to change the tone, length, or formatting. From essays to blog posts to reports, uplevel your documents in Canvas. Generate & iterate on code: Easily convert your ideas to working prototypes for web apps, Python scripts, and more. Ask Gemini to generate & preview React or HTML code directly in Canvas in a familiar code editor, and review Gemini's changes on each turn. Canvas is available globally across all Gemini-supported languages. Select Canvas in the prompt bar to start creating! Why: Canvas lets you experience the power of collaboration with Gemini because you can see your ideas and iterations take shape in real time. Go from blank slate to a share-worthy creation in minutes. This means you can focus on your vision to create something awesome and leave the heavy lifting of generating, editing, and fixing things to Gemini.
Mistral AI · New model·Mar 17, 2025
We released Mistral Small 3.1 (mistral-small-2503).
Google · Gemini app·Mar 13, 2025
What: Starting today, an improved version of Gemini 2.0 Flash Thinking (experimental) will become available to Gemini app users. Built on the foundation of 2.0 Flash, this model delivers improved performance and better advanced reasoning capabilities with efficiency and speed. Starting in English, 2.0 Flash Thinking (experimental) now works with your favorite Gemini features and connected apps such as YouTube, Maps, Search and more. Gemini Advanced users will also have access to a 1M token context window with this model. Why: We're investing in thinking and reasoning capabilities because we believe they unlock deeper intelligence and deliver enhanced performance for tasks requiring complex reasoning, such as coding, scientific discovery, and advanced math.
Google · Gemini API·Mar 12, 2025
Launched an experimental Gemini 2.0 Flash model capable of image generation and editing. Released gemma-3-27b-it, available on AI Studio and through the Gemini API, as part of the Gemma 3 launch. Added support for YouTube URLs as a media source. Added support for including an inline video of less than 20MB.
Google · Gemini API·Mar 11, 2025
Released the Google Gen AI SDK for TypeScript and JavaScript to public preview.
Google · Gemini API·Mar 7, 2025
Released gemini-embedding-exp-03-07, an experimental Gemini-based embeddings model in public preview.
Mistral AI · New model·Mar 6, 2025
We released Mistral OCR (mistral-ocr-2503) and document understanding.
Google · Gemini app·Mar 3, 2025
What: Gemini can now connect to more apps and services on Android. With the new Spotify extension, you can play your favorite songs and discover playlists for any mood With the Phone, Messages, and WhatsApp extensions, you can quickly make calls and draft, refine, and send messages with your default phone and messaging apps, as well as WhatsApp With the Utilities extension, you can use Gemini to set alarms, control your device settings, and even open your camera to take a quick selfie Phone, Messages, WhatsApp, Utilities, Calendar, Tasks, and Keep extensions are now available in all Gemini languages Why: We're continuing to build new ways for Gemini to connect to the apps you rely on every day. Through these extensions, Gemini can help you get things done and take actions for you, better than ever. Learn more about extensions.
xAI · New model·Mar 1, 2025
The image generation model is available on API. Visit Image Generations for more details on using the model.
Google · Gemini API·Feb 28, 2025
Support for Search as a tool added to gemini-2.0-pro-exp-02-05, an experimental model based on Gemini 2.0 Pro.
Anthropic · Claude API·Feb 27, 2025
You can now reference images and PDFs directly through a URL instead of having to base64-encode them. Learn more in Vision and PDF support. We've added support for a none option to the tool_choice parameter in the Messages API that prevents Claude from calling any tools. We've launched an OpenAI-compatible API endpoint, allowing you to test Claude models by changing just your API key, base URL, and model name in existing OpenAI integrations.
Google · Gemini API·Feb 25, 2025
Released gemini-2.0-flash-lite, a generally available (GA) version of Gemini 2.0 Flash-Lite, which is optimized for speed, scale, and cost efficiency.
Anthropic · Claude API·Feb 24, 2025
Claude Sonnet 3.7 can produce near-instant responses or show its extended thinking step-by-step. One model, two ways to think. Learn more about all Claude models in Models overview. We've added vision support to Claude Haiku 3.5, enabling the model to analyze and understand images. We've released a token-efficient tool use implementation, improving overall performance when using tools with Claude. We've changed the default temperature in the Console for new prompts from 0 to 1 for consistency with the default temperature in the API. We've released updated versions of our tools that decouple the text edit and bash tools from the computer use system prompt: * bash_20250124: Same functionality as previous version but is independent from computer use.
Google · Gemini app·Feb 20, 2025
What: You can seamlessly upload multiple Google Docs, PDFs, and Word documents from Google Drive or your device into Gemini for quick summaries, personalized feedback, and actionable insights. Currently, Gemini offers a 32K context window (around 50 pages of text), with usage limits. Gemini Advanced users continue to have a 1 million token context window (around 1,500 pages of text) and much higher usage limits. Why: Unlock deeper insights from your documents and streamline your workflows with Gemini to save time and boost your productivity.
Google · Gemini API·Feb 19, 2025
Support for additional regions (Kosovo, Greenland and Faroe Islands). Support for additional regions (Kosovo, Greenland and Faroe Islands).
Google · Gemini API·Feb 18, 2025
Gemini 1.0 Pro is no longer supported. For the list of supported models, see Gemini models.
Mistral AI · New model·Feb 17, 2025
We released Mistral Saba (mistral-saba-2502).
Google · Gemini app·Feb 12, 2025
What: Starting today, Gemini can now use your past chats to give you more helpful responses. Ask a question about something you've discussed, or have Gemini summarize a previous conversation. It will use the information from relevant chats to craft its response. This means no more starting over, and you can build on top of previous conversations or projects you already started. You’re in control: easily review, delete, or decide how long to keep your chat history. You can also turn off Gemini Apps Activity altogether by going to My Activity. Gemini may indicate when it uses your past chats in Sources and related content. This experience is starting to roll out in English for Gemini Advanced subscribers via Google One AI Premium Plan. Why: Conversations change, ideas evolve, and sometimes we need to revisit what’s been said. Think of Gemini as your personal AI assistant, always there to pick up where you left off. It becomes your thought partner, evolving alongside you to unlock new levels of helpfulness, that’s unique to you.
Google · Gemini API·Feb 11, 2025
Updates on the OpenAI libraries compatibility.
Anthropic · Claude API·Feb 10, 2025
This header provides the organization ID associated with the API key used in the request.
Google · Gemini API·Feb 6, 2025
Released imagen-3.0-generate-002, a generally available (GA) version of Imagen 3 in the Gemini API. Released the Google Gen AI SDK for Java for public preview.