Skip to content
AIMarketCap

Releases

566 matching — model launches, API releases, product updates, financial events and news.

Clear filters

May 20262

May 19, 2026Gemini 3.5 Flash launched at Google I/OGoogle (Alphabet) · Gemini 3.5 Flash

Google launches Gemini 3.5 Flash across Google AI Studio, Antigravity and Gemini Enterprise Agent Platform.

May 9, 2026ERNIE 5.1 releasedBaidu · ERNIE 5.1

Baidu released a more compact flagship model with stronger agents, reasoning and creative abilities.

April 202619

Apr 29, 2026Granite 4.1 8B weights releasedIBM · Granite 4.1 8B

IBM published the 8B base and instruct Granite 4.1 models.

Apr 29, 2026Granite 4.1 30B weights releasedIBM · Granite 4.1 30B

IBM published the 30B base and instruct Granite 4.1 models.

Apr 29, 2026Granite 4.1 3B weights releasedIBM · Granite 4.1 3B

IBM published the 3B base and instruct Granite 4.1 models.

Apr 29, 2026Granite 4.1 collection releasedIBM · Granite 4.1

IBM released Granite 4.1 language, vision, speech, embedding and Guardian models under Apache 2.0.

Apr 28, 2026NVIDIA releases Nemotron 3 Nano OmniNVIDIA · Nemotron 3 Nano Omni · 3 Nano Omni

NVIDIA released an open multimodal model unifying vision, audio and language for efficient agents.

Apr 26, 2026Sora product discontinuedOpenAI · Sora app

OpenAI discontinued the standalone Sora product as its video strategy shifted to newer experiences and models.

OpenAI Model deprecationOpenAISora app
Apr 24, 2026DeepSeek V4 preview launchesDeepSeek · DeepSeek V4

DeepSeek introduced the V4 family through Pro and Flash preview API models.

Apr 23, 2026Grok Voice Think Fast 1.0 releasedSpaceXAI · Grok Voice Think Fast 1.0 · 1.0

SpaceXAI released its first named Think Fast speech-to-speech voice-agent model.

Apr 23, 2026GPT-5.5 and GPT-5.5 Pro releasedOpenAI · GPT-5.5

OpenAI released GPT-5.5 and GPT-5.5 Pro as upgraded frontier models for complex knowledge work and reasoning.

Apr 21, 2026ChatGPT Images 2.0 releasedOpenAI · ChatGPT Images 2.0

OpenAI released ChatGPT Images 2.0 with stronger world knowledge, instruction following, detail, and an images-with-thinking mode.

Apr 20, 2026Kimi K2.6 released and open-sourcedMoonshot AI · Kimi K2.6

Moonshot released Kimi K2.6 with improved agent capabilities and published model weights.

Apr 16, 2026Claude Opus 4.7 releasedAnthropic · Claude Opus 4.7 · 4.7

Anthropic released Claude Opus 4.7 with improved software engineering, vision, complex multi-step work, and new cybersecurity safeguards.

Apr 15, 2026ERNIE-Image releasedBaidu · ERNIE-Image

Baidu released an open-weight 8B DiT text-to-image model.

Apr 15, 2026Gemini 3.1 Flash Audio launchedGoogle (Alphabet) · Gemini 3.1 Flash Audio

Google launches Gemini 3.1 Flash Audio for real-time and batch audio understanding and generation.

Apr 8, 2026Muse Spark introducedMeta · Muse Spark · Spark

Meta Superintelligence Labs introduced Muse Spark, a multimodal model designed for efficient broad intelligence.

Apr 7, 2026Claude Mythos Preview released to Project Glasswing partnersAnthropic · Claude Mythos Preview · Preview

Anthropic released Claude Mythos Preview as a gated frontier model for vetted Project Glasswing partners, focusing its advanced cybersecurity capabilities on defensive use.

Apr 2, 2026MAI-Image-2 launched in Microsoft FoundryMicrosoft · MAI-Image-2

Microsoft AI launches MAI-Image-2 for faster, higher-quality image generation in Copilot and Microsoft Foundry.

Apr 2, 2026MAI-Transcribe-1 launched in Microsoft FoundryMicrosoft · MAI-Transcribe-1

Microsoft AI launches MAI-Transcribe-1 for high-speed multilingual speech-to-text through Microsoft Foundry.

Apr 2, 2026Gemma 4 releasedGoogle (Alphabet) · Gemma 4

Google releases Gemma 4 open models and makes them available through Google Cloud's AI development stack.

March 202618

Mar 27, 2026SAM 3.1 releasedMeta · Segment Anything Model 3.1 · 3.1

Meta released an efficiency update adding object multiplexing and up to 32-frame-per-second multi-object video tracking on one H100.

Mar 26, 2026Transcribe model family introducedCohere · Transcribe family

Cohere entered speech recognition with the open Transcribe model family.

Mar 26, 2026Cohere Transcribe releasedCohere · Cohere Transcribe

Cohere introduced an open-source speech-recognition model for accurate multilingual transcription.

Mar 26, 2026TRIBE v2 releasedMeta · TRIBE v2 · 2

Meta released a tri-modal foundation model that predicts high-resolution human brain activity from video, audio and language.

Mar 23, 2026Voxtral TTS releasedMistral AI · Voxtral TTS · 4B

Mistral AI released Voxtral TTS, a 4B multilingual text-to-speech model.

Mar 17, 2026GPT-5.4 mini and nano releasedOpenAI · GPT-5.4 mini

OpenAI released GPT-5.4 mini and GPT-5.4 nano as smaller, lower-latency variants.

Mar 16, 2026Leanstral releasedMistral AI · Leanstral · 1

Mistral AI released Leanstral, an open model for formal theorem proving and Lean code.

Mar 16, 2026Mistral Small 4 releasedMistral AI · Mistral Small 4 · 4

Mistral AI released Mistral Small 4, a next-generation compact open model for multimodal and agentic workloads.

Mar 16, 2026NVIDIA expands open model families for agentic, physical and healthcare AINVIDIA · NVIDIA Nemotron

NVIDIA introduced Nemotron 3 Omni and VoiceChat, GR00T N1.7, Alpamayo 1.5 and Proteina-Complexa.

Mar 11, 2026NVIDIA releases Nemotron 3 SuperNVIDIA · Nemotron 3 Super · 3 Super

NVIDIA released the open 120-billion-parameter Nemotron 3 Super model for complex agentic AI systems.

Mar 10, 2026Grok 4.20 and Grok 4.20 Multi-agent launchSpaceXAI · Grok 4.20 · 4.20

SpaceXAI released Grok 4.20 and a multi-agent variant through its platform.

Mar 10, 2026Canopy Height Maps v2 releasedMeta · Canopy Height Maps v2

Meta and World Resources Institute released a DINOv3-based global forest-canopy model with improved accuracy and detail.

Mar 5, 2026GPT-5.4 and GPT-5.4 Pro releasedOpenAI · GPT-5.4

OpenAI released GPT-5.4 and GPT-5.4 Pro for professional work, coding, computer use, and long-context tasks.

Mar 4, 2026Phi-4 Reasoning Vision releasedMicrosoft · Phi-4 Reasoning Vision

Microsoft Research releases Phi-4 Reasoning Vision, a 15-billion-parameter open-weight multimodal reasoning model.

Mar 3, 2026Gemini 3.1 Flash-Lite launchedGoogle (Alphabet) · Gemini 3.1 Flash-Lite

Google launches Gemini 3.1 Flash-Lite for low-cost, high-throughput multimodal workloads.

Mar 2026Checkpoint Engine releasedMoonshot AI · Checkpoint Engine

Moonshot published its distributed model-checkpointing engine.

Mar 2026FlashKDA releasedMoonshot AI · FlashKDA

Moonshot published optimized kernels for Kimi Delta Attention.

Mar 2026MoonEP releasedMoonshot AI · MoonEP

Moonshot released its expert-parallel communication library for large MoE training and inference.

February 20268

Feb 26, 2026Gemini 3.1 Flash Image launchedGoogle (Alphabet) · Gemini 3.1 Flash Image

Google launches Gemini 3.1 Flash Image for fast, controllable image generation and editing.

Feb 19, 2026Gemini 3.1 Pro launchedGoogle (Alphabet) · Gemini 3.1 Pro

Google launches Gemini 3.1 Pro with advances in reasoning, coding and long-running agentic tasks.

Feb 17, 2026Claude Sonnet 4.6 releasedAnthropic · Claude Sonnet 4.6 · 4.6

Anthropic released Claude Sonnet 4.6 with upgrades across coding, computer use, long-context reasoning, agent planning, knowledge work, and design.

Feb 12, 2026GPT-5.3-Codex-Spark research preview releasedOpenAI · GPT-5.3-Codex-Spark

OpenAI released a research preview of GPT-5.3-Codex-Spark, a smaller model optimized for near-real-time coding.

Feb 6, 2026ERNIE 5.0 officially releasedBaidu · ERNIE 5.0

Baidu released a 2.4T-parameter unified model trained jointly across text, image, video and audio.

Feb 5, 2026Claude Opus 4.6 releasedAnthropic · Claude Opus 4.6 · 4.6

Anthropic released Claude Opus 4.6 with stronger coding, longer-running agentic work, improved knowledge work, and a one-million-token context window in beta.

Feb 5, 2026GPT-5.3-Codex releasedOpenAI · GPT-5.3-Codex

OpenAI released GPT-5.3-Codex, an interactive agentic coding model for complex software and professional work.

Feb 4, 2026Voxtral Transcribe 2 releasedMistral AI · Voxtral Transcribe 2 · 2

Mistral AI released two second-generation Voxtral transcription models for production speech-to-text workloads.

January 20265

Jan 29, 2026PaddleOCR-VL-1.5 releasedBaidu · PaddleOCR-VL-1.5

Baidu released an upgraded 0.9B document-parsing vision-language model.

Jan 27, 2026Kimi K2.5 releasedMoonshot AI · Kimi K2.5

Moonshot released Kimi K2.5, adding native multimodal capabilities and stronger agentic coding.

Jan 27, 2026DeepSeek-OCR2 releasedDeepSeek · DeepSeek-OCR2

DeepSeek published its second-generation visual causal-flow OCR model.

Jan 14, 2026FunctionGemma releasedGoogle (Alphabet) · FunctionGemma

Google releases FunctionGemma, a compact open model specialized for function calling and on-device agent workflows.

Jan 5, 2026NVIDIA releases new physical AI models and Jetson T4000NVIDIA · Cosmos Reason 2

NVIDIA released Cosmos Transfer 2.5, Predict 2.5 and Reason 2, updated GR00T N1.6, and introduced Jetson T4000.

December 202516

Dec 18, 2025GPT-5.2-Codex releasedOpenAI · GPT-5.2-Codex

OpenAI released GPT-5.2-Codex, an agentic coding model optimized for long-horizon software engineering.

Dec 17, 2025Mistral OCR 3 releasedMistral AI · Mistral OCR 3 · 3

Mistral AI released Mistral OCR 3 with improved document parsing and structured extraction.

Dec 17, 2025Gemini 3 Flash launchedGoogle (Alphabet) · Gemini 3 Flash

Google launches Gemini 3 Flash as a fast, production-oriented model in the Gemini 3 family.

Dec 16, 2025SAM Audio releasedMeta · SAM Audio

Meta released a unified multimodal model for separating sounds from complex mixtures using natural prompts.

Dec 15, 2025NVIDIA debuts Nemotron 3 familyNVIDIA · NVIDIA Nemotron · 3

NVIDIA introduced Nemotron 3 Nano, Super and Ultra open models and companion reinforcement-learning and evaluation tools.

Dec 11, 2025Rerank 4 releasedCohere · Rerank 4

Cohere launched Rerank 4, its fourth-generation retrieval model family in Pro and Fast tiers.

Dec 11, 2025GPT-5.2 releasedOpenAI · GPT-5.2

OpenAI released GPT-5.2 with improvements for professional knowledge work and long-context tasks.

Dec 9, 2025Devstral 2 releasedMistral AI · Devstral 2 · 2

Mistral AI released Devstral 2, its second-generation open code-agent model.

Dec 2, 2025Mistral 3 model generation releasedMistral AI · Mistral Large 3 · 3

Mistral AI released the Mistral 3 generation, including open-weight Mistral Large 3 and Ministral 3 models in 3B, 8B and 14B sizes.

Dec 2, 2025Amazon Nova Multimodal Embeddings releasedAmazon · Amazon Nova Multimodal Embeddings

AWS released a unified embedding model for text, documents, images, video and audio.

Dec 2, 2025Amazon Nova 2 Sonic releasedAmazon · Amazon Nova 2 Sonic · 2 Sonic

AWS released its second-generation speech-to-speech foundation model.

Dec 2, 2025Amazon Nova 2 Omni preview announcedAmazon · Amazon Nova 2 Omni · 2 Omni

AWS introduced a unified multimodal reasoning and image-generation model in preview.

Dec 2, 2025Amazon Nova 2 Pro preview announcedAmazon · Amazon Nova 2 Pro · 2 Pro

AWS introduced its most intelligent Nova 2 model for complex multistep tasks in preview.

Dec 2, 2025Amazon Nova 2 Lite releasedAmazon · Amazon Nova 2 Lite · 2 Lite

AWS released a fast multimodal reasoning model with extended thinking, built-in tools and a one-million-token context window.

Dec 1, 2025DeepSeek-V3.2-Speciale temporary endpoint launchesDeepSeek · DeepSeek-V3.2-Speciale

DeepSeek offered a temporary high-compute reasoning endpoint until December 15.

Dec 1, 2025DeepSeek-V3.2 releasedDeepSeek · DeepSeek-V3.2

DeepSeek upgraded chat and reasoner services to the production V3.2 model.

November 202513

Nov 27, 2025DeepSeekMath-V2 releasedDeepSeek · DeepSeekMath-V2

DeepSeek released its self-verifiable mathematical reasoning model.

Nov 24, 2025Claude Opus 4.5 releasedAnthropic · Claude Opus 4.5

Anthropic released Claude Opus 4.5 for coding, agents, computer use, research, and professional work.

Nov 20, 2025Gemini 3 Pro Image launches as Nano Banana ProGoogle (Alphabet) · Gemini 3 Pro Image

Google launches Gemini 3 Pro Image, branded Nano Banana Pro, for professional image generation and editing.

Nov 19, 2025Grok 4.1 Fast and Agent Tools API launchSpaceXAI · Grok 4.1 Fast · 4.1 Fast

xAI released Grok 4.1 Fast and an Agent Tools API with X search, web search, code execution and collections search.

Nov 19, 2025SAM 3D Body releasedMeta · SAM 3D Body · Body

Meta released a model for estimating human body pose and shape from a single image.

Nov 19, 2025SAM 3D Objects releasedMeta · SAM 3D Objects · Objects

Meta released a model for reconstructing 3D objects and scenes from a single image.

Nov 19, 2025Segment Anything Model 3 releasedMeta · Segment Anything Model 3 · 3

Meta released SAM 3 for text-, exemplar- and visual-prompted detection, segmentation and tracking in images and video.

Nov 18, 2025Gemini 3 Pro launchedGoogle (Alphabet) · Gemini 3 Pro

Google launches Gemini 3 Pro for advanced multimodal reasoning, coding and agentic work.

Nov 17, 2025Grok 4.1 launchesSpaceXAI · Grok 4.1 · 4.1

xAI released Grok 4.1 across grok.com, X, iOS and Android with improvements to creative, emotional and collaborative interactions.

Nov 13, 2025GPT-5.1 releasedOpenAI · GPT-5.1

OpenAI released GPT-5.1 for developers with improved coding, steerability, and agentic performance.

Nov 11, 2025ERNIE 4.5 VL 28B A3B Thinking releasedBaidu · ERNIE 4.5 VL 28B A3B Thinking

Baidu released a sparse multimodal reasoning model activating only 3B parameters.

Nov 10, 2025Omnilingual ASR releasedMeta · Omnilingual ASR

Meta released speech-recognition models covering more than 1,600 languages plus an open corpus for 350 underserved languages.

Nov 6, 2025Kimi K2 Thinking releasedMoonshot AI · Kimi K2 Thinking

Moonshot released the reasoning-focused K2 Thinking model for long-horizon agentic tasks.

October 20259

Oct 20, 2025DeepSeek-OCR releasedDeepSeek · DeepSeek-OCR

DeepSeek published its optical-compression model for OCR and document understanding.

Oct 20, 2025Chronos-2 releasedAmazon · Chronos-2 · 2

Amazon released a universal time-series foundation model for univariate, multivariate and covariate-informed forecasting.

Oct 16, 2025PaddleOCR-VL releasedBaidu · PaddleOCR-VL

Baidu released a compact multilingual vision-language model for document parsing.

Oct 15, 2025Veo 3.1 launchedGoogle (Alphabet) · Veo 3.1

Google launches Veo 3.1 with richer audio, stronger narrative control and reference-image guidance.

Oct 15, 2025Claude Haiku 4.5 releasedAnthropic · Claude Haiku 4.5

Anthropic released Claude Haiku 4.5, offering near-frontier coding performance at lower cost and latency.

Oct 13, 2025MAI-Image-1 announcedMicrosoft · MAI-Image-1

Microsoft AI announces MAI-Image-1, its first image-generation model developed entirely in-house.

Oct 7, 2025Gemini 2.5 Computer Use released in previewGoogle (Alphabet) · Gemini 2.5 Computer Use

Google releases Gemini 2.5 Computer Use in preview for agents that interact with graphical user interfaces.

Oct 2, 2025Granite 4.0 open models releasedIBM · Granite 4.0

IBM launched hybrid Mamba/transformer Granite 4.0 models optimized for efficient agentic workloads.

IBM Open weightsIBMGranite 4.0
Oct 2025Kimi Linear 48B weights releasedMoonshot AI · Kimi Linear 48B

Moonshot published a 48B Kimi Linear checkpoint.

September 202510

Sep 30, 2025Sora 2 releasedOpenAI · Sora 2

OpenAI released Sora 2, a video and audio generation model with improved physical realism and control.

OpenAI Model releaseOpenAISora 2
Sep 29, 2025DeepSeek-V3.2-Exp releasedDeepSeek · DeepSeek-V3.2-Exp

DeepSeek released the experimental V3.2 model with sparse attention.

Sep 29, 2025NVIDIA releases GR00T N1.6 and robotics simulation librariesNVIDIA · GR00T N1.6 · N1.6

NVIDIA released GR00T N1.6, Cosmos Reason NIM and new Newton physics integration for Isaac Lab.

Sep 29, 2025Claude Sonnet 4.5 releasedAnthropic · Claude Sonnet 4.5

Anthropic released Claude Sonnet 4.5 with major gains in coding, computer use, agents, reasoning, and mathematics.

Sep 25, 2025Gemini Robotics 1.5 introducedGoogle (Alphabet) · Gemini Robotics 1.5

Google DeepMind introduces Gemini Robotics 1.5 and an embodied-reasoning model for more capable robot agents.

Sep 25, 2025Gemini 2.5 Flash-Lite reaches general availabilityGoogle (Alphabet) · Gemini 2.5 Flash-Lite

Gemini 2.5 Flash-Lite becomes generally available for high-volume, latency-sensitive applications.

Sep 22, 2025DeepSeek-V3.1-Terminus releasedDeepSeek · DeepSeek-V3.1-Terminus

DeepSeek improved language consistency and its code and search agents.

Sep 19, 2025Grok 4 Fast releasedSpaceXAI · Grok 4 Fast · 4 Fast

xAI released Grok 4 Fast with a 2M-token context window, unified reasoning modes and lower-cost agentic search.

Sep 12, 2025ERNIE 4.5 receives PLAS sparse-attention updateBaidu · ERNIE 4.5

Baidu introduced a sparse-attention inference update for faster long-context performance.

Sep 9, 2025OpenAI gpt-oss models enter OCI betaOracle · OpenAI gpt-oss on OCI

OCI Generative AI added beta access to OpenAI's gpt-oss open-weight models.

101–200 of 566 releases