Skip to content
AIMarketCap

Releases

123 matching — model launches, API releases, product updates, financial events and news.

Clear filters

June 20251

Jun 4, 2025CoreWeave, NVIDIA, and IBM set Blackwell MLPerf training recordsCoreWeave · CoreWeave Compute

The companies submitted the largest NVIDIA Blackwell MLPerf Training v5.0 cluster and reported record Llama 3.1 405B training performance.

May 20252

May 16, 2025Unauthorized Grok prompt modification disclosedSpaceXAI · Grok

xAI disclosed that an unauthorized system-prompt change caused Grok to inject unrelated claims about South African racial politics; it published prompts and announced stronger review and monitoring controls.

May 14, 2025AlphaEvolve introducedGoogle (Alphabet) · AlphaEvolve

Google DeepMind introduces AlphaEvolve, a Gemini-powered evolutionary coding agent for algorithm discovery and optimization.

April 20251

Apr 2, 2025CoreWeave posts first cloud NVIDIA GB200 MLPerf inference resultsCoreWeave · CoreWeave Compute

CoreWeave submitted the first cloud-provider MLPerf Inference v5.0 results for NVIDIA GB200, reporting more than 800 tokens per second on Llama 3.1 405B.

March 20252

Mar 27, 2025Tracing the thoughts of a large language model publishedAnthropic · Claude

Anthropic published interpretability research tracing internal mechanisms underlying multilingual reasoning, planning, and unfaithful explanations.

Mar 2025Kimina-Prover project introducedMoonshot AI · Kimina-Prover

Moonshot introduced the Kimina-Prover line of open models for formal theorem proving.

February 20257

Feb 19, 2025Muse world and action model introducedMicrosoft · Muse

Microsoft Research and Xbox introduce Muse, a generative world and action model trained on gameplay sequences.

Feb 11, 2025MASK belief-alignment benchmark releasedScale AI · MASK

Scale published MASK as part of its safety research program for testing whether models knowingly state beliefs inconsistent with their internal representations.

Scale AI Benchmark resultScale AIMASK
Feb 11, 2025FORTRESS benchmark releasedScale AI · FORTRESS

Scale released FORTRESS to evaluate frontier-model safeguards against dual-use national-security and public-safety risks.

Feb 10, 2025Anthropic Economic Index launchedAnthropic

Anthropic launched the Economic Index to measure how AI is being used across occupations and tasks.

Anthropic Research paperAnthropic
Feb 7, 2025PARTNR benchmark and dataset releasedMeta · PARTNR

Meta released PARTNR, a benchmark, dataset and planning model for human-robot collaboration on household tasks.

Meta AI Research paperMetaPARTNR
Feb 6, 2025Brain2Qwerty research publishedMeta · Brain2Qwerty · v1

Meta researchers introduced a non-invasive deep-learning system for decoding typed sentences from EEG and MEG brain recordings.

Feb 3, 2025Constitutional Classifiers research publishedAnthropic · Claude

Anthropic published and tested Constitutional Classifiers, a safeguard approach designed to resist broad classes of jailbreaks.

January 20255

Jan 23, 2025Humanity's Last Exam results and benchmark releasedScale AI · Humanity's Last Exam

Scale and the Center for AI Safety published Humanity’s Last Exam, assembled from nearly 1,000 contributors across more than 500 institutions in 50 countries.

Jan 20, 2025DeepSeek R1 reasoning family launchesDeepSeek · DeepSeek R1

DeepSeek introduced its reinforcement-learning reasoning model family and technical report.

Jan 16, 2025MatterGen research publishedMicrosoft · MatterGen

Microsoft Research publishes MatterGen, a generative model that designs stable inorganic materials conditioned on target properties.

Jan 2025Kimi k1 reasoning model disclosedMoonshot AI · Kimi k1

Moonshot described k1 as a long-context reinforcement-learning reasoning system.

2025TRIBE research model recognized at Algonauts 2025Meta · TRIBE · v1

Meta's first TRIBE brain-response model formed the foundation for its award-winning Algonauts 2025 entry; Meta's later TRIBE v2 announcement identifies the original model as 2025 work.

Meta AI Benchmark resultMetaTRIBE

December 20244

Dec 26, 2024DeepSeek V3 family launchesDeepSeek · DeepSeek V3

DeepSeek introduced its 671B-parameter V3 mixture-of-experts architecture.

Dec 18, 2024Alignment faking research publishedAnthropic · Claude 3.5 Sonnet (October 2024)

Anthropic and Redwood Research published evidence of alignment-faking behavior in a controlled Claude 3 Opus experiment.

Dec 4, 2024BioEmu generative protein model introducedMicrosoft · BioEmu

Microsoft Research introduces BioEmu for efficiently sampling diverse protein conformational ensembles.

Dec 4, 2024Genie 2 world model introducedGoogle (Alphabet) · Genie 2

Google DeepMind introduces Genie 2, capable of generating diverse playable 3D environments from a prompt image.

November 20241

Nov 2024Kimi k0-math research model disclosedMoonshot AI · Kimi k0-math

Moonshot disclosed k0-math as an experimental reinforcement-learning model for mathematical reasoning.

October 20242

Oct 18, 2024Janus multimodal family introducedDeepSeek · Janus

DeepSeek introduced the Janus family for unified multimodal understanding and generation.

Oct 18, 2024Spirit LM introducedMeta · Spirit LM

Meta released research on a multimodal language model that freely mixes text and speech.

Meta AI Research paperMetaSpirit LM
26–50 of 123 releases