Releases
75 matching — model launches, API releases, product updates, financial events and news.
June 20253
Jun 12, 2025
Meta invests $14.3 billion in Scale AI at valuation over $29 billionFunding round
Meta made a $14.3 billion strategic investment for a 49% non-voting minority stake in Scale AI, valuing the independent company at over $29 billion and expanding the commercial partnership.
Jun 12, 2025
Alexandr Wang moves to Meta; Jason Droege becomes interim CEOLeadership change
Founder and CEO Alexandr Wang left day-to-day management to join Meta’s AI efforts while remaining a Scale director; Chief Strategy Officer Jason Droege became interim CEO.
Jun 2025
Physical AI Data Engine introducedProduct launch
Scale formalized a data-engine offering for robotics and embodied AI, covering multimodal collection, annotation, and real-world model improvement.
March 20251
Mar 5, 2025
Scale selected to lead Project ThunderforgePartnership
The U.S. Defense Innovation Unit selected Scale to lead Project Thunderforge with Anduril and Microsoft, integrating AI agents into operational military planning.
February 20253
Feb 23, 2025
Scale signs five-year Qatar government AI partnershipPartnership
Scale entered a five-year partnership with Qatar’s Ministry of Communications and Information Technology to deploy AI applications, analytics, automation, and workforce programs across government.
Feb 11, 2025
MASK belief-alignment benchmark releasedBenchmark result
Scale published MASK as part of its safety research program for testing whether models knowingly state beliefs inconsistent with their internal representations.
Feb 11, 2025
FORTRESS benchmark releasedBenchmark result
Scale released FORTRESS to evaluate frontier-model safeguards against dual-use national-security and public-safety risks.
January 20252
Jan 23, 2025
Humanity's Last Exam results and benchmark releasedBenchmark result
Scale and the Center for AI Safety published Humanity’s Last Exam, assembled from nearly 1,000 contributors across more than 500 institutions in 50 countries.
Jan 2025
Scale expands frontier evaluation analyticsFeature update
Scale added richer model comparison, failure analysis, cross-sectional reporting, and automated error-discovery capabilities to its frontier evaluation framework.
November 20241
Nov 4, 2024
Defense Llama releasedModel release
Scale released Defense Llama, a Llama 3-based model fine-tuned for controlled U.S. national-security use cases inside Scale Donovan.
September 20242
Sep 2024
Scale and Center for AI Safety develop Humanity's Last ExamPartnership
Scale and the Center for AI Safety launched an international effort to create a broad expert-written benchmark for frontier AI systems.
Sep 2024
EnigmaEval benchmark releasedBenchmark result
Scale and research collaborators introduced EnigmaEval, a multimodal reasoning benchmark built from novel puzzle-competition problems.
August 20241
Aug 2024
MultiChallenge benchmark releasedBenchmark result
Scale released MultiChallenge to measure model performance on realistic multi-turn conversation problems involving memory, instruction retention, editing, and self-consistency.
May 20243
May 29, 2024
Scale Evaluation reaches general availabilityProduct launch
Scale made its model-evaluation platform generally available for model developers, enterprises, and public-sector organizations.
May 29, 2024
SEAL Leaderboards launchedProduct launch
Scale launched expert-evaluated leaderboards for coding, instruction following, mathematics, and multilinguality using private, contamination-resistant prompt sets.
May 21, 2024
Series F financing at $13.8 billion valuationFunding round
Scale closed a $1 billion primary-and-secondary financing led by Accel at a $13.8 billion valuation.
January 20241
2024
Scale expands GenAI Platform into custom enterprise agentsFeature update
Scale expanded its generative-AI platform and services into custom agentic applications connected to enterprise data, controls, and workflows.
November 20231
Nov 8, 2023
SEAL research lab launchedProduct launch
Scale launched its Safety, Evaluations, and Alignment Lab to build standardized frontier-model benchmarks, evaluation products, and red-team methods.
August 20233
Aug 24, 2023
OpenAI selects Scale as preferred fine-tuning partnerPartnership
OpenAI named Scale a preferred partner for enterprise fine-tuning of GPT-3.5, combining Scale’s data engine with OpenAI models.
Aug 11, 2023
Scale supports first public generative-AI red-team challenge at DEF CONEvent
Scale provided evaluation infrastructure for the White House-supported Generative Red Team Challenge at DEF CON 31, testing frontier models with thousands of participants.
Aug 2023
Scale introduces LLM test and evaluation platformProduct launch
Scale introduced an expert-driven testing and evaluation platform for measuring frontier-model capability, safety, bias, and robustness.
May 20233
May 10, 2023
Scale Enterprise Generative AI Platform launchedProduct launch
Scale launched EGP as a model-agnostic enterprise platform for deploying secure generative-AI applications with proprietary data, fine-tuning, evaluation, and red teaming.
May 10, 2023
Scale Donovan launchedProduct launch
Scale launched Donovan, a secure decision-support platform using large language models, retrieval, and mission data for government and defense users.
May 10, 2023
Donovan deployed on classified U.S. Army networkPartnership
Scale deployed Donovan for the U.S. Army XVIII Airborne Corps, describing it as the first large-language-model platform deployed on a classified network for the Corps.
April 20231
Apr 26, 2023
Anthropic and Scale AI announce enterprise partnershipPartnership
Scale partnered with Anthropic to help enterprises customize and deploy generative-AI applications powered by Claude.

