Releases
75 matching — model launches, API releases, product updates, financial events and news.
September 20261
Sep 22, 2026
Scale and Google Cloud publish enterprise-agent reference architecturePartnership
Scale and Google Cloud published a joint architecture for deploying Scale’s GenAI Portfolio on Google Cloud with Gemini Enterprise, identity, governance, and cloud infrastructure.
July 20261
Jul 30, 2026
Francis deSouza appointed CEOLeadership change
Scale’s board appointed former Google Cloud and Illumina executive Francis deSouza as chief executive officer, effective August 10, 2026.
May 20265
May 20, 2026
Scale GenAI Platform added to GSA SchedulePartnership
Scale partnered with Carahsoft to make its GenAI Platform, test-and-evaluation capabilities, and deployment services available through U.S. government procurement schedules.
May 14, 2026
Scale marks ten years in operationEvent
Scale marked its tenth anniversary, reporting more than 700 customers served, 15 billion human decisions applied to model development, work across 150 languages and locales, and over $1 billion paid to contributors.
May 7, 2026
SWE Atlas completed with testing and refactoring evaluationsFeature update
Scale completed SWE Atlas by adding Test Writing and Refactoring alongside Codebase QnA, producing a 284-task software-engineering evaluation suite.
May 6, 2026
Pentagon expands Scale enterprise agreement ceiling to $500 millionPartnership
The U.S. defense department increased the potential ceiling of its Scale production agreement from $100 million to $500 million following rapid adoption across military components.
May 1, 2026
Scale signs Department of Energy Genesis Mission MOUPartnership
Scale signed a memorandum of understanding with the U.S. Department of Energy to explore AI-ready scientific data infrastructure, evaluations, and model applications for the Genesis Mission.
April 20261
Apr 20, 2026
Scale acquires ICG SolutionsAcquisition
Scale acquired ICG Solutions and its LUX real-time streaming-data analytics platform to expand its national-security AI stack; ICG initially remained a wholly owned subsidiary.
March 20266
Mar 23, 2026
MultiChallenge evaluation pipeline updatedFeature update
Scale Labs updated MultiChallenge with a stronger judge model, refined tasks, and refreshed results to improve agreement with expert ratings.
Mar 16, 2026
Scale partners with Universal Robots for industrial physical AIPartnership
Scale integrated its Physical AI Data Engine with Universal Robots’ AI Trainer to collect production-grade visual and force-feedback data on industrial robots.
Mar 9, 2026
Agentic Rubrics introducedResearch paper
Scale Labs introduced repository-grounded rubrics for evaluating and reranking candidate software patches without executing tests.
Mar 9, 2026
VeRO evaluation framework introducedResearch paper
Scale Labs introduced VeRO, a reproducible harness using versioned agent snapshots, controlled budgets, structured traces, and reference procedures.
Mar 9, 2026
Scale Labs launchedProduct launch
Scale expanded SEAL into Scale Labs, a broader research organization focused on evaluation, agents, multimodal systems, post-training, deployments, and oversight.
Mar 4, 2026
SWE Atlas launches with Codebase QnABenchmark result
Scale launched SWE Atlas, initially releasing Codebase QnA to evaluate how coding agents investigate and reason about real software systems.
February 20261
Feb 27, 2026
Scale RL Environments launchedProduct launch
Scale launched high-fidelity simulated environments for training and testing agents on tool use, computer use, coding, and professional workflows.
January 20262
Jan 23, 2026
Long-Horizon Augmented Workflows releasedBenchmark result
Scale researchers released LHAW, a framework for generating underspecified long-horizon tasks and evaluating whether agents clarify ambiguity and recover performance.
Jan 22, 2026
Scale reports strongest financial year with over $1 billion in new businessNews
Scale reported that 2025 was its strongest financial year, with more than $1 billion in new business, profitable data operations, and accelerating applications and government work.
December 20251
Dec 2025
ALIF government AI fluency program launches in QatarPartnership
Qatar’s Ministry of Communications and Information Technology launched the Arabic-first ALIF AI fluency program with Scale AI for government professionals.
November 20251
Nov 2025
Scale expands offices across four global hubsEvent
Scale announced expansion across New York, London, Washington, D.C., Doha, and St. Louis to support enterprise, government, and international growth.
September 20255
Sep 19, 2025
MCP Atlas releasedBenchmark result
Scale released MCP Atlas to evaluate how well AI agents combine tools across real Model Context Protocol servers.
Sep 19, 2025
Agentic Leaderboards launchedFeature update
Scale expanded the SEAL Leaderboards into agent evaluation, beginning with SWE-Bench Pro and MCP Atlas.
Sep 19, 2025
SWE-Bench Pro releasedBenchmark result
Scale released a harder, contamination-resistant benchmark using real and commercial repositories to measure software agents on complex multi-file engineering tasks.
Sep 2025
Scale updates mission around reliable AI systemsNews
Scale updated its mission to developing reliable AI systems for the world’s most important decisions, reflecting its expansion from data infrastructure into applications and evaluation.
Sep 2025
Pentagon awards Scale enterprise agreement with $100 million ceilingPartnership
The U.S. defense department awarded Scale a five-year production agreement with an initial $100 million ceiling for data operations, computer vision, GenAI, and Donovan capabilities.
July 20251
Jul 16, 2025
Scale restructures generative-AI data businessNews
Scale reduced approximately 200 employee roles and contractor assignments after determining it had expanded its generative-AI capacity too quickly, while prioritizing enterprise, government, and international growth.
June 20253
Jun 12, 2025
Meta invests $14.3 billion in Scale AI at valuation over $29 billionFunding round
Meta made a $14.3 billion strategic investment for a 49% non-voting minority stake in Scale AI, valuing the independent company at over $29 billion and expanding the commercial partnership.
Jun 12, 2025
Alexandr Wang moves to Meta; Jason Droege becomes interim CEOLeadership change
Founder and CEO Alexandr Wang left day-to-day management to join Meta’s AI efforts while remaining a Scale director; Chief Strategy Officer Jason Droege became interim CEO.
Jun 2025
Physical AI Data Engine introducedProduct launch
Scale formalized a data-engine offering for robotics and embodied AI, covering multimodal collection, annotation, and real-world model improvement.
March 20251
Mar 5, 2025
Scale selected to lead Project ThunderforgePartnership
The U.S. Defense Innovation Unit selected Scale to lead Project Thunderforge with Anduril and Microsoft, integrating AI agents into operational military planning.
February 20253
Feb 23, 2025
Scale signs five-year Qatar government AI partnershipPartnership
Scale entered a five-year partnership with Qatar’s Ministry of Communications and Information Technology to deploy AI applications, analytics, automation, and workforce programs across government.
Feb 11, 2025
MASK belief-alignment benchmark releasedBenchmark result
Scale published MASK as part of its safety research program for testing whether models knowingly state beliefs inconsistent with their internal representations.
Feb 11, 2025
FORTRESS benchmark releasedBenchmark result
Scale released FORTRESS to evaluate frontier-model safeguards against dual-use national-security and public-safety risks.
January 20252
Jan 23, 2025
Humanity's Last Exam results and benchmark releasedBenchmark result
Scale and the Center for AI Safety published Humanity’s Last Exam, assembled from nearly 1,000 contributors across more than 500 institutions in 50 countries.
Jan 2025
Scale expands frontier evaluation analyticsFeature update
Scale added richer model comparison, failure analysis, cross-sectional reporting, and automated error-discovery capabilities to its frontier evaluation framework.
November 20241
Nov 4, 2024
Defense Llama releasedModel release
Scale released Defense Llama, a Llama 3-based model fine-tuned for controlled U.S. national-security use cases inside Scale Donovan.
September 20242
Sep 2024
Scale and Center for AI Safety develop Humanity's Last ExamPartnership
Scale and the Center for AI Safety launched an international effort to create a broad expert-written benchmark for frontier AI systems.
Sep 2024
EnigmaEval benchmark releasedBenchmark result
Scale and research collaborators introduced EnigmaEval, a multimodal reasoning benchmark built from novel puzzle-competition problems.
August 20241
Aug 2024
MultiChallenge benchmark releasedBenchmark result
Scale released MultiChallenge to measure model performance on realistic multi-turn conversation problems involving memory, instruction retention, editing, and self-consistency.
May 20243
May 29, 2024
Scale Evaluation reaches general availabilityProduct launch
Scale made its model-evaluation platform generally available for model developers, enterprises, and public-sector organizations.
May 29, 2024
SEAL Leaderboards launchedProduct launch
Scale launched expert-evaluated leaderboards for coding, instruction following, mathematics, and multilinguality using private, contamination-resistant prompt sets.
May 21, 2024
Series F financing at $13.8 billion valuationFunding round
Scale closed a $1 billion primary-and-secondary financing led by Accel at a $13.8 billion valuation.
January 20241
2024
Scale expands GenAI Platform into custom enterprise agentsFeature update
Scale expanded its generative-AI platform and services into custom agentic applications connected to enterprise data, controls, and workflows.
November 20231
Nov 8, 2023
SEAL research lab launchedProduct launch
Scale launched its Safety, Evaluations, and Alignment Lab to build standardized frontier-model benchmarks, evaluation products, and red-team methods.
August 20233
Aug 24, 2023
OpenAI selects Scale as preferred fine-tuning partnerPartnership
OpenAI named Scale a preferred partner for enterprise fine-tuning of GPT-3.5, combining Scale’s data engine with OpenAI models.
Aug 11, 2023
Scale supports first public generative-AI red-team challenge at DEF CONEvent
Scale provided evaluation infrastructure for the White House-supported Generative Red Team Challenge at DEF CON 31, testing frontier models with thousands of participants.
Aug 2023
Scale introduces LLM test and evaluation platformProduct launch
Scale introduced an expert-driven testing and evaluation platform for measuring frontier-model capability, safety, bias, and robustness.
May 20233
May 10, 2023
Scale Enterprise Generative AI Platform launchedProduct launch
Scale launched EGP as a model-agnostic enterprise platform for deploying secure generative-AI applications with proprietary data, fine-tuning, evaluation, and red teaming.
May 10, 2023
Scale Donovan launchedProduct launch
Scale launched Donovan, a secure decision-support platform using large language models, retrieval, and mission data for government and defense users.
May 10, 2023
Donovan deployed on classified U.S. Army networkPartnership
Scale deployed Donovan for the U.S. Army XVIII Airborne Corps, describing it as the first large-language-model platform deployed on a classified network for the Corps.
April 20231
Apr 26, 2023
Anthropic and Scale AI announce enterprise partnershipPartnership
Scale partnered with Anthropic to help enterprises customize and deploy generative-AI applications powered by Claude.

