Releases
4 matching — model launches, API releases, product updates, financial events and news.
September 20254
Sep 19, 2025
MCP Atlas releasedBenchmark result
Scale released MCP Atlas to evaluate how well AI agents combine tools across real Model Context Protocol servers.
Sep 19, 2025
Agentic Leaderboards launchedFeature update
Scale expanded the SEAL Leaderboards into agent evaluation, beginning with SWE-Bench Pro and MCP Atlas.
Sep 19, 2025
SWE-Bench Pro releasedBenchmark result
Scale released a harder, contamination-resistant benchmark using real and commercial repositories to measure software agents on complex multi-file engineering tasks.
Sep 19, 2025
Grok 4 Fast releasedModel release
xAI released Grok 4 Fast with a 2M-token context window, unified reasoning modes and lower-cost agentic search.

