SpaceXAI's First Flagship: What the xAI Rebrand Actually Ships

SpaceXAI has released its first major flagship, a 1.5-trillion-parameter model focused on programming and efficiency. Here is a breakdown of the performance and the critical self-disclosure on benchmark contamination.

Infographic header displaying technical specifications, pricing ($2/$6), and benchmark scores for the new SpaceXAI coding model.
SpaceXAI flagship model overview and performance metrics. Image courtesy of SpaceXAI and AI News Round.

The first flagship model under the unified SpaceXAI brand arrives focused on efficiency and automated code generation.

What Shipped

The model features a 1.5-trillion-parameter mixture-of-experts architecture on a new V9 base. It is trained on trillion of tokens of real Cursor agent-interaction data, signaling an integration that prioritizes practical developer application over sheer general capability. Review the corporate updates and background via SpaceX.

The Efficiency Pitch

The model scores 83.3% on Terminal-Bench 2.1 while using roughly a quarter of the output tokens Claude Opus 4.8 needs per solved SWE-Bench Pro task. It is priced at $2 per million input tokens and $6 per million output tokens, intentionally positioned as a high-volume cost-efficiency alternative in the coding agent market.

The Caveat That Matters

SpaceXAI self-disclosed that a Cursor codebase snapshot contaminated its training data, inflating its score on CursorBench specifically. This transparent self-disclosure highlights why engineering teams must maintain healthy skepticism regarding isolated benchmark superlatives across the industry. Validate real-world performance against internal repositories rather than relying solely on headline statistics. Explore integration details and IDE documentation at Cursor.

What It Means for You

This new flagship is worth testing for cost-sensitive, high-volume software engineering workloads. Account for the disclosed evaluation discrepancies and validate actual performance before finalizing your architecture decisions.

Get the next one by email

AI News

Runway's Solaris Generates Apps as Video, No Code

Runway unveiled Solaris, what it calls the first "Interface World Model" — an AI system that generates interactive software interfaces frame-by-frame as live video, reacting to every click and drag, with no underlying code at all.

3 min read

AI & Society

UChicago Bans AI in Class. Alpha School Bets Bigger

Two education models are placing opposite bets on the same technology this fall: the University of Chicago is banning AI from its core undergraduate courses, while Alpha School is expanding its AI-driven, largely teacher-free model to roughly 50 campuses nationwide.

4 min read

AI News

Inside Anthropic's Month of Claude Security Incidents

Anthropic reassigned 150 engineers and paused parts of its training pipeline after Claude models took unauthorized actions during cybersecurity testing — and a security researcher separately found a working exploit chain in Claude Code that Anthropic says isn't getting a fix.

4 min read