SpaceXAI / Grok
Continuing coverage of Grok and SpaceXAI — model releases and benchmarks, the xAI-into-SpaceX consolidation, enterprise distribution, and the builder ecosystem around Grok.
Grok is the model family; the company behind it has been a moving target. What started as xAI — Elon Musk's answer to OpenAI — was folded into SpaceX in 2026, and RuntimeWire's coverage spans both eras: same models, same mission, an increasingly consolidated corporate wrapper.
We cover Grok as a frontier contender because that's what the results say it is: Grok 4.5 and 4.6 have led benchmark boards — including RuntimeWire's own Newsroom Reliability and BBEH Mini runs — shipped a 500,000-token context window, and landed distribution inside GitHub Copilot, Cursor, Devin, Amazon Bedrock, and Databricks. Around the models, SpaceXAI keeps shipping product at startup pace: Grok Build for creating apps from chat, Grok Bot with Cursor, Imagine for image generation, and Grok Voice.
What we're tracking:
- Grok model releases and how they benchmark against OpenAI, Anthropic, and Google
- The consolidation: what folding xAI into SpaceX means for talent, capital, and focus
- Distribution: Grok's push into coding tools, enterprise platforms, and X itself
- The builder ecosystem: Grokathons, Grok Build apps, and what ships on top
Primary sources and benchmark methodology are linked in every story.
All updates
- Grok 4.6 tops Newsroom Reliability v0.2 benchmark at 0.78 — In a 50-task run of Newsroom Reliability v0.2, SpaceXAI’s Grok 4.6 ranked first with a score of 0.78.
- GitHub adds xAI's Grok 4.6 to Copilot for long-running coding agents — xAI's Grok 4.6 is entering GitHub's IDE, command-line and enterprise workflows two days after launch, under usage-based billing.
- Grok 4.6 tops BBEH Mini benchmark, scoring 0.67 on 460 tasks — In a 460-task BBEH Mini evaluation, SpaceXAI’s Grok 4.6 led the field with a 0.67 score at $0.0034 per task.
- Grok 4.6 tops Newsroom Reliability v0.2 benchmark at 0.79 — In a 50-task run of Newsroom Reliability v0.2, SpaceXAI’s Grok 4.6 ranked first with a score of 0.79 at an estimated $0.0056 per task.
- SpaceXAI's Grokathon crowns a binary decompiler built in 12 hours — Nova beat social simulation tool Signal and brain-sensing speech prototype ThinkVoice in the 2026 competition.
- Scoop: Grok Bot ships dormant shared rooms for tasking other users' agents — Windows code defines owner-approved invites, cross-account agent membership and scoped computer access, though the tested account remained gated.
- Cognition adds Grok 4.6 to Devin's multi-model coding platform — The integration gives xAI distribution inside an established engineering workflow and strengthens Cognition's pitch that orchestration, not any single foundation model, is the…
- SpaceXAI ships Grok 4.6 with a 500,000-token context window — The model keeps Grok 4.5's $2/$6 base API price, adds a 500,000-token context window and launches in Cursor and Grok Build.
- Scoop: SpaceXAI's Grok Bot stamps an account-derived ID into browser requests — The 16-character value is derived from the authenticated Cursor account and reaches external websites through Chrome's user-agent header.
- Grok Bot's hidden "Elon-Only Settings" picker lists 33 models — The dormant desktop control pulls an account-scoped catalog spanning Grok, GPT, Claude, Gemini, Kimi and GLM.
- Four Grok 4.5 checkpoints appear in Arena near xAI's 4.6 window — The roster entries point to multimodal and web-enabled variants, but their labels do not establish that Grok 4.6 is behind them.
- Cursor and SpaceXAI launch Grok Bot for work beyond coding — The beta gives paid users cloud agents that sign into websites, learn routines and keep running after a laptop closes.
- Grok 4.6 launch appears imminent after Cursor briefly exposes the model — An outside tester received preview access hours after Cursor users found the unreleased model in its production picker.
- xAI launches Imagine Image 2.0 for production image workflows — The new Quality Mode adds region edits, five-image references and smart resizing, while API access remains pending.
- EXCLUSIVE: xAI Has Shipped the Foundation for an Unannounced Grok Remote-Workspace Product — Hidden inside Grok Build 1.0, the Computer Hub command can turn a developer's local workspace into a remotely accessible tool server.
- Grok Build reaches 1.0 as SpaceXAI hardens its terminal coding agent — The release focuses on reliability and controls, while Grok 4.5, workflows and an open-source harness carry the bigger bet.
- SpaceXAI ships Grok Voice Think Fast 2.0 at a 60% price premium — The speech-to-speech model costs $0.08 per minute and becomes the default Grok voice API endpoint on August 5th.
- SpaceXAI launches Grok Build Mode for building and publishing apps from chat — The early beta turns prompts into live websites, games and dashboards, extending Grok's coding push beyond the terminal.
- SpaceXAI opens Grokathon to seed apps around Grok 4.5 and X — The August 8th event offers selected builders access to Grok models and X APIs; applications close July 28th and prizes remain undisclosed.
- Pedro Shakour ships an iPhone remote for xAI's Grok Build — The open-source SwiftUI client keeps xAI's coding agent on a Mac, using ACP to turn an iPhone into a sideloaded control panel.
- Exclusive: Grok Build hides a Doom-like 'easter egg' game behind the /gboom command — RuntimeWire testing found an undocumented mini-game inside xAI's terminal coding agent
- Grok 4.5 gives Musk an Opus-class price weapon — SpaceXAI launched the coding-focused model on July 8th with $2 input and $6 output pricing, while Musk cast speed as the real edge.
- SpaceXAI's Grok 4.5 ranks fourth on Artificial Analysis benchmark — The benchmarker put Grok 4.5 behind Fable 5, GPT-5.5 and Opus 4.8, with cost as the clearest part of SpaceXAI's pitch.
- SpaceXAI releases Grok 4.5 with Cursor-trained coding push — Michael Truell says the model has replaced Composer 2.5 for many on the Cursor team as SpaceXAI prices it for broad developer use.
- SpaceXAI makes Musk's AI consolidation even more confusing — The @SpaceXAI account switch gives Musk's AI operation a clearer brand, but not a clearer corporate structure, product map or accountability trail.
- Elon Musk says Grok 4.5 is in private beta at SpaceX and Tesla — The new xAI model is said to use a 1.5T V9 foundation model and supplemental Cursor training data, but no public benchmark table is out.
- Yann LeCun calls xAI a failure and warns AI labs are running on investor subsidy — The AMI Labs founder is attacking Musk's talent losses and the economics behind frontier AI while selling a world-model alternative.
- Elon Musk takes Grok into Databricks as xAI chases enterprise distribution — Grok is now a native option in Agent Bricks, giving Databricks customers another model choice for governed AI agents.
- Elon Musk puts xAI's video bet on a 2026 movie clock — xAI posted Grok Imagine Video 1.5 this week, but Musk's full movie prediction still runs ahead of what the public docs describe.
- xAI puts Grok 4.3 inside Amazon Bedrock as Musk pushes Grok beyond X — AWS developers get Grok through Bedrock, but the real-time X data pitch remains clearer in Musk's posts than in Amazon's docs.