Posts
Elon Musk
@elonmusk · Jul 10, 2026
The most important thing about Grok Build and the 4.5 release is that it is genuinely so useful for real-world work
X Freeze@XFreeze· Jul 10, 2026Grok 4.5 just topped Perplexity’s WANDR orchestrator evaluation
It scored higher than every other tested configuration at just $4.76 per trial....roughly half the cost of Opus 4.8
Grok 4.5 is becoming the powerful brain coordinating entire agentic workflows
It’s now available

398
233
1.4K
552.1K
Elon Musk reposted

Andrew Milich
@milichab · Jul 10, 2026
Built with two prompts and /goal!
Jayden Davis@JaydenDavisNC· Jul 10, 2026Grok 4.5 in Grok Build is insane for 3D game dev🤯
I ran a detailed prompt using the /goal feature and I created this in only two prompts. Took roughly an hour.
The model rigged all the enemy humanoids and made a fleshed out map and AI logic.
49
95
533
190.1K
Elon Musk
@elonmusk · Jul 10, 2026
Anyone can visit the Starship factory and launch site in Texas, as it is right next to on the public highway. It is incredibly inspiring to see!
Sawyer Merritt@SawyerMerritt· Jul 10, 2026SpaceX has released the next episode of its new docuseries about Starship.
781
670
4.3K
642.9K
Elon Musk
@elonmusk · Jul 10, 2026
Background on the Starship story
663
633
3.4K
713K
Elon Musk reposted

Tesla Megapack
@Tesla_Megapack · Jul 10, 2026
10 GWh of our industrial energy storage products are now operating across Australia!
That's equivalent to 160,000 Model Ys. And we're just getting started
89
303
2.2K
175.2K
Elon Musk reposted

Tesla Manufacturing
@gigafactories · Jul 10, 2026
End of an era: Decommissioning the original Model S & X assembly line in just 46 days
565
1.2K
11.8K
1.1M
Elon Musk
@elonmusk · Jul 10, 2026
Try Grok 4.5 in Perplexity
Perplexity@perplexity_ai· Jul 10, 2026Grok 4.5 is now available as an orchestrator model in Computer for Consumer Pro and Max subscribers.
We evaluated it against five other orchestrator configurations on WANDR. It scored higher than every other configuration at roughly half the cost of Opus 4.8.

670
447
2.4K
1.1M
Elon Musk
@elonmusk · Jul 10, 2026
Grok is closing the loop on real-world use cases
Thibault Jaigu@ThibaultJaigu· Jul 10, 2026@OpenAI released 3 new models yesterday and we immediately tested it on our internal benchmark.
All three outperforming gpt-5.5 but @SpaceXAI still the clear winner with Grok-4.5

541
299
1.7K
774.8K
Elon Musk
@elonmusk · Jul 10, 2026
True
Mia@MiaAI_lab· Jul 10, 2026Here's one way to get the most from Grok 4.5 ✨
Switching effort to "low" gives strong results on most tasks, near-zero quality loss, big usage savings — plus it's the fastest.
559
321
2K
812.7K
Elon Musk
@elonmusk · Jul 10, 2026
Grok Build
X Freeze@XFreeze· Jul 10, 2026Grok 4.5 with Grok Build just ranked #1 on the SWE-Atlas-QnA benchmark with a score of 84
That puts it level with GPT-5.6 (max) Codex and ahead of Claude Code Fable 5 (max), Opus 4.8 (max), and every other tested coding setup
Grok Build is now the most powerful harness for

426
232
1.3K
729.4K
Elon Musk
@elonmusk · Jul 10, 2026
Try @grok
Grok@grok· Jul 10, 2026Grok 4.5 is now available to try on the free tier. Use Grok Build with any X or SuperGrok account.
We’re excited to hear your feedback.

476
256
1.3K
614K
Elon Musk
@elonmusk · Jul 10, 2026
Grok 4.5 has the best real-world ROI
Julian Solemsli Rian@LORD_RIAN_· Jul 10, 2026Grok 4.5 just did what no other lab has managed: pushed the limits of frontier intelligence AND made it accessible to everyone.
Hopefully this shifts the AI race toward winning on cost too -> frontier intelligence for all, not just the few.
Huge respect to the @SpaceXAI team.
551
339
2.3K
905.7K
Elon Musk
@elonmusk · Jul 10, 2026
😂
Doge Tipping@Dogetothemoon· Jul 9, 2026"Grok will be able to call Imagine as a tool in agentic mode now.
To develop games, right?"
573
236
1.6K
590.4K
Elon Musk reposted

ben hylak
@benhylak · Jul 9, 2026
i tried tesla full self driving over the weekend.
it’s magical. can’t believe it got that good without it being a bigger deal.
173
178
2.1K
171.4K
Elon Musk reposted

Chief Nerd
@TheChiefNerd · Jul 10, 2026
GAVIN BAKER: “Tesla and Elon have done more to decarbonize the world than all environmental activists combined.”
152
166
896
103.4K
Elon Musk reposted

Artificial Analysis
@ArtificialAnlys · Jul 9, 2026
SpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4.8 (48%) at roughly a quarter of their cost per task - the first model to complete more than half of workflow objectives without breaking any business rules
AutomationBench-AA, our independent leaderboard for @zapier’s AutomationBench, tests whether AI agents can automate real SaaS workflows while adhering to business rules. The test set is private to prevent contamination.
Models complete 657 tasks across 40 simulated app environments including Gmail, Google Sheets, Slack, Salesforce, and HubSpot, and the headline score is the share of objectives completed without violating any guardrails.
Key takeaways:
➤ Grok 4.5 completes more objectives than any other model: It completes 79.9% of task objectives and strictly passes 21.9% of tasks. This is the highest we’ve measured on both outcomes, exceeding Claude Fable 5’s 73.3% objective completion and Claude Opus 4.8’s 19.3% of fully-completed tasks
➤ Grok 4.5 pushes out the Pareto frontier of score vs. cost per task: At $0.34 per task, it is both cheaper and higher-scoring than every other leading model - Claude Fable 5 ($1.35 per task), Claude Opus 4.8 ($1.46), GPT-5.5 (xhigh, $1.28), and Gemini 3.5 Flash (high, $0.49)
➤ It is extremely token-efficient: Grok 4.5 uses ~8k output tokens per task, the fewest of any leading model - less than a quarter of Claude Opus 4.8 (32k) and a third of Gemini 3.5 Flash (24k). Its total token usage of 0.44M per task is among the lowest on the leaderboard. Low cost is driven by this efficiency as well as low token pricing
➤ Grok 4.5 uses fewer turns with many parallel tool use: Grok 4.5 resolves tasks in ~16 turns, fewer than GPT-5.5 (xhigh, 25) and less than half of Gemini 3.5 Flash (high, 35), while making the most tool calls per task of any leading model (52.5). It batches 3.3 tool calls per turn, compared to ~2.5 for Claude Opus 4.8 and ~2.0 for GPT-5.5 (xhigh)
➤ Guardrails still get broken: Grok 4.5 triggers 0.63 violations per task, above Claude Opus 4.8 (0.55) and Gemini 3.5 Flash (0.46). At 13.0 objectives completed per violation, it trails Gemini 3.5 Flash (15.0) and Claude Opus 4.8 (13.5)
➤ Its strongest lead is in the hardest domain: Grok 4.5 completes 71% of Finance objectives, the domain with the lowest average score, ahead of Claude Fable 5 (64%) and Claude Opus 4.8 (62%)
Congratulations to @SpaceXAI and @elonmusk on topping the leaderboard!

134
222
1.8K
863.5K
Elon Musk
@elonmusk · Jul 10, 2026
Model comparison
Cursor@cursor_ai· Jul 9, 2026See how every model compares:
448
175
1K
689.3K
Elon Musk
@elonmusk · Jul 10, 2026
629
349
1.7K
850K
Elon Musk
@elonmusk · Jul 10, 2026
Grok Imagine
868
298
2K
756.9K
Elon Musk reposted

SimWorld
@simworld_ai · Jul 9, 2026
Built from scratch by Grok 4.5 + Grok Build in UE5.8: a cyberpunk L-corner street with neon facades, rain, signs, and crowds walking through the scene.
End-to-end, the run took 10.75M tokens, 36.5 minutes, and only ~$12.4 at API pricing.
What impressed me most is not just the final render, but the process: the coding agent built it in many small steps, saved 30+ map checkpoints, verified the scene, and iteratively fixed issues autonomously.
Congrats to the @SpaceXAI team for building such an impressive coding model and harness. This feels like a real step toward agentic 3D world creation. @elonmusk @milichab @skcd42 @yunta_tsai
#UnrealEngine
83
126
766
137.1K