Posts
SortierungNeueste zuerst
Zufälliger PostElon Musk
@elonmusk · 10. Juli 2026
🎯
Dogan Ural@doganuraldesign· 10. Juli 2026Grok

835
462
3610
705.967
Elon Musk
@elonmusk · 10. Juli 2026
The most important thing about Grok Build and the 4.5 release is that it is genuinely so useful for real-world work
X Freeze@XFreeze· 10. Juli 2026Grok 4.5 just topped Perplexity’s WANDR orchestrator evaluation
It scored higher than every other tested configuration at just $4.76 per trial....roughly half the cost of Opus 4.8
Grok 4.5 is becoming the powerful brain coordinating entire agentic workflows
It’s now available

398
233
1431
552.104
Elon Musk hat regepostet

Andrew Milich
@milichab · 10. Juli 2026
Built with two prompts and /goal!
Jayden Davis@JaydenDavisNC· 10. Juli 2026Grok 4.5 in Grok Build is insane for 3D game dev🤯
I ran a detailed prompt using the /goal feature and I created this in only two prompts. Took roughly an hour.
The model rigged all the enemy humanoids and made a fleshed out map and AI logic.
49
95
533
190.054
Elon Musk
@elonmusk · 10. Juli 2026
Anyone can visit the Starship factory and launch site in Texas, as it is right next to on the public highway. It is incredibly inspiring to see!
Sawyer Merritt@SawyerMerritt· 10. Juli 2026SpaceX has released the next episode of its new docuseries about Starship.
781
670
4309
642.912
Elon Musk
@elonmusk · 10. Juli 2026
Background on the Starship story
663
633
3360
713.036
Elon Musk hat regepostet

Tesla Megapack
@Tesla_Megapack · 10. Juli 2026
10 GWh of our industrial energy storage products are now operating across Australia!
That's equivalent to 160,000 Model Ys. And we're just getting started
89
303
2207
175.157
Elon Musk hat regepostet

Tesla Manufacturing
@gigafactories · 10. Juli 2026
End of an era: Decommissioning the original Model S & X assembly line in just 46 days
565
1174
11.769
1,1 Mio.
Elon Musk
@elonmusk · 10. Juli 2026
Try Grok 4.5 in Perplexity
Perplexity@perplexity_ai· 10. Juli 2026Grok 4.5 is now available as an orchestrator model in Computer for Consumer Pro and Max subscribers.
We evaluated it against five other orchestrator configurations on WANDR. It scored higher than every other configuration at roughly half the cost of Opus 4.8.

670
447
2360
1,1 Mio.
Elon Musk
@elonmusk · 10. Juli 2026
Grok is closing the loop on real-world use cases
Thibault Jaigu@ThibaultJaigu· 10. Juli 2026@OpenAI released 3 new models yesterday and we immediately tested it on our internal benchmark.
All three outperforming gpt-5.5 but @SpaceXAI still the clear winner with Grok-4.5

541
299
1676
774.750
Elon Musk
@elonmusk · 10. Juli 2026
True
Mia@MiaAI_lab· 10. Juli 2026Here's one way to get the most from Grok 4.5 ✨
Switching effort to "low" gives strong results on most tasks, near-zero quality loss, big usage savings — plus it's the fastest.
559
321
1970
812.651
Elon Musk
@elonmusk · 10. Juli 2026
Grok Build
X Freeze@XFreeze· 10. Juli 2026Grok 4.5 with Grok Build just ranked #1 on the SWE-Atlas-QnA benchmark with a score of 84
That puts it level with GPT-5.6 (max) Codex and ahead of Claude Code Fable 5 (max), Opus 4.8 (max), and every other tested coding setup
Grok Build is now the most powerful harness for

426
232
1259
729.394
Elon Musk
@elonmusk · 10. Juli 2026
Try @grok
Grok@grok· 10. Juli 2026Grok 4.5 is now available to try on the free tier. Use Grok Build with any X or SuperGrok account.
We’re excited to hear your feedback.

476
256
1283
613.951
Elon Musk
@elonmusk · 10. Juli 2026
Grok 4.5 has the best real-world ROI
Julian Solemsli Rian@LORD_RIAN_· 10. Juli 2026Grok 4.5 just did what no other lab has managed: pushed the limits of frontier intelligence AND made it accessible to everyone.
Hopefully this shifts the AI race toward winning on cost too -> frontier intelligence for all, not just the few.
Huge respect to the @SpaceXAI team.
551
339
2298
905.684
Elon Musk
@elonmusk · 10. Juli 2026
😂
Doge Tipping@Dogetothemoon· 9. Juli 2026"Grok will be able to call Imagine as a tool in agentic mode now.
To develop games, right?"
573
236
1618
590.406
Elon Musk hat regepostet

ben hylak
@benhylak · 9. Juli 2026
i tried tesla full self driving over the weekend.
it’s magical. can’t believe it got that good without it being a bigger deal.
173
178
2147
171.447
Elon Musk hat regepostet

Chief Nerd
@TheChiefNerd · 10. Juli 2026
GAVIN BAKER: “Tesla and Elon have done more to decarbonize the world than all environmental activists combined.”
152
166
896
103.382
Elon Musk hat regepostet

Artificial Analysis
@ArtificialAnlys · 9. Juli 2026
SpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4.8 (48%) at roughly a quarter of their cost per task - the first model to complete more than half of workflow objectives without breaking any business rules
AutomationBench-AA, our independent leaderboard for @zapier’s AutomationBench, tests whether AI agents can automate real SaaS workflows while adhering to business rules. The test set is private to prevent contamination.
Models complete 657 tasks across 40 simulated app environments including Gmail, Google Sheets, Slack, Salesforce, and HubSpot, and the headline score is the share of objectives completed without violating any guardrails.
Key takeaways:
➤ Grok 4.5 completes more objectives than any other model: It completes 79.9% of task objectives and strictly passes 21.9% of tasks. This is the highest we’ve measured on both outcomes, exceeding Claude Fable 5’s 73.3% objective completion and Claude Opus 4.8’s 19.3% of fully-completed tasks
➤ Grok 4.5 pushes out the Pareto frontier of score vs. cost per task: At $0.34 per task, it is both cheaper and higher-scoring than every other leading model - Claude Fable 5 ($1.35 per task), Claude Opus 4.8 ($1.46), GPT-5.5 (xhigh, $1.28), and Gemini 3.5 Flash (high, $0.49)
➤ It is extremely token-efficient: Grok 4.5 uses ~8k output tokens per task, the fewest of any leading model - less than a quarter of Claude Opus 4.8 (32k) and a third of Gemini 3.5 Flash (24k). Its total token usage of 0.44M per task is among the lowest on the leaderboard. Low cost is driven by this efficiency as well as low token pricing
➤ Grok 4.5 uses fewer turns with many parallel tool use: Grok 4.5 resolves tasks in ~16 turns, fewer than GPT-5.5 (xhigh, 25) and less than half of Gemini 3.5 Flash (high, 35), while making the most tool calls per task of any leading model (52.5). It batches 3.3 tool calls per turn, compared to ~2.5 for Claude Opus 4.8 and ~2.0 for GPT-5.5 (xhigh)
➤ Guardrails still get broken: Grok 4.5 triggers 0.63 violations per task, above Claude Opus 4.8 (0.55) and Gemini 3.5 Flash (0.46). At 13.0 objectives completed per violation, it trails Gemini 3.5 Flash (15.0) and Claude Opus 4.8 (13.5)
➤ Its strongest lead is in the hardest domain: Grok 4.5 completes 71% of Finance objectives, the domain with the lowest average score, ahead of Claude Fable 5 (64%) and Claude Opus 4.8 (62%)
Congratulations to @SpaceXAI and @elonmusk on topping the leaderboard!

134
222
1836
863.476
Elon Musk
@elonmusk · 10. Juli 2026
Model comparison
Cursor@cursor_ai· 9. Juli 2026See how every model compares:
448
175
1000
689.304
Elon Musk
@elonmusk · 10. Juli 2026
629
349
1720
850.004
Elon Musk
@elonmusk · 10. Juli 2026
Grok Imagine
868
298
1951
756.877