Elon Musk Archive

Posts

Yun-Ta Tsai
Yun-Ta Tsai
@yunta_tsai · 22 sept. 2026
I have some sympathy for people who work on LLMs. A lot of the job has become the pursuit of bar charts. For years, Full Self-Driving only had to be good where it actually matters: the world. Ourselves, friends, families, customers, and unscripted miles. Real weather, real
Elon Musk
Elon Musk
@elonmusk · 22 sept. 2026
Exactly
138
60
761
42,2 k
Elon Musk a reposté
NIK
NIK
@ns123abc · 22 sept. 2026
Grok TRIPLED on Terminal-Bench 4.0 in only two months (12.4% -> 38.0%), overtaking GPT-5.6 Sol
NIK
75
90
662
167,1 k
Elon Musk a reposté
Fascinating
Fascinating
@fasc1nate · 22 sept. 2026
Remembering Roger Boisjoly, the engineer who correctly identified a fatal flaw in the Challenger shuttle design months before the disaster, but nobody gave a damn. His exact words to his wife, Darlene: "It's going to blow up" 73 seconds before it did. More rare historical photos:
Fascinating
87
166
1,1 k
170,6 k
Elon Musk a reposté
Aidan
Aidan
@A_d_n_R_d_i_g · 22 sept. 2026
Love Trump or hate him, it’s genuinely embarrassing the White House didn’t have this before now. Decline is a choice.
Aidan
337
843
13,6 k
462 k
Elon Musk a reposté
StarbaseTX
StarbaseTX
@StarbaseTX · 21 sept. 2026
Saturday was a good day for Boca Chica Beach. 985 volunteers came out for International Coastal Cleanup Day and removed 2,880 lbs of trash from our shoreline. Thank you to @SpaceX, @TXAdoptABeach, and Cameron County, and to every friend, family member, and neighbor who came out to volunteer. This is what community looks like. See you at the next one!
StarbaseTXStarbaseTXStarbaseTXStarbaseTX
128
242
2,6 k
189,9 k
Elon Musk a reposté
British Intel
British Intel
@TheBritishIntel · 21 sept. 2026
I stand with @JohnCleese.
British Intel
211
1,4 k
14,5 k
209,6 k
Elon Musk a reposté
samuel
samuel
@marsrepublica · 21 sept. 2026
In just the last 90 days: 1. Grok 4.3 — barely top 10. “xAI is dead beyond compute leases.” 2. Grok 4.5 — massive comeback. “Maybe a chance. But Grok will never catch up to the frontier.” 3. Grok 4.6 — they’re frontier. “Still not top 3.” 4. Grok 4.7 — now top 3 in frontier coding, while being much cheaper & faster. 5. Grok 4.8 next month..... The model machine is just starting up. Grok will be the workhorse of the upcoming agentic era.
samuel
172
230
2,3 k
1,7 M
Elon Musk a reposté
Kun Chen
Kun Chen
@kunchenguid · 22 sept. 2026
day 1 observations for grok 4.7 ignore the reports that say “it’s terrible” and the only thing they reference is a public benchmark. the same benchmarks told us opus 5 was better that fable - they are useless also ignore the reports that compare models with 3d games - that’s not real work. it's made for attention on social media i used grok 4.7 for a whole day as my firstmate, and it has been a really solid model with visible improvements over 4.5 (i'm ignoring 4.6 because 4.5 has been working better in my experience) key differences with 4.7 - 1. it follows system prompt very, very closely i noticed firstmate showing many new behaviors that i've never seen before, such as asking me to name specific red CI checks that i'm ok with bypassing, and refuse a simple "yolo" instruction i traced it and it's indeed how i instructed it in firstmate's system prompt, but none of the other models followed it closely enough to make this behavior visible - grok 4.7 is the first to pick that up there were a few other similar examples as well. so to me this is a clear behavioral difference 2. it's very "stable" if you've used astra then you know what a "spiky" model is. it can have some genius moments but you occasionally also wonder "how could it be so dumb and doesn't get me". grok 4.7 is the opposite of that throughout the whole day so far, i'll be honest i haven't get a "wow this is absolutely genius" moment yet, but grok 4.7 has been very steady with no big surprises. its behavior feels predictable, which does help it gain trust from me quickly 3. it's a conservative model it doesn't like to take actions without asking, and would explicitly say so this is a bit of a double edged sword, because it means i sometimes have to state the obvious "yes i do want that", but in hindsight a lot of those cases are indeed a bit ambiguous and i may not have preferred the model to just move forward without my confirmation 4. it's a bit slower and costs more than 4.5, visibly turns are taking a bit longer and my quota is draining at a visibly faster pace. i haven't quantified exactly where this is coming from yet so overall, i think it's showing some clearly different traits, and i mostly like the changes. i'm going to keep it as my primary firstmate and observe more if you've been using it, what qualitative insights have you gathered from real usage so far?
49
36
340
30,5 k
Elon Musk a reposté
Alex Finn
Alex Finn
@AlexFinn · 22 sept. 2026
Grok 4.7 just released and it's an EXCELLENT model It was trained FOR Grok Bot Meaning this is a fully agentic model trained to do your knowledge work better than you can In this video I show you how to use Grok 4.7 and a Grok Bot workflow that will 10x your productivity:
Alex Finn
93
123
1,1 k
230,7 k
Elon Musk a reposté
tetsuo
tetsuo
@tetsuoai · 22 sept. 2026
@elonmusk I'm using it right now to build a game engine in C. It's really good.
37
40
566
182,9 k
Elon Musk
Elon Musk
@elonmusk · 22 sept. 2026
True
samuelsamuel@marsrepublica· 21 sept. 2026
In just the last 90 days: 1. Grok 4.3 — barely top 10. “xAI is dead beyond compute leases.” 2. Grok 4.5 — massive comeback. “Maybe a chance. But Grok will never catch up to the frontier.” 3. Grok 4.6 — they’re frontier. “Still not top 3.” 4. Grok 4.7 — now top 3 in frontier
samuel
459
329
2,7 k
745,3 k
Elon Musk a reposté
Zack Jackson
Zack Jackson
@ScriptedAlchemy · 21 sept. 2026
Been testing Grok 4.7 for the past week or more. It’s been a great improvement over 4.6. It worked for over 70 hours straight on a goal. Much better attention to detail. The 500k context window really makes a difference imo. 4.7 much better at things like skill selection; workflows. Found it particularly good with pstack. Didn’t have access to it in Grokbot, but I really wanted to test it there. Together it’s a great daily driver.
62
68
633
180 k
Elon Musk
Elon Musk
@elonmusk · 22 sept. 2026
Grok 4.7
Elon Musk
456
224
2 k
389,6 k
Sawyer Merritt
Sawyer Merritt
@SawyerMerritt · 22 sept. 2026
Tesla has for years openly invited other automakers to license FSD. None of them have accepted.
Blake Scholl 🛫Blake Scholl 🛫@bscholl· 21 sept. 2026
Prediction: Tesla will keep FSD an in-house advantage for a couple years, and then once others start getting close will begin licensing it for other cars—just as they opened up the Supercharger network.
Elon Musk
Elon Musk
@elonmusk · 22 sept. 2026
Exactly
0
0
1
18
Elon Musk
Elon Musk
@elonmusk · 22 sept. 2026
Grok 4.7
Artificial AnalysisArtificial Analysis@ArtificialAnlys· 21 sept. 2026
Grok 4.7 is behind only Anthropic models on AA-Briefcase, ranking just behind Opus 5 at ~50% of its Cost per Task Grok 4.7’s improvements over Grok 4.6 are clear in AA-Briefcase-Lite, our public due diligence scenario where models are tasked with building market models and
Artificial Analysis
227
106
892
367,8 k
Elon Musk
Elon Musk
@elonmusk · 22 sept. 2026
Cool
AuggieAuggie@aug5thmusic· 21 sept. 2026
This was unexpected. Grok 4.7 scored 100% on my music error detection test, the same as GPT-6 Astra.
Auggie
179
102
723
356,6 k
Elon Musk
Elon Musk
@elonmusk · 21 sept. 2026
Grok 4.7 is a strong combination of intelligence, speed & low cost
SpaceXAISpaceXAI@SpaceXAI· 21 sept. 2026
Grok 4.7 works longer on difficult tasks, checks its work more carefully, and comes with our strongest safeguards to date.
SpaceXAI
2 k
2,3 k
24,5 k
28,2 M
Elon Musk a reposté
X Freeze
X Freeze
@XFreeze · 21 sept. 2026
Grok 4.7 xHigh is now at the top with just ONE point away from Claude Fable 5.1 Max on Artificial Analysis’ AA-Briefcase benchmark Claude Fable 5.1 Max — 59% Grok 4.7 xHigh — 58% Just a 1-point difference at the very top And Grok is ahead of GPT-6 Astra, GPT-5.6, Gemini, Kimi, GLM and nearly every other frontier model on the benchmark
X Freeze
108
155
1,1 k
0
Tak 🦞
Tak 🦞
@cherry_mx_reds · 21 sept. 2026
Grok 4.7 made this in blender. I didn't tell it what to make but only that it should be something it can accomplish in 10 minutes or so. Honestly not bad at all. Not bad at all. Did they train Grok on blender?
Tak 🦞
Elon Musk
Elon Musk
@elonmusk · 21 sept. 2026
Cool
112
47
699
41,8 k
Elon Musk a reposté
Matt Shumer
Matt Shumer
@mattshumer_ · 21 sept. 2026
I’ve been testing Grok 4.7… it’s a really great daily driver, and a huge step up over 4.6. Definitely worth trying!
SpaceXAISpaceXAI@SpaceXAI· 21 sept. 2026
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
SpaceXAI
68
72
630
234,8 k
Posts — Elon Musk Archive