Posts
SortierungNeueste zuerst
Zufälliger Post
Yun-Ta Tsai
@yunta_tsai · 22. Sept. 2026
I have some sympathy for people who work on LLMs. A lot of the job has become the pursuit of bar charts.
For years, Full Self-Driving only had to be good where it actually matters: the world. Ourselves, friends, families, customers, and unscripted miles. Real weather, real
Elon Musk
@elonmusk · 22. Sept. 2026
Exactly
138
60
761
42.179
Elon Musk hat regepostet

NIK
@ns123abc · 22. Sept. 2026
Grok TRIPLED on Terminal-Bench 4.0 in only two months (12.4% -> 38.0%), overtaking GPT-5.6 Sol

75
90
662
167.149
Elon Musk hat regepostet

Fascinating
@fasc1nate · 22. Sept. 2026
Remembering Roger Boisjoly, the engineer who correctly identified a fatal flaw in the Challenger shuttle design months before the disaster, but nobody gave a damn. His exact words to his wife, Darlene: "It's going to blow up" 73 seconds before it did.
More rare historical photos:

87
166
1134
170.596
Elon Musk hat regepostet

Aidan
@A_d_n_R_d_i_g · 22. Sept. 2026
Love Trump or hate him, it’s genuinely embarrassing the White House didn’t have this before now.
Decline is a choice.

337
843
13.599
461.997
Elon Musk hat regepostet

StarbaseTX
@StarbaseTX · 21. Sept. 2026
Saturday was a good day for Boca Chica Beach.
985 volunteers came out for International Coastal Cleanup Day and removed 2,880 lbs of trash from our shoreline.
Thank you to @SpaceX, @TXAdoptABeach, and Cameron County, and to every friend, family member, and neighbor who came out to volunteer.
This is what community looks like. See you at the next one!




128
242
2634
189.887
Elon Musk hat regepostet

British Intel
@TheBritishIntel · 21. Sept. 2026
I stand with @JohnCleese.

211
1441
14.527
209.617
Elon Musk hat regepostet

samuel
@marsrepublica · 21. Sept. 2026
In just the last 90 days:
1. Grok 4.3 — barely top 10. “xAI is dead beyond compute leases.”
2. Grok 4.5 — massive comeback. “Maybe a chance. But Grok will never catch up to the frontier.”
3. Grok 4.6 — they’re frontier. “Still not top 3.”
4. Grok 4.7 — now top 3 in frontier coding, while being much cheaper & faster.
5. Grok 4.8 next month.....
The model machine is just starting up.
Grok will be the workhorse of the upcoming agentic era.

172
230
2332
1,7 Mio.
Elon Musk hat regepostet

Kun Chen
@kunchenguid · 22. Sept. 2026
day 1 observations for grok 4.7
ignore the reports that say “it’s terrible” and the only thing they reference is a public benchmark. the same benchmarks told us opus 5 was better that fable - they are useless
also ignore the reports that compare models with 3d games - that’s not real work. it's made for attention on social media
i used grok 4.7 for a whole day as my firstmate, and it has been a really solid model with visible improvements over 4.5 (i'm ignoring 4.6 because 4.5 has been working better in my experience)
key differences with 4.7 -
1. it follows system prompt very, very closely
i noticed firstmate showing many new behaviors that i've never seen before, such as asking me to name specific red CI checks that i'm ok with bypassing, and refuse a simple "yolo" instruction
i traced it and it's indeed how i instructed it in firstmate's system prompt, but none of the other models followed it closely enough to make this behavior visible - grok 4.7 is the first to pick that up
there were a few other similar examples as well. so to me this is a clear behavioral difference
2. it's very "stable"
if you've used astra then you know what a "spiky" model is. it can have some genius moments but you occasionally also wonder "how could it be so dumb and doesn't get me". grok 4.7 is the opposite of that
throughout the whole day so far, i'll be honest i haven't get a "wow this is absolutely genius" moment yet, but grok 4.7 has been very steady with no big surprises. its behavior feels predictable, which does help it gain trust from me quickly
3. it's a conservative model
it doesn't like to take actions without asking, and would explicitly say so
this is a bit of a double edged sword, because it means i sometimes have to state the obvious "yes i do want that", but in hindsight a lot of those cases are indeed a bit ambiguous and i may not have preferred the model to just move forward without my confirmation
4. it's a bit slower and costs more than 4.5, visibly
turns are taking a bit longer and my quota is draining at a visibly faster pace. i haven't quantified exactly where this is coming from yet
so overall, i think it's showing some clearly different traits, and i mostly like the changes. i'm going to keep it as my primary firstmate and observe more
if you've been using it, what qualitative insights have you gathered from real usage so far?
49
36
340
30.528
Elon Musk hat regepostet

Alex Finn
@AlexFinn · 22. Sept. 2026
Grok 4.7 just released and it's an EXCELLENT model
It was trained FOR Grok Bot
Meaning this is a fully agentic model trained to do your knowledge work better than you can
In this video I show you how to use Grok 4.7 and a Grok Bot workflow that will 10x your productivity:
93
123
1058
230.731
Elon Musk hat regepostet

tetsuo
@tetsuoai · 22. Sept. 2026
@elonmusk I'm using it right now to build a game engine in C. It's really good.
37
40
566
182.882
Elon Musk
@elonmusk · 22. Sept. 2026
True
samuel@marsrepublica· 21. Sept. 2026In just the last 90 days:
1. Grok 4.3 — barely top 10. “xAI is dead beyond compute leases.”
2. Grok 4.5 — massive comeback. “Maybe a chance. But Grok will never catch up to the frontier.”
3. Grok 4.6 — they’re frontier. “Still not top 3.”
4. Grok 4.7 — now top 3 in frontier

459
329
2661
745.260
Elon Musk hat regepostet

Zack Jackson
@ScriptedAlchemy · 21. Sept. 2026
Been testing Grok 4.7 for the past week or more. It’s been a great improvement over 4.6.
It worked for over 70 hours straight on a goal. Much better attention to detail.
The 500k context window really makes a difference imo. 4.7 much better at things like skill selection; workflows. Found it particularly good with pstack.
Didn’t have access to it in Grokbot, but I really wanted to test it there. Together it’s a great daily driver.
62
68
633
179.961
Elon Musk
@elonmusk · 22. Sept. 2026
Grok 4.7

456
224
1961
389.602

Sawyer Merritt
@SawyerMerritt · 22. Sept. 2026
Tesla has for years openly invited other automakers to license FSD. None of them have accepted.
Blake Scholl 🛫@bscholl· 21. Sept. 2026Prediction: Tesla will keep FSD an in-house advantage for a couple years, and then once others start getting close will begin licensing it for other cars—just as they opened up the Supercharger network.
Elon Musk
@elonmusk · 22. Sept. 2026
Exactly
0
0
1
18
Elon Musk
@elonmusk · 22. Sept. 2026
Grok 4.7
Artificial Analysis@ArtificialAnlys· 21. Sept. 2026Grok 4.7 is behind only Anthropic models on AA-Briefcase, ranking just behind Opus 5 at ~50% of its Cost per Task
Grok 4.7’s improvements over Grok 4.6 are clear in AA-Briefcase-Lite, our public due diligence scenario where models are tasked with building market models and

227
106
892
367.833
Elon Musk
@elonmusk · 22. Sept. 2026
Cool
Auggie@aug5thmusic· 21. Sept. 2026This was unexpected. Grok 4.7 scored 100% on my music error detection test, the same as GPT-6 Astra.

179
102
723
356.559
Elon Musk
@elonmusk · 21. Sept. 2026
Grok 4.7 is a strong combination of intelligence, speed & low cost
SpaceXAI@SpaceXAI· 21. Sept. 2026Grok 4.7 works longer on difficult tasks, checks its work more carefully, and comes with our strongest safeguards to date.

2044
2282
24.475
28,2 Mio.
Elon Musk hat regepostet

X Freeze
@XFreeze · 21. Sept. 2026
Grok 4.7 xHigh is now at the top with just ONE point away from Claude Fable 5.1 Max on Artificial Analysis’ AA-Briefcase benchmark
Claude Fable 5.1 Max — 59%
Grok 4.7 xHigh — 58%
Just a 1-point difference at the very top
And Grok is ahead of GPT-6 Astra, GPT-5.6, Gemini, Kimi, GLM and nearly every other frontier model on the benchmark

108
155
1076
0

Tak 🦞
@cherry_mx_reds · 21. Sept. 2026
Grok 4.7 made this in blender.
I didn't tell it what to make but only that it should be something it can accomplish in 10 minutes or so.
Honestly not bad at all. Not bad at all.
Did they train Grok on blender?
Elon Musk
@elonmusk · 21. Sept. 2026
Cool
112
47
699
41.776
Elon Musk hat regepostet

Matt Shumer
@mattshumer_ · 21. Sept. 2026
I’ve been testing Grok 4.7… it’s a really great daily driver, and a huge step up over 4.6.
Definitely worth trying!
SpaceXAI@SpaceXAI· 21. Sept. 2026Grok 4.7 is here.
It's a notable improvement over Grok 4.6 at the same price and speed.

68
72
630
234.815