Improvements to Grok @Bot
We've shipped several quality-of-life improvements to Grok Bot. Mobile notifications are now grouped by Bot and use their specific icon.
892
1,160
10,583
22,843,348
Wow, this is an all-star lineup!
1,335
1,927
17,916
5,139,919
🚨Starlink is providing free internet service through Sept. 17 to customers in Hawaii impacted by Hurricane Lala. SpaceX is also working with local officials to deploy Starlink for emergency crews and residents in the hardest-hit areas. Thank you @elonmusk and @Starlink! 🇺🇸🇺🇸
183
501
3,243
437,636
🎯
Just imagine economists from the 1600s arguing about the max GDP of the western hemisphere as a % of UK GDP. Now that’s funny and obvious. Same for space but space has many orders of magnitude more energy and mass to turn into productivity. It’s just a matter of when, and things are accelerating
553
958
6,233
4,070,934
Try Grok 4.6! Grok 4.7, which is a major upgrade, is coming soon.
Grok 4.6 is up to 5× cheaper than GPT-5.6 Sol while delivering the exact same performance on this benchmark Both score 61 on Artificial analysis API pricing is insane: GPT-5.6 Sol: • $5 / 1M input tokens • $30 / 1M output tokens Grok 4.6: • $2 / 1M input tokens • $6 / 1M output tokens That makes Grok 2.5× cheaper on input and 5× cheaper on output Same performance here....at a fraction of the price Grok 4.6’s price-to-performance is ridiculous
1,159
1,156
9,785
3,315,109
Elon Musk retweeted
73% of women on the titanic survived. Only 19% of men did. Read that again. Even the richest men alive, class, power, status, all useless when the ship started sinking. Hierarchy collapsed. Gender didn’t. “Women and children first” Not a suggestion. A rule men didn’t break. Men stood there. No panic No begging. Just accepted it. They didn’t ask: “Do I know her?” “Is she worth it?” “Will she remember me?” No one asked if it was fair. No camera. No applause. No second chances. Zero return. Full cost. Understand this: When survival gets limited, a man’s life becomes negotiable. Men like that built the world. And today? Masculinity gets mocked. Call it whatever you want. But when everything collapsed, men paid first. This isn’t opinion. This is history.
2,779
15,791
79,368
5,236,033
Yes
The space economy should probably be thought of as an order of magnitude greater (or two), in large part because it will be absolutely necessary for the continued scaling of AI compute and power It will also likely carry the vast majority of internet traffic (which will also become far more AI than human) It is the final frontier
1,113
1,428
7,882
5,309,816
It will be much bigger
"Second Space Age": Goldman Forecasts $1.8 Trillion Space Economy By 2035 zerohedge.com/technology/sec…
1,413
2,075
17,022
7,104,149
So much has been built since then!
When @MKBHD interviewed Elon 8 years ago, Fremont was Tesla’s only car factory. Since then, $TSLA added 3 more car factories, a Tesla Semi factory, 3 Megapacktories, a lithium refinery, 4680 battery cell production, the world’s best-selling car: Model Y, and a Level 4 Cybercab.
598
969
8,145
4,323,362
Elon Musk retweeted
BREAKING: Starlink has now partnered with 48 airlines worldwide, covering more than 7,000 aircraft that are already equipped, undergoing installation, or under contract.
512
975
6,358
537,806
Grok
Grok 4.6 just ranked #2 in agentic U.S. administrative/regulatory legal research and the cost is insane • Grok 4.6 — 62.12% | $1.52/test • Claude Opus 5 — 65.15% | $6.58/test • GPT-5.6 Sol — 60.61% | $19.69/test Grok actually beats GPT-5.6 Sol on accuracy while costing roughly 13× less per test And it comes within just 3 percentage points of Opus 5, while Opus costs more than 4× as much This is exactly where Grok 4.6 is getting ridiculous It’s delivering frontier-level agentic performance at a fraction of the cost Grok is making high-end AI dramatically cheaper to actually use
890
1,088
5,767
3,054,651
Starship Flight 13 is being recovered from sea
After approx. 24 days at sea, the SpaceX Recovery team successfully guided Starship to a location just off the coast of Christmas Island. A team of SpaceX engineers is on their way to conduct additional analysis on the vehicle in calmer waters before attempting to return it to Starbase
1,929
3,614
34,190
5,272,792
Grok 4.6 takes top spot on this benchmark
🥇 We evaluated Grok 4.6 on MedAgentBench, a benchmark for agentic clinical EHR tasks, and it took the top spot. Grok 4.6 posts the highest pass@1 we've measured to date, moving ahead of the previous leader, GPT-5.6 Sol, to claim first place on the leaderboard. 🏥 Grok 4.6 averages ~95.9% pass@1 (avg of 3 runs), edging out the prior best (GPT-5.6-sol at ~94.7%) and improving on Grok 4.5 by roughly 2.5 points. On a benchmark where the model acts as an autonomous agent in a simulated EHR, calling FHIR APIs to complete clinical tasks across 10 task types, the result is also remarkably consistent run to run (95.3% to 96.3%). 📊 Leaderboards: medicalsphere.ai/benchmarks
1,287
1,359
7,469
3,476,030
Elon Musk retweeted
Grok 4.6 is my daily driver now. This is why: When I’m using interactive agents in a CLI/terminal (or even Slack), what matters most to me is speed & token throughput. Intelligence and differences in performance on benchmarks is negligible to me when speed is sacrificed. If I am there to steer & guide the agent, or make some quick fixes, then I don’t want to wait 5+ min per turn. Opus 5 (fast) is a good solution to this, but it is WAY too expensive. Grok 4.6 is perfect though, incredibly fast, almost as smart as Fable, and relatively much cheaper than anything else right now. On the flip side, for cloud agents working in Factories & background tasks triggered on crons, I don’t mind using intelligent, slow models like Fable / Sol. I’d even opt for a model router for cloud tasks that biases to heavy, smart models. I guess the meta point I’m making here is that model choice is largely a function of whether there’s a human waiting in the loop.
93
158
1,329
540,516
Grok @Bot
I think Grok @Bot is a glimpse into the future of personal AI agents. Here's my new tutorial where I show you how to set up 5 useful bots: 1. An advisor to create and manage your bots 2. A YouTube researcher to find outlier videos 3. An X scout to find viral and funny tweets 4. A digital Marie Kondo to clean up your inbox and save money on paid subscriptions 5. A personal concierge to save money on trips I also tested a Gamer bot to see if Grok Bot can install and play classic games like Doom, Red Alert, and Commander Keen. Plus, I discuss the biggest barrier to Grok Bot adoption and whether it can replace ChatGPT as my daily driver. 📌 Watch now: piped.video/MkVcHbviYOw
838
1,002
6,563
6,394,732
Elon Musk retweeted
Create a full scene with Grok Imagine for a chance to win 👇
Homer had a lyre. You have Grok Imagine. Create a compelling scene from Homer’s The Odyssey that shows what Grok @Imagine’s video and voice capabilities can do. We’re awarding $100K, $50K, and $25K to the top three videos submitted by quoting this post. 🧵
278
350
2,354
956,267
Elon Musk retweeted
1/2 the turns and 25% of the tokens. Task/Token is something to watch. Is Grok 4.6 inside Grok Bot? Has this been revealed yet? Have I missed it? I’ve been wondering what magic allows grok bot to just work so consistently.
Agentic work is where @grok 4.6 lands hardest, taking the top spot on the Artificial Analysis Agentic Index at 59, tied with Claude Opus 5 Max. The index measures tool use, planning, autonomy and complex problem solving rather than single answers Grok 4.6 completes tasks in ~53 turns and ~0.5bn input tokens on average, against ~103 turns and ~2.0bn for Claude Opus 5 Max Cost of $0.84 per task, putting it on the intelligence versus cost per task Pareto frontier Enterprises buying agents pay per completed task, not per benchmark point. Turn efficiency is what determines whether a long-running workflow is affordable at volume. Two labs now sit at the top of this index with very different cost structures. Buyers get real choice on price for the first time in agentic deployment.
88
185
1,290
745,136
Grok is very good at agentic tasks!
Agentic work is where @grok 4.6 lands hardest, taking the top spot on the Artificial Analysis Agentic Index at 59, tied with Claude Opus 5 Max. The index measures tool use, planning, autonomy and complex problem solving rather than single answers Grok 4.6 completes tasks in ~53 turns and ~0.5bn input tokens on average, against ~103 turns and ~2.0bn for Claude Opus 5 Max Cost of $0.84 per task, putting it on the intelligence versus cost per task Pareto frontier Enterprises buying agents pay per completed task, not per benchmark point. Turn efficiency is what determines whether a long-running workflow is affordable at volume. Two labs now sit at the top of this index with very different cost structures. Buyers get real choice on price for the first time in agentic deployment.
1,186
1,310
8,226
3,991,657
Create with Grok Imagine
Homer had a lyre. You have Grok Imagine. Create a compelling scene from Homer’s The Odyssey that shows what Grok @Imagine’s video and voice capabilities can do. We’re awarding $100K, $50K, and $25K to the top three videos submitted by quoting this post. 🧵
1,185
1,567
9,860
6,266,593
We’re working with Southaven officials on the development of a new police and fire training facility, investing $40 million to help strengthen the local emergency response systems that are critical for the safety and well-being of the communities we call home.
236
518
5,657
631,892