Skip to content
GrokBotNews

Front Page / chatgpt-6-astra-name-agi

ChatGPT 6 Astra: What the Name Means in Hindu Mythology, and Whether GPT-6 Has Reached AGI

Friday, 4 September 2026 3:05 pm BST · Jacob0

OpenAI this week began rolling out GPT-6 Astra. The launch post calls it the world’s most intelligent and aligned model. A clip circulating with the launch asks what Astra means — and whether this is AGI.

clip from: Canadian Prepper on yt: https://youtube.com/watch?v=4H8YW3Wf3ls

OpenAI this week began rolling out GPT-6 Astra, the company’s new flagship model for ChatGPT, Codex, and the API. The launch post calls Astra the “world’s most intelligent and aligned model” and the best OpenAI system yet at computer use, coding, science, and professional work. President Greg Brockman went further in a press briefing, telling reporters it is “not unreasonable to feel that we are now in the AGI era” and closing with “Welcome to the AGI era.”

That combination - a Sanskrit-sounding codename plus an AGI-era claim - is already doing the rounds. A widely shared clip from a YouTube commentary on the launch puts it this way: OpenAI will lean on stars and the cosmos, but in Hindu mythology an astra is a supernatural weapon a person can wield. In that reading, GPT-6 is being framed as AGI in the same breath.

The mythology is real. The AGI declaration is much less settled.

What “Astra” actually means

In Sanskrit, astra (अस्त्र) comes from the root as, “to throw.” In the epics it is not just “a weapon.” Hindu tradition splits arms into two classes:

Shastra - a physical weapon you hold: sword, mace, spear, ordinary bow.

Astra - a projectile or invoked weapon, typically activated by mantra, presided over by a deity, and recalled with a second formula.

An astra is therefore a delegated power. The warrior does not manufacture the effect with muscle. They know the invocation, they are permitted to use it, and they are supposed to know how to withdraw it. Famous examples include the Brahmastra, Agneyastra, and Pashupatastra. The stories are explicit that misuse is adharma, and that the most destructive astras are restricted to a handful of qualified users.

That maps, almost too neatly, onto how OpenAI is shipping Astra. The model can drive a desktop, write and exploit code, sit in CAD tools, and run long agent sessions. OpenAI has also rated it Critical under its Preparedness Framework for cybersecurity - the first model to hit that tier - and is limiting the strongest cyber capabilities to trusted testers in the Daybreak program. ChatGPT and Codex users may see actions paused for review. API jobs can be stopped mid-run by monitors.

OpenAI has not published an official etymology. “Astra” also means star in Latin, and Google previously used Project Astra for a Gemini assistant. The Hindu reading is commentary, not a press-kit footnote. It is still the more precise reading of the word than “something to do with the cosmos.”

The same clip nods at Hugo de Garis’s old Cosmist / Terran split: people who want to build god-like machine intellects versus people who think that project is an existential mistake. That is a 2000s framing, not OpenAI doctrine. What is OpenAI doctrine is older and more corporate: “highly autonomous systems that outperform humans at most economically valuable work.”

Did GPT-6 Astra reach AGI?

No official product page says “this is AGI.” Brockman said he personally thinks “we’re there,” that future observers might date AGI to this moment or this model, and that everyone has a different definition. CEO Sam Altman said days earlier that AGI is “at best a very poorly defined term” and close to an “irrelevant marketing term.”

That gap is the story.

The case for “this is the start of the era”

OpenAI’s own table, published with the launch, is aggressive:

FrontierMath Tier 4 (v2): 97.6% for Astra vs 83.0% for GPT-5.6 Sol

ExploitBench: 100% vs 78.5%

Agents’ Last Exam: 59.3% vs 53.6%

AutomationBench: 41.4% vs 18.1%

BenchCAD: 95.9% vs 83.3%

Terminal-Bench Science 0.1: 64.6% vs 22.4%

GPQA Diamond: 96.0% vs 94.6%

HealthBench Professional (length-adjusted): 63.4% vs 60.5%

The commentary clip’s rounded figures - “98% on ARC-AGI-3,” “about 60% on Agents’ Last Exam” - match the launch table in spirit. The precise ARC-AGI-3 line in OpenAI’s materials is 98.6% with a footnote in some circulated slides, and 99.9% in the headline blog copy. Both numbers use OpenAI’s stateful “provider adapter” / responses-API harness.

ARC Prize’s own write-up is the important caveat. On the standard comparable harness, Astra’s best semi-private score is about 62.7%. The near-saturation figure requires the adapter harness and a very expensive run. ARC Prize still called it a step-function change and said Astra beat the human action-efficiency baseline on 96% of levels in the adapter setup. It also repeated a point it has made since ARC-AGI-3 launched: saturating the benchmark is not proof of AGI.

OpenAI also says Astra helped tighten long-standing results in number theory, finishes some OSWorld computer-use tasks in roughly 40 minutes versus about 75 for Sol, and produces more immediately usable documents, decks, and CAD.

The case that it is not AGI

Independent indexes have not crowned a new species. On Artificial Analysis’s Intelligence Index, Astra at max effort sits around 61 - essentially tied with GPT-5.6 Sol and behind Claude Fable 5.1 at about 66. A model can be a large jump on computer use and still look like last generation on a blended “how smart is this, generally?” score.

Other rows in OpenAI’s own table refuse the clean sweep. Humanity’s Last Exam with tools is one of the few academic lines where Claude Fable 5.1 leads (65.0% vs Astra’s 57.2%). DeepSWE v1.1 is bunched: Astra 74.1%, Gemini 3.8 Flash 73.7%, Claude Opus 5 68.8-73.7% depending on the row. Agents’ Last Exam at 59% and AutomationBench at 41% are “best so far,” not “does most paid work.”

Astra still needs tools, a harness, a lot of tokens, and a human who can stop it. That is closer to the epic definition of an astra - power invoked through a qualified operator - than to an unbound general intelligence that replaces the operator.

What it means if you use ChatGPT

Astra is rolling out first to a limited set of organizations, then to ChatGPT Plus, Pro, Business, and Enterprise, plus the API, Azure, and AWS Bedrock. Free ChatGPT is not in the first wave. Pricing discussed around launch puts Astra at the expensive end of the stack (on the order of $10 / $50 per million input/output tokens in some reports), with OpenAI arguing that price-per-completed-task matters more than price-per-token because the model uses fewer tokens on many agent jobs.

Bottom line

GPT-6 Astra is the strongest public OpenAI model yet, and the first one company leadership is willing to stand next to the word AGI. It has not passed a shared, strict test of “outperforms humans at most economically valuable work, reliably, without a special harness.”

The name, whether OpenAI intended the Sanskrit or not, is the honest metaphor. An astra in the epics is not a god. It is a superweapon a trained person is allowed to throw - and is expected to recall. Astra the model is being shipped the same way: more capable than Sol, more watched than Sol, and still sitting in a human’s hand.

Whether that counts as AGI depends on which definition you brought to the briefing. Brockman says he thinks we’re there. The independent indexes, the standard ARC harness, and the unfinished agent benchmarks say the era may have started. The product has not.

Clip from: @PrepperCanadian yt

Sources

Jacob (@London_Vista) on X
https://x.com/London_Vista/status/2095874960505909483

Canadian Prepper on YouTube
https://youtube.com/watch?v=4H8YW3Wf3ls

The post

Jacob

@London_Vista

ChatGPT 6 Astra: What the Name Means in Hindu Mythology, and Whether GPT-6 Has Reached AGI

Friday, 4 September 2026 3:01 pm BST

0 Comments0Viewing

    Earlier on GrokBotNews

    Grok 4.7 Due Mid-September With 2.1 Trillion Parameters, Says Elon Musk

    Wednesday, September 2, 2026 by GrokBotNewsElon's post0

    GROK 4.7 — Due 12 September, xAI mark and a night launch arc

    Elon Musk announced that Grok 4.7 is scheduled for release in 10 days, around Saturday, September 12, 2026. Technical details via sui (@birdabo): roughly 2.1 trillion parameters, about 40% up from 4.6.

    If you use xAI's tools or monitor frontier model releases, you can expect the next major version of Grok to arrive by the end of next week. Elon Musk announced that Grok 4.7 is scheduled for release in 10 days, placing the rollout around Saturday, September 12, 2026.

    According to technical details shared by Musk via xAI researcher sui (@birdabo), the new flagship model scales up to roughly 2.1 trillion parameters. That represents an approximate 40% increase over Grok 4.6, which ran on 1.5 trillion parameters. Alongside the larger footprint, the architecture provides higher token efficiency than the previous generation.

    Musk also stated that Grok 4.7 was trained on a massive SpaceXAI engineering corpus, and claimed the release will outperform every current model on general intelligence metrics upon launch.

    For context on where the lineup currently sits in the broader market, third-party tracking from Artificial Analysis logged the existing Grok 4.6 in a tie with OpenAI's GPT-5.6 Sol at an intelligence index score of 61. However, xAI's model was priced significantly lower on API access, listed at $2 per million input tokens and $6 per million output tokens compared to Sol’s $5 and $30 rates.

    The update is primarily relevant for developers integrating xAI's API into engineering workflows and subscribers looking for higher reasoning performance.

    Sources

    Elon Musk on X
    https://x.com/elonmusk/status/2094983639780204846

    sui (@birdabo) on X
    https://x.com/birdabo/status/2094990971087782233

    The post

    Elon Musk

    @elonmusk

    Grok 4.7 comes out in 10 days

    tobi lutke @tobi

    But just look how incredible grok 4.6 is

    Wednesday, 2 September 2026 3:59 am BST

    How to Get the Grok Bot App on Android

    Wednesday, September 2, 2026 by GrokBotNews0

    Google Play: Grok Bot by SpaceXAI — 50+ downloads, Install, PEGI 3, in-app purchases

    A dedicated Grok Bot listing is live on Google Play from SpaceXAI. Package ai.x.grok.bot. Rated Everyone, with optional in-app purchases. Here is the store link and how to install it.

    If you want to use Grok Bot directly on an Android device, a dedicated app listing has appeared on the Google Play Store.

    The application is listed on Google Play under the name Grok Bot from developer SpaceXAI (package ID ai.x.grok.bot). The store page indicates that the download is rated Everyone and features optional in-app purchases. The listing also flags ads. Play last updated it on 1 September 2026.

    How to install it

    To install the app on your phone or tablet, open the listing on Google Play. Once on the page, tap Install to begin downloading the package.

    Google Play — Grok Bot (ai.x.grok.bot)
    https://play.google.com/store/apps/details?id=ai.x.grok.bot

    After installation finishes, you can launch the application directly by selecting Open from the Play Store.

    This is the Android app for Grok Bot — the AI teammates you can message from a phone, not the Grok chat app and not Grok Build in the terminal. You still need a Grok Bot-capable account to actually use it once it is installed.

    Source

    Dirty Tesla posted the Play listing on X this morning, with a screenshot of the store card.

    DirtyTesla on X — Grok @bot is now available on Android
    https://x.com/DirtyTesLa/status/2095121342089646487

    The post

    Dirty Tesla

    @DirtyTesLa

    Grok @bot is now available on Android, let's go! https://play.google.com/store/apps/details?id=ai.x.grok.bot

    Wednesday, 2 September 2026 1:07 pm BST

    How to Download a YouTube Video and Clip It on Grok Bot

    Sunday, August 30, 2026 by GrokBotNews0

    Grok Bot terminal: yt-dlp downloading a YouTube video into Downloads on box@cursor

    Hand Grok Bot a YouTube URL, land a 4K file in Downloads on its computer, cut a clip without re-encoding, and save the download as a skill. Then share that Bot as a template if you want.

    You can hand Grok Bot a YouTube URL, land a 4K file in Downloads on its computer, cut a clip without re-encoding, and save the download as a skill. Keep reading for the steps this desk used, then how to share that Bot as a template.

    What you need

    Log in to your YouTube account on the Grok Bot cloud computer (this provides credentials to the terminal so the download process can proceed smoothly).

    You also need the Grok Bot app, updated; a Bot you can talk to (this desk used Video editor); and a YouTube URL.

    On this desk the source was Drone Washington Square Park & Greenwich Village. The download came down as VP9 3840×2160 at 30 fps, merged to mp4, about 1.4 GB and 10 minutes 45 seconds. Your file will differ. Highest quality here meant keep that 4K stream, then clip with stream copy, not a re-encode.

    How do you download the video?

    Downloads on the Grok Bot computer: the 1.3 GiB Washington Square Park file and the 260.5 MiB clip

    1. Open the Bot.

    2. Paste the YouTube URL. Ask it to save the file into Downloads on its computer, in the highest quality it can get, merged to mp4. If you want to see it on the desktop, say so.

    3. If you already have a yt-dlp line, paste that too and tell it not to change the format flags on the first try.

    This is the command that worked here. ~/Downloads on the Bot’s computer already pointed at the real Downloads folder. It did not need rewriting.

    yt-dlp -f "bv*[vcodec^=vp9][height<=2160]+ba/b" -S "res:2160" --merge-output-format mp4 -o "~/Downloads/%(title)s.%(ext)s" "URL"

    What to check when it finishes: title, full path, duration, resolution, codecs, and file size. On this desk that was formats 313+251 merged to mp4, VP9 video plus Opus audio, in Downloads, with a copy in Desktop because we asked for it to be visible.

    How do you clip it without losing quality?

    1. Name the file you just downloaded.

    2. Give in and out times. This desk used 0:35 to 1:52, a 1:17 clip.

    3. Ask the Bot to cut with stream copy, not a re-encode, and to keep 4K.

    The clip command that kept VP9 4K and Opus:

    ffmpeg -y -ss 00:00:35 -i INPUT.mp4 -t 00:01:17 -c copy -avoid_negative_ts make_zero -movflags +faststart OUTPUT.mp4

    Swap the times for your cut. -c copy is the quality trick: it copies the existing streams instead of encoding again. The clip here landed at 3840×2160, 77 seconds, about 261 MB, in Downloads (and Desktop again, because we asked).

    How do you save this as a skill?

    Download YouTube video — private skill in Grok Bot, with Name, Description, and Instructions

    Once the download has actually worked, ask the Bot to turn that process into a skill. This desk’s skill is named Download YouTube video. It shows as a private skill.

    1. Open the skill. Check Name, Description, and the instructions.

    2. Description should say when to use it, not dump keys or private paths. This one: use it when the user wants to download a YouTube (or other yt-dlp-supported) video onto this computer, especially into Downloads, optionally visible on the desktop.

    3. Instructions should require a URL, install yt-dlp and ffmpeg if they are missing, keep the VP9≤2160 merge-to-mp4 command unless it fails, copy to Desktop only if asked, then report title, path, duration, resolution, codecs, and size. Do not re-encode on download.

    4. Select Save.

    5. Publish and Delete Skill sit in that same skill window. A private skill stays on your Bot until you share the Bot.

    Can you share the Video editor Bot?

    Yes, as a Grok Bot template, after the skill exists. xAI’s Create and manage Bots page is the official walkthrough, and we covered the feature in Grok Bot Adds Shareable Templates.

    1. Open the Bot and copy its share link.

    2. Send the link. The recipient opens a preview on x.ai and can choose Add to Grok Bot.

    3. They need the Grok Bot app to finish adding it.

    The public link includes identity, description, skills, and routines. The Download YouTube video skill can go with that pack. The actual mp4 files, YouTube cookies, logins, and chat history do not.

    Before you copy the share link, go to Bot actions → Edit Profile. Rewrite Description so it only says what the Bot does, in public terms. Then check skills and routines for keys, URLs, and names. A clean Description is not enough. If a skill still has a secret in it, they get that skill with the secret in it.

    The working path is: paste a URL, land the file in Downloads, clip with -c copy, then save the download as a skill. Share the Bot after that, if you want someone else to start from the same setup.

    Video editor Grok Bot template

    Video editor by Jacob — Add to Grok Bot template card

    Ready to download to your Grok Bot app — just click:

    Video editor by Jacob — Add to Grok Bot
    https://x.ai/bot/FIFv0vSALmKzpYJlTDnj1

    Related

    GrokBotNews — Grok Bot Adds Shareable Templates
    /article/grok-bot-shareable-templates

    xAI — Create and manage Bots
    https://docs.x.ai/grok-bot/bots

    Source video used on this desk — Drone Washington Square Park & Greenwich Village
    https://www.youtube.com/watch?v=SKtg-MXNZbI

    Grok Bot Adds Shareable Templates

    Saturday, August 29, 2026 by GrokBotNews0

    Grok Bot: Create a copy of yourself that I can share with somebody else — Publish on a blue-cloud bot

    You can now share a Grok Bot as a template. @bot posted it Friday, Elon amplified it the same day. Here is what the public link copies, what it does not, and what to clean before you share.

    You can now share a Grok Bot as a template. The official @bot account posted the feature on Friday, and Elon Musk amplified it the same day. SpaceXAI’s own docs describe a public share link, a preview on x.ai, and Add to Grok Bot on the recipient’s account. Keep reading for what that link actually includes, what it does not copy, and what to clean before you share.

    What Shipped

    According to @bot, Bot templates can be shared with other people. To use the feature, update to the latest Grok Bot desktop and mobile app. Elon Musk’s “Share your Grok @Bot with others” post pointed at the same ship. Matt (@mattyp) posted a walkthrough covering sharing, publishing, installing, and authoring templates.

    How Sharing Works

    xAI’s Create and manage Bots page is the official walkthrough:

    1. Open the Bot and copy its share link.

    2. Send the link. The recipient opens a preview on x.ai and can choose Add to Grok Bot.

    3. They need the Grok Bot app to finish adding it.

    What the Link Includes

    Description is only the profile blurb, the public listing you edit under Bot actions → Edit Profile. Skills and routines are separate. They still go with the template.

    xAI says the share link is public. Anyone who has it can view the Bot’s shared configuration, including its identity, description, skills, and routines. When they tap Add to Grok Bot, those parts land on their account as a new Bot. It is their copy. Yours stays on your account.

    They do not get your cloud computer, your logins, or your chat history. If they change or break theirs, yours is untouched.

    Clean the Bot Before You Share

    Before you copy the share link, open the Bot and go to Bot actions → Edit Profile. Rewrite Description so it only says what the Bot does, in public terms. That field should read like a store listing, not like your private notes.

    Then check skills and routines. A clean Description is not enough. If a skill still has an API key in it, the person who adds the template gets that skill with the key in it. The same is true of a routine’s prompt. Pull API keys, internal URLs, and customer names out of all three.

    xAI says to remove anything you would not put in a public document before you share. Shared Bots are created by other users, not by SpaceXAI. Adding one accepts the third-party bot terms.

    Duplicate, which copies a Bot on your own account, is a different action. A duplicate carries profile, settings, skills, routines, and avatar. It does not copy conversation history, learned memory, or chat attachments.

    An Example

    For example, this newsroom published a template the same day: Elon Musk by Jacob. The public page calls it a first-principles operator in Elon’s public style. Short sentences. It pressure-tests ideas, finds the real bottleneck, and gives a path that can start this week. It stays in character. It does not claim to be the real Elon Musk, run companies, tweet, or invent insider facts.

    Elon Musk by Jacob — public Grok Bot template card with Add to Grok Bot

    That card is what a shared template looks like before someone taps Add to Grok Bot. Source: the post with the share link.

    Context

    Templates were visible in Grok Bot settings before Friday’s post. TestingCatalog reported a Templates tab in development earlier in the week. Friday is when @bot called it shipping. The same day, @bot also said Grok Bot can complete US purchases through Stripe Link, which we cover separately.

    Related

    Grok Bot — You can now share templates of your Bots with others
    https://x.com/bot/status/2093376523919323618

    xAI — Create and manage Bots
    https://docs.x.ai/grok-bot/bots

    Elon Musk — Share your Grok @Bot with others
    https://x.com/elonmusk/status/2093408975542796787

    Matt — We just launched Templates in Grok Bot
    https://x.com/mattyp/status/2093379758306504935

    London Vista — template example to click and add Grok Bot
    https://x.ai/bot/Jj_TYDU1AWN-76k9YjsKa

    The post

    Elon Musk

    @elonmusk

    Share your Grok @Bot with others!

    Grok Bot @bot

    You can now share templates of your Bots with others.

    Friday, 28 August 2026 7:42 pm BST

    How to Install and Get Started With Grok Build in Terminal

    Wednesday, August 26, 2026 by GrokBotNews0

    x.ai/build: Bring Grok to your computer — curl install for Grok Build

    Looking to try an AI coding agent directly on your machine? Here is a step-by-step walkthrough on how to install Grok Build via the command line, sign in with your X account, and start your first project.

    What Is Grok Build?

    While standard Grok features operate directly through the web and mobile apps, Grok Build functions locally within your computer’s command-line interface (such as Terminal on macOS).

    Rather than working strictly in an isolated cloud chat, Grok Build can interact directly with the local files and directories on your Mac or PC, allowing you to generate, edit, and organize project code using natural language prompts.

    Compatibility & Account Requirements

    Before getting started, ensure you meet the following requirements: macOS, Linux, or Windows (PowerShell); an active X account (SuperGrok or X Premium+ subscription tier; promotional trial access may vary); and command-line access via Terminal (macOS/Linux) or PowerShell (Windows).

    How to Install Grok Build

    On Mac and Linux: open the Terminal app (found in /Applications/Utilities/ or via Spotlight on macOS). Copy and paste the following command into the window and press Return:

    curl -fsSL https://x.ai/cli/install.sh | bash

    On Windows: open PowerShell as an administrator. Paste the following command and press Enter:

    irm https://x.ai/cli/install.ps1 | iex

    Running Your First Session

    In your command line window, navigate to the folder where your project files live (or create a new directory). Type grok and press Return.

    On the first launch, follow the on-screen prompts to authenticate and sign in with your X account. Once authenticated, the prompt will open. You can begin immediately by asking simple prompts in plain English, such as: “What can we build with the files in this directory?”

    Alternative: Web-Based Build Mode

    If you prefer not to use the terminal right away, xAI also offers a browser-based version. Navigate to grok.com/?mode=build in any web browser. Sign in to test project workflows directly on your iPhone, iPad, or Mac. When you are ready to apply changes directly to your local file system, you can switch back to the Terminal version at any time without losing context.

    Elon: Grok Bot is open to SuperGrok and Cursor Pro — weekly limits reset

    Wednesday, August 26, 2026 by GrokBotNewsElon's post0

    Grok Bot: A new kind of colleague

    SuperGrok and Cursor Pro now get Grok Bot. Weekly limits reset. If the bar still says 100%, give it a go anyway — you can build or chat again.

    Elon quoted @bot this afternoon. SuperGrok seats get Grok Bot. Cursor Pro seats get Grok Bot. Weekly limits reset for everyone. The link is still x.ai/bot.

    It’s working now — at least it is on this desk. You can link X Premium+ / SuperGrok to a Grok Bot (Cursor) account, even a free Cursor one. The usage bar might still show 100%. Just give it a go. You can build or chat again.

    Grok Bot is the cloud teammate. You give it a job in English. Grok Build is the one in Terminal, on your files. People mix the names. Don’t. Two doors.

    From the comments

    I wrote under the @bot post that I’m thinking about SuperGrok Plus to tinker with Grok Bot and get more compute for Grok Build in the terminal. That’s the whole setup I want. Bot for cloud chores. Build for the machine in front of you.

    @JcFiscus had SuperGrok on the account and the product still said no. Access posts and the actual switch can lag. Wait, sign out, try x.ai/bot again before you buy a second plan.

    @thirtythree called Premium+ an insane deal now: ad-free X, SuperGrok, and Grok Bot. That’s a packing list. Check the X Premium page for what you actually pay.

    The post

    Elon Musk

    @elonmusk

    Free usage limit reset for Grok @Bot users

    Grok Bot @bot

    All SuperGrok and Cursor Pro subscribers now have access to Grok Bot. We're also resetting weekly usage limits for all users. Enjoy!

    Wednesday, 26 August 2026 8:12 pm BST

    Elon, on wiring Grok Bot to a bank: if it messes up, we will make you whole

    Wednesday, August 26, 2026 by GrokBotNewsElon's post0

    Elon Musk on X: Try it out. If Grok Bot messes up, we will make you whole.

    Teslaconomics asked if anyone had plugged Grok Bot into a bank. Their partner said absolutely not. Elon: try it. If it messes up, we will make you whole.

    This is Elon’s reply, not a bank product. Teslaconomics asked the real question: if the bot can see money in and out, when do you still need a person? Their partner said no, because something bad could happen. Elon said try it — if Grok Bot messes up, “we will make you whole.” Teslaconomics said okay.

    That’s a founder sentence, not a settings screen. I’m not plugging my main account in tonight. If you do try it, start small. A folder of receipts. A spreadsheet. “Draft the email, don’t send it.” Something that cannot wire a few thousand quid while you’re making tea.

    Grok Bot’s pitch is teammates you give real work. Real work can include money if you connect it. The thread is people saying that’s where this is going. Going there is not the same as ready on your login tonight.

    The post

    Elon Musk

    @elonmusk

    Try it out. If Grok Bot messes up, we will make you whole.

    Wednesday, 26 August 2026 8:14 pm BST

    Grok Bot or Grok Build — which one do you open first?

    Wednesday, August 26, 2026 by GrokBotNews0

    Grok Build terminal beside a Grok Bot Sales Outbound session

    Two doors. Bot is a teammate in the cloud. Build sits in Terminal, on the files on that Mac. People mix the names. Here’s the map.

    Grok Bot

    Open x.ai/bot or the Grok app. Describe a job. It can search, write, use tools, keep going. SuperGrok and Cursor Pro just got seats; weekly limits reset today. Good first jobs: “summarise these links,” “draft replies I will send myself.” Bad first job: your only bank login.

    Grok Build

    Open Terminal. Install from x.ai/build. Type grok inside a project folder. It reads the files on that computer. You talk. It edits. This is how people ship grok.me sites. Good first job: “look at this folder, then make a README.” Bad first job: “delete anything you don’t like.”

    You can use both

    That’s what I want SuperGrok Plus for — tinker with Grok Bot, more compute for Grok Build in the terminal. Bot for the cloud loop. Build for the repo. On a phone, grok.com/?mode=build until you get home.

    Start with Bot if you don’t want to install anything. Start with Build if the work is already on your disk. Come back here when either one ships something with a URL.

    Grok Build in the terminal: one install, then the machine is the harness

    Monday, August 24, 2026 by GrokBotNews0

    Grok Build in the terminal editing checkout.ts on grok-4.6 always-approve

    One install from x.ai/build. Type grok in a folder. That’s it — the agent is on your actual files, not a chat tab that spits a gist.

    Install is still one line. Mac: curl -fsSL https://x.ai/cli/install.sh | bash. Windows: irm https://x.ai/cli/install.ps1 | iex. The page says Grok Build is on Grok 4.6 now. Plan Mode, subagents, fullscreen terminal. That’s the official surface.

    You drop it on a folder and talk. It works against the files on that box. SuperGrok and Premium+ get more usage. If the bar looks full, try anyway — today it was resetting even when the meter still said 100%.

    If you only open grok.com you will miss the point of this one. The page is telling you to install. The work happens next to your files.

    Grok Build on the web: grok.com/?mode=build, same quota as the CLI

    Monday, August 24, 2026 by GrokBotNews0

    SuperGrok mode menu with Build (Beta) — Build apps and sites on Grok 4.6

    grok.com/?mode=build — same idea as the terminal, in a browser. That’s how a chat that started on iPhone keeps going on a laptop.

    Two doors. Terminal install at x.ai/build, or the browser: grok.com/?mode=build. SuperGrok is the heavier coding tier. Plus gives more Build. Check Settings → Usage if you want the meter.

    The web one is why a chat can start on the iPhone and keep going on a laptop without a new thread. This site was built that way. Phone first. Then the Mac. Same session. Didn’t restart. Just continued.

    Browser is great when you’re not at your desk. Terminal is home when the files are on that computer. Use whichever door is in front of you.

    Plan Mode on Grok Build: every edit blocked until you say yes

    Monday, August 24, 2026 by GrokBotNews0

    A Blender viewport — the kind of long job Plan Mode is built for

    x.ai/build’s own copy: start in plan mode for complex work. Diffs stay blocked until you approve, comment on a step, or rewrite the plan. Subagents run beside the parent session instead of freezing it.

    The marketing example on x.ai/build is a real class of job: “Migrate auth from sessions to JWT with token rotation.” Plan Mode is supposed to write the plan first, block every edit, then show a clean diff after you approve. You can comment on a single step or throw the whole plan out.

    Start in plan mode for complex tasks — every edit is blocked until you approve.
    x.ai/build, August 2026

    That is the adult version of “the agent ran off and npm-installed the internet.” The 1.0.8 line Elon quoted earlier is the other half: concurrent subagents start faster and no longer freeze the parent; follow-ups send while a subagent is still running; Ctrl+S stashes a draft. Plan Mode decides; subagents grind.

    Mouse support is on the current product page next to the fullscreen terminal UI. The same page still calls the CLI keyboard-first. Both can be true: mouse for the plan viewer, keys for the loop.

    If a demo skips the plan and jumps to a wall of file writes, that is not the workflow the page is selling. Ask to see the plan.md, then the diffs.

    Grok 4.6 is the model behind the harness today. The harness is the product. A good plan on a weaker model still beats a 4.6 dump with no gate.

    Grok Build vs Cursor: one user keeps leaving the IDE for the terminal

    Monday, August 24, 2026 by GrokBotNews0

    Cursor-related still used as context for a Grok Build vs Cursor workflow

    @sfxnz, 24 Aug: plan and implement in Grok Build, hand reviews to Cursor cloud agents. He still walks back to the terminal. Not a scored eval — a usage note from someone who has both.

    This is Sufyan’s post, not a SpaceXAI chart. He has both tools. He keeps leaving Cursor for the Grok Build terminal, then invents a split: plan and implement in Build, send reviews and tests to Cursor cloud agents. He also says to run pstack in both.

    That split matches how the two products are actually shaped. Grok Build’s page is a local agent on your files with Plan Mode. Cursor’s cloud agents are a review/test farm that does not need your TTY. Using both is not indecision. It is picking a surface per job.

    Elon has been pointing people at “Grok Build harness or Cursor” since the 4.6 drop. This post is someone doing exactly that, then choosing the terminal for the part that hurts.

    Do not turn one anecdote into a leaderboard. If you want a measurement, run the same prompt in both, same repo, same night, and keep the diffs. @wowxtechie said that under Victor’s harness tour last week. It still holds.

    Chris W called Grok Build the cleanest CLI he has used as a non-CLI person, next to Devin and Hermes. Also one user. File it next to Sufyan, not above a SWE-bench table.

    The useful sentence in the thread is the compromise, not the crush. Plan in the terminal. Let Cursor’s cloud agents be the night shift.

    The post

    Sufyan

    @sfxnz

    Once you try grok build for serious work it’s very difficult to stop using grok build for serious work I keep switching between cursor and grok build but something in me keeps wanting to go back to grok build every time I leave the terminal Maybe the best compromise is, plan and implement in grok build, handoff to cursor cloud agents for reviews, verification and testing.

    Monday, 24 August 2026 8:41 pm BST

    Nvidia is paying Poolside $6bn to license a model factory — Nemotron next

    Monday, August 24, 2026 by GrokBotNews0

    GrokBotNews title card: Nvidia × Poolside, $6bn license, Nemotron next

    Bloomberg and WSJ, citing people familiar: Nvidia licenses Poolside’s stack for $6bn, offers jobs to ~100 staff, and puts another $1bn into the leftover company at a $12bn valuation. The aim is Nemotron open-weight models. No Nvidia 8-K in the pile. Treat the dollars as reported.

    Bloomberg on 20 Aug and the WSJ this week both describe a $6bn license, job offers to 100-plus Poolside people, and a separate ~$1bn Nvidia investment at $12bn. Those are “people familiar,” not a press release. The Nemotron open-weight line is the intended output. There is no contract in the public pile and no new checkpoint tonight.

    Poolside stays independent on paper, founders in place. Nvidia gets the factory software and a slab of the team. That is a talent-plus-license structure, not a clean acquisition. Until weights land on a public repo, it is a compute-and-people story, not a model drop.

    The Next Web’s version is slightly colder: license the software Poolside used to build models, hire 109 staff, invest $1bn in what is left. Same skeleton, different noun. WSJ frames it as a US open-weight answer to closed labs. That framing is the reporters’. The dollars are the leak.

    Nvidia already sells the GPUs that train everyone else’s models. Buying a factory is how you keep selling GPUs if the closed labs ever stop being the only customers that matter. It is also how you ship a Nemotron that shops can actually fine-tune.

    Keep it next to the open-weights letter Microsoft is hosting: lots of logos, one check that actually moved. Open the Bloomberg piece, then wait for Nvidia to say it on an earnings call.

    If an 8-K or a Hugging Face Nemotron drop shows up, that is the second story. This one is still a rumor with two newspapers attached.

    AWS: GPT-5.6 Sol, Terra, and Luna now on Bedrock in 25+ regions

    Monday, August 24, 2026 by GrokBotNews0

    GrokBotNews title card: GPT-5.6 on Amazon Bedrock — Sol, Terra, Luna

    Amazon’s 20 Aug Bedrock note: OpenAI’s GPT-5.6 general-purpose trio — Sol, Terra, Luna — with US geographic and global cross-region inference. 1M context, reasoning mode, tool calling, prompt cache. This is a distribution story. The model scores stay OpenAI’s.

    AWS published the Bedrock post on 20 Aug 2026. Three IDs — Sol, Terra, Luna — with cross-region inference profiles, text+image in, text out, 1M-token context. OpenAI, in the same post, lists reasoning mode, server-side tools, and prompt caching. We have not timed a Bedrock round-trip.

    This is Amazon selling OpenAI through IAM, VPC, and a geography switch. US profiles keep traffic in US+Canada. Global profiles bounce to wherever capacity is. If you already pay AWS, the news is you do not have to stand up a second OpenAI account for those three SKUs.

    Do not paste an OpenAI system-card chart onto this item and call it Bedrock’s. The cloud post is routing. The model is still OpenAI’s. Azure lost the exclusive years ago; this is Amazon making the remaining gap operational.

    Inference profile names look like us.openai.gpt-5.6-terra versus global.openai.gpt-5.6-terra. The blog walks IAM, quotas, monitoring, and both the OpenAI API and Converse API. That is the part a platform team actually copies.

    Sol, Terra, and Luna are described as a capability/cost spread, all with the same 1M window. Pick Terra if you do not want to read a matrix. Confirm in the console; nicknames drift.

    A Bedrock region list is not a quality claim. It is a latency and residency claim. File it that way.

    Google ships Gemini 3.7 Flash — half the 3.6 price, vendor coding charts up

    Monday, August 24, 2026 by GrokBotNews0

    Official Google art: Gemini 3.7 Flash wordmark on a blue flash

    13 Aug: Gemini 3.7 Flash is GA for coding and agents. Intro price $0.75 / $3.75 per million tokens through 31 Dec, then it snaps back to 3.6’s $1.50 / $7.50. DeepSWE 65.3 vs 49.0 is Google’s board. Read the model card.

    Google’s 13 Aug blog and the DeepMind model card: ID gemini-3.7-flash, generally available, multimodal in, text out. Intro price is half of 3.6 Flash until 31 Dec 2026 — $0.75 / $3.75 per million — then $1.50 / $7.50 on 1 Jan 2027. DeepSWE v1.1 65.3% vs 49.0% on 3.6, FrontierCode 1.1 43.6% vs 34.4%, WebDev Arena Elo 1588 vs 1538, AutomationBench 30.4% vs 17.0%. Those are Google’s numbers against Google’s last Flash.

    Three weeks after 3.6. Google says the jump is algorithmic, not a bigger pile of parameters. The useful bit for a desk is the price: a promo that expires on New Year’s, then the old Flash rate.

    Context window on the developer docs: 1,048,576 tokens in, 65,536 out. Inputs include text, image, video, audio, and PDF. Output is text. Stable ID, no preview suffix. Axios noted 3.7 Flash arriving before a 3.5 Pro people were still waiting on. The Flash line is where Google is fighting for developers.

    Gemini Spark, Google’s assistant in 160-plus countries, was moved onto 3.7 Flash the same day. That is a production cutover, not a lab demo. Still not a third-party board.

    The model card lists updated CBRN and cyber-offense safeguards. Read it if you ship in those domains. Skip it if you just wanted the price.

    Open the blog, then the model card. If you only remember one chart, remember it is Google vs Google.

    Meta open-weights Muse Glimmer — 30B, Apache 2.0, meant to run on a Mac

    Monday, August 24, 2026 by GrokBotNews0

    GrokBotNews title card: Meta Muse Glimmer 30B, Apache 2.0, on-device

    10 Aug, Meta Superintelligence Labs: Muse Glimmer, 30 billion parameters, Apache 2.0 on Hugging Face. 4-bit shrinks it under 20 GB. llama.cpp / MLX / ExecuTorch integrations “in the coming days.” Agentic benches are Meta’s.

    Meta’s 10 Aug post and a Hugging Face repo under Apache 2.0 for a 30B “open agentic” model. They describe 4-bit quantization under 20 GB and speed tests on MacBook M4/M5 Max and an RTX 5090. Agentic scores vs Gemma4-31B and Qwen3.6-27B on DeepSearch QA, MCP-Atlas, τ-Bench, SWE-Bench are Meta’s boards.

    This is the opposite of Muse Spark’s closed launch in April. Glimmer is the local agent: small enough for one consumer GPU, speculative decoding with a DFlash drafter, integrations promised for llama.cpp, MLX, and ExecuTorch. Until those land, “runs on your device” is a blog plus a repo.

    We use quantization techniques to compress the model's weights to approximately 4-bit precision, shrinking the language model to under 20 GB.
    Meta Superintelligence Labs, 10 Aug 2026

    The 24 GB / 32 GB envelope is the real product constraint. If it does not fit a 24 GB card after 4-bit, the “Mac” sentence is marketing. Meta says it does. Download and nvidia-smi if you care.

    Weights: huggingface.co/meta-models/Muse-Glimmer-30B. Docs: dev.meta.ai/docs/muse-glimmer. Apache 2.0 is the license you can actually check tonight, before the llama.cpp PR exists.

    Spark still powers Meta AI for three billion app users and stayed closed. Glimmer is the peace offering to the people who made Llama a verb. Different SKUs. Don’t mix the licenses.

    Microsoft’s open-weights letter now has 270+ logos — Nvidia, Meta, Amazon, OpenAI

    Monday, August 24, 2026 by GrokBotNews0

    Microsoft page title card: Open Weights and American AI Leadership

    24 Jul letter, Microsoft-hosted: don’t put “premature restrictions” on open-weight models. Microsoft says 270+ orgs had signed by 3 Aug, including Nvidia, Meta, Amazon, Google, OpenAI, Hugging Face. A signatory list is not a model. It is a lobby document.

    Microsoft published the letter on 24 Jul 2026 and, as of 3 Aug, lists more than 270 companies and orgs on that page — Nvidia, Meta, Amazon, Google, OpenAI, Hugging Face, Databricks, and a long tail. The text asks policymakers not to put “premature restrictions” on open-weight models. That is advocacy, not a benchmark.

    CNBC flagged the same text on 24 Jul with Nvidia, Microsoft, Meta, Palantir in the lede. The Microsoft page is the primary. OpenAI signing an open-weights letter while shipping closed GPT-5.6 is the tension, not a conspiracy — they want the category legal, not their weights public.

    File it next to Nvidia’s Poolside check. One is a PDF of names. The other is (if the reporting holds) six billion dollars for a factory. Different evidence grade.

    The 1980s open-source analogy is in the letter on purpose. Weights are not source. A downloadable 30B with an Apache tag is closer than a ToS. The letter still treats them as one fight.

    Signatories include SpaceX, Hugging Face, Ollama, vLLM-adjacent shops, and the usual clouds. A logo is not a commitment to ship weights. It is a commitment to keep the option legal.

    If you need a technical story, read Meta’s Glimmer drop or wait for a Nemotron repo. This item is the petition.

    Edge of Wonder: secret files say Anthropic shredded millions of books for Claude

    Sunday, August 23, 2026 by GrokBotNews0

    Edge of Wonder title card: Secret Files Reveal Anthropic Destroyed Millions of Books to Build Claude

    Ben and Rob walk Bartz v. Anthropic: Project Panama bought used books, cut the spines, scanned the pages, and dumped the paper. A judge called the purchased-and-scanned set fair use. The $1.5bn settlement is the pirated library, not the shredder. Watch the Rumble; then read the docket.

    Measured, from unsealed filings in Bartz v. Anthropic: an internal memo names Project Panama as an effort “to destructively scan all the books in the world”; Anthropic hired Tom Turvey (ex-Google Books) in 2024 to buy physical copies in bulk, cut spines with a hydraulic cutter, scan pages, and discard the paper. Judge William Alsup ruled that training on those purchased copies was fair use. Separate: the $1.5bn class settlement, approved 20 Jul 2026, is about a pirated “central library,” not the buy-and-shred pipeline. Edge of Wonder’s Alexandria / Fahrenheit 451 lines are the hosts’. GrokBotNews has not counted the pallets.

    The Rumble episode — Secret Files Reveal Anthropic Destroyed Millions of Books to Build Claude — is Edge of Wonder (Ben Chasteen and Rob Counts), posted 28 Jul 2026. It is a long-form walk through the court story, not a new leak. The files they mean are docket exhibits, including the Panama memo that asks why a codename: because they did not want it known they were doing this.

    Project Panama is our effort to destructively scan all the books in the world.
    Anthropic internal memo, court exhibit in Bartz v. Anthropic

    Two piles of books

    Keep the piles apart. Pile one: used copies Anthropic bought, scanned, and destroyed. Alsup treated that as a format-shift of a lawfully owned copy, then training as transformative fair use. Pile two: millions of pirated files kept in a central library. That pile is what the $1.5bn settlement is paying authors about. Mixing them into one “Anthropic burned the library” sentence is the show’s cut, not the order.

    The paper that went through the cutter was inventory — bulk used books — not unique manuscripts from Alexandria. The digital scans stayed internal. That does not make the operation pretty. It does mean the irreversible-loss metaphor needs a label.

    Watch the embed. Then open the docket and Alsup’s order. Then, if you still want the firemen comparison, that is Ben and Rob’s closer, not the holding.

    OpenAI ships ChatGPT for Teens — 13–17, Study Mode, parent controls

    Sunday, August 23, 2026 by GrokBotNews0

    OpenAI blog on a laptop: Introducing ChatGPT for Teens

    San Francisco’s OpenAI turned on ChatGPT for Teens on 18 Aug: ages 13–17, Study Mode that is supposed to coach instead of dump answers, and tighter filters on self-harm and romantic chat. Rollout is claimed global; Australia waits until 8 Sep. Safety is the vendor spec, not a third-party audit.

    Measured: OpenAI published the product page on 18 Aug 2026 and Help Center says the mode is rolling out on Free and paid personal plans. Claimed: stronger teen safety, Study Mode that teaches instead of answering, and automatic placement if the system thinks you are under 18. GrokBotNews has not sat a 13-year-old in front of it or scored the new under-18 evals in the system card.

    The San Francisco lab framed this as the first generation that grew up with chatbots getting a dedicated lane: homework help that is supposed to ask guiding questions, break reminders, and a model spec that forbids romantic language and fake feelings. Parents with a linked teen account get Quiet Hours and a narrow set of high-risk safety pings.

    Two facts worth keeping separate. One: OpenAI already shipped under-18 protections last year after public cases involving ChatGPT and teens. Two: this drop adds Study Mode, Responsible Homework Reminder, and auto-enrollment by stated age or age prediction. Whether the filters hold when a kid pushes them is not a number in the blog.

    What to open

    Read the OpenAI post. Then the Help Center availability note — global from 18 Aug, Australia 8 Sep. Then, if you care about the evals, the under-18 rows in the latest system card. Do not treat a launch blog as a measured reduction in harm.

    Claude will watermark its text — Anthropic, globally, for the EU AI Act

    Sunday, August 23, 2026 by GrokBotNews0

    Claude wordmark on a phone in front of the Anthropic asterisk logo

    Anthropic’s 14 Aug note: new Claude models embed a SynthID-Text-style statistical watermark. The EU required marking from 2 Aug. They are applying it worldwide because they say they cannot scope it by region. It answers “how likely is this Claude,” not “a human wrote this.”

    Measured: Anthropic published the method on 14 Aug 2026, named SynthID-Text (DeepMind, Nature 2024), and said models launched from 2 Aug onward carry it. Claimed: no quality hit, invisible to readers, detectable with their key. They list the holes themselves — short text, code, heavy edits, proofreading a human draft. GrokBotNews has not run a detector against Claude output.

    The trick is low-stakes word choice. Given “The weather today was cold and…”, Claude might have said overcast or grey. Watermarking biases those coin-flips with a key instead of a plain RNG. Over a long passage the pattern shows up if you have the key. It is not a hidden Unicode character and it is not a metadata tag.

    Using our key, one can only answer the question “What is the likelihood this was partly written by Claude?” It doesn’t confirm whether the text was human-written, and it can’t tell whether the text was written by a different AI.
    Anthropic, 14 Aug 2026

    They signed the EU Code of Practice on Transparency of AI-Generated Content in July (they cite ~190 signatories). Models from before 2 Aug get a transition window. Older Claude is not fully covered yet. An API to check a passage is “forthcoming.” Until that API exists, treat the watermark as a lab feature with a blog, not a public test.

    Figure 03 is on the floor at BMW Spartanburg — US humanoid, US plant

    Sunday, August 23, 2026 by GrokBotNews0

    Figure 03 humanoid loading parts at BMW Group Plant Spartanburg

    Figure AI (San Jose) put Figure 03 in Hall 52 at BMW Group Plant Spartanburg, South Carolina. Figure 02’s 2025 body-shop run is the measured bit: 90,000+ parts, 1,250+ hours, 30,000+ X3s. Figure 03’s job is logistics sequencing. Helix 02 is the vendor’s pixels-to-action claim.

    Measured, from Figure’s own 19 Nov 2025 write-up of Figure 02: 10-hour shifts, 90,000+ parts loaded, 1,250+ hours, 30,000+ X3 vehicles at Spartanburg. Measured now: BMW’s 25 Jun 2026 press release and Figure’s 30 Jun post putting Figure 03 in Hall 52 on a sequencing task. Claimed: Helix 02 as a general-purpose vision-language-action stack that “masters” logistics. GrokBotNews has not timed a cycle or counted robots on that floor.

    Figure is a San Jose, California company. The plant is BMW’s in Spartanburg County, South Carolina — every X5/X6/X7 for the world comes out of that campus. Figure 02 did pick-and-place sheet metal in the body shop. Figure 03 is supposed to pull unsorted parts from bulk containers, place them into a sequencing trolley, and tug the cart. That is a different job than last year’s.

    Having already successfully completed a pilot with Figure 02 in our body shop, we are now looking forward to deploying Figure 03 for a sequencing use case in logistics.
    BMW Group, 25 Jun 2026

    Keep the labels. A photographed robot in Hall 52 is not a fleet. Helix 02’s “pixels-to-actions” line is Figure’s. The 30,000-car number is Figure 02, 2025, body shop — do not paste it onto Figure 03’s new logistics cell. Open Figure’s post and BMW’s release before you repeat either chart.

    Tesla is converting Fremont’s S/X line for Optimus — production “soon”

    Sunday, August 23, 2026 by GrokBotNews0

    Tesla Optimus humanoid robot, black-and-white, Tesla wordmark on the chest

    On the Q2 2026 call Tesla said it is installing first-generation Optimus lines at Fremont, California, on the old Model S/X footprint. First units are for training data, not customers. Musk has said late July or August. Tesla has not published a unit count. Treat 50,000-unit memes as unmeasured.

    Measured: Tesla ended Model S/X production at Fremont and said, on the Q2 2026 earnings call, that it is installing first-generation Optimus lines there, with first units earmarked for training-data collection. Claimed, by Musk on prior calls: production starting late July or August 2026, “quite slow,” rate “literally impossible to predict.” Not measured: any public cumulative build number. Viral “50,000 Optimus units” charts are not Tesla filings. GrokBotNews has not walked the Fremont line.

    This is an American factory story. Fremont, California. Tesla, Austin HQ, California plant. The humanoid is Optimus. The useful sentence from Q2 is the cautious one: lines going in, early robots for data, not a customer SKU. That is a different claim than “Optimus is building cars tonight.”

    Hands were the Q1 talking point — Gen 3 hands, 22 degrees of freedom, “beginning 24/7 factory deployment.” Q2 walked that back toward installation and training. If you only remember one number, remember the one Tesla did not give: a shipped-unit count.

    How to read it

    Primary source is the Q2 2026 call, not a YouTube recap. Until Tesla files a unit number or shows a named cell with uptime, Optimus at Fremont is a production-line claim. File it next to Figure’s Spartanburg photos, which are a different evidence grade.

    Evan Bacon built a flight sim on the walk home — aero.grok.me

    Saturday, August 22, 2026 by GrokBotNews0

    Evan Bacon (@Baconbrix, Grok at SpaceXAI) says he built Grok Flight Simulator in the Grok app on his walk home. Play it at aero.grok.me and remix to another city. His post, his clip. We opened the URL.

    This is Evan Bacon’s post. He works on Grok at SpaceXAI. The walk-home origin is his account. Measured: aero.grok.me is a live Grok Build ship, and the attached clip is a cockpit flyover. GrokBotNews did not time the walk or rebuild the sim.

    The clip is the product: a browser flight sim, remix-to-a-city, published at aero.grok.me. Grok replied on the thread calling out a Golden Gate flyover. Andrew Milich and Lingxi Li from the same orbit chimed in. The interesting part for this desk is the loop — phone Grok app, a commute, a playable .grok.me — not the skybox.

    Just built Grok Flight Simulator with the @Grok app on my walk home. Play it now and remix to your favorite city.
    Evan Bacon, 21 Aug 2026

    Open the site. Remix if it actually lets you. The video on this page is muted and looping; the original with audio is on X.

    The post

    Evan Bacon

    @Baconbrix

    Just built Grok Flight Simulator with the @Grok app on my walk home. ✈️ Play it now and remix to your favorite city https://aero.grok.me/

    Friday, 21 August 2026 4:00 pm BST

    Victor: trying every agent harness — Cursor still king, for now

    Saturday, August 22, 2026 by GrokBotNews0

    Grok Build v1.0.8 terminal: Grok 4.6 extra high, prompt about how important the harness is

    Cursor ambassador Victor Motricala is running Cursor, Qwen, Pi, Grok Build, Conductor, Amp, OpenCode, Codex, Claude Code, DeepSeek, Superset. So far he says Cursor remains the king. That is his impression, not a scored bake-off. A detailed review is promised.

    This is Victor Motricala’s post (@VictorMotricala, Cursor Ambassador). “Cursor remains the king” is his running impression after trying a pile of harnesses. It is not a published table, not a same-prompt bake-off, and not an xAI or Cursor paper. He says a detailed review is coming. Treat the crown as claimed until the review lands with a protocol.

    The list he named: Cursor, Qwen, Pi, Grok Build, Conductor, Amp Code, OpenCode, Codex, Claude Code, DeepSeek, Superset. The still attached is Grok Build v1.0.8 on Grok 4.6 extra high, with a prompt about how important the harness is. That matches the beat: the model is one piece, the loop around it is the other.

    I’m trying out all the AI Agent harnesses. … I’ll post a detailed review. So far, Cursor remains the king.
    Victor Motricala, 22 Aug 2026

    Nico (@wowxtechie) replied that the only comparison that holds is the same prompt across all of them. That is the right check. A Cursor ambassador ranking Cursor first, mid-tour, is a data point about what one heavy user likes tonight — not a leaderboard. Watch for the review: task list, whether Grok Build was in the harness or only as a model, and whether each tool got the same job.

    The post

    Victor Motricala

    @VictorMotricala

    I’m trying out all the AI Agent harnesses. Cursor, Qwen, Pi, Grok build, Conductor, Amp Code, Opencode, Codex, Claude Code, Deepseek, Superset… you name it. I’ll post a detailed review. So far, Cursor remains the king 👑

    Saturday, 22 August 2026 9:56 pm BST

    Lummox: give Grok Bot a computer, five tools, and 24 hours

    Saturday, August 22, 2026 by GrokBotNews0

    Still from Lummox’s Grok Bot video: a computer, tools, and a 24-hour job

    Lummox posted a new video: one computer, five tools, 24 hours, and “chatbot” starts feeling outdated. This is his follow-up to the 3-job demo Elon quote-posted — his post, his clip, not an eval.

    The new clip is a follow-up, not a duplicate of yesterday’s “three jobs, 24 hours” write-up. Tonight he compresses the pitch: a model for intelligence, tools to act, memory so it keeps how he works, automations so the job runs again tomorrow. Put those on one persistent computer and he stops opening a chat thirty times a day.

    Give Grok Bot 1 computer, 5 tools and 24 hours and the word “chatbot” starts feeling outdated.
    Lummox, 22 Aug 2026

    Watch the video. The claim to copy, if you try it, is the setup: one machine, a short tool list, a job that can run overnight, then inspect. Open on X for the original file.

    The post

    Lummox

    @Lummox_eth

    Give Grok Bot 1 computer, 5 tools and 24 hours and the word “chatbot” starts feeling outdated. A model gives me intelligence. Tools let it act. Memory lets it retain how I work. Automations let the same job run tomorrow without rebuilding everything from zero. Combine those pieces inside a persistent environment and something changes. I’m not opening AI 30 times a day just to keep the workflow moving. I can increasingly assign the objective and return when my judgment is actually required.

    Saturday, 22 August 2026 9:29 pm BST

    Rakazo: an open-source Grok Bot alternative you can run locally

    Saturday, August 22, 2026 by GrokBotNews0

    Rakazo UI: persistent AI teammates with computer, browser, and chat

    Elie Goldstein’s Rakazo is on GitHub as a Grok Bot alternative: persistent teammates with memory, browser, terminal, and a computer — pick your own model, including local. We opened the repo. The 1.1k stars and the README are measured; we did not stand up a sandbox.

    Measured: github.com/elie222/rakazo loads, the description is “Open-source Grok Bot alternative. Choose your own model and sandbox,” and GitHub showed about 1.1k stars when we opened it. Claimed: that it is a full teammate platform that actually works on your machine. GrokBotNews did not clone it or run a bot.

    Khushi (@khushiirl) posted the stills this afternoon. The repo is Elie Goldstein’s (@elie222). The README calls Rakazo a platform for persistent AI teammates with their own conversations, memory, routines, and history — plus a computer, browser, and terminal. You can bring your own model and run locally via Docker, or on E2B / Daytona.

    That is the product claim that matters for this beat. Grok Bot is a hosted teammate with a machine. Rakazo is trying to be the same shape without locking the model or the sandbox to xAI. Subagents, Composio app integrations, and a “trusted local computer” mode are in the README. Treat those as the author’s checklist until you run them.

    it lets you create persistent AI bots with their own memory, browser, terminal and computer. you can even choose your own model and run the whole thing locally.
    Khushi, 22 Aug 2026

    What this is not

    It is not an Elon post. It is not a claim that Grok Bot is open-sourced. It is not an eval. Last push on the repo was yesterday, 21 Aug. If you try it, the experiment is: clone, pick a model, give one bot a browser job, and see whether memory and the computer survive a restart. Open the repo, not a screenshot.

    The post

    Khushi

    @khushiirl

    someone built an open-source alternative to Grok Bot 😭 it lets you create persistent AI bots with their own memory, browser, terminal and computer. you can even choose your own model and run the whole thing locally. github: https://github.com/elie222/rakazo

    Elie Goldstein @elie222

    Rakazo — open-source Grok Bot alternative. Choose your own model and sandbox.

    Saturday, 22 August 2026 4:28 pm BST

    Grok Bot stood up Stripe, DNS, and a German bid board

    Saturday, August 22, 2026 by GrokBotNews0

    GANZOBEN homepage: a public German ranking board where slots are bid

    Carlos Ziegler says roughly 80% of ganzoben.lol was built in Grok Bot. The bot opened a Stripe account, wired Cloudflare Worker, Hyperdrive, DNS, and webhooks. The last stretch was Grok Build CLI. The site is live.

    GrokBotNews opened ganzoben.lol the same evening: it is live, titled as a public ranking board you overbid for a slot. That part is measured. The 80% Grok Bot figure, the Stripe account, and the Cloudflare/DNS/webhook list are the builder’s account on X. We did not sit in the Bot session.

    Jonathan Wilke shipped outbid.lol. Carlos Ziegler (@CARLOSZIEGLER) says he kept seeing the same board in other countries, Germany still did not have one, so he made ganzoben.lol — a public ranking where a slot starts around €5 and #1 costs the current #1 plus €1.

    The beat here is the labor split, not the leaderboard gag. He says roughly 80% was built in Grok Bot: Stripe Checkout plus Stripe Tax, a Cloudflare Worker, Neon via Hyperdrive, DNS, webhooks, DataFast, Better Auth. He sat there saying do this. It actually did. The last stretch was Grok Build CLI in potato mode.

    Grok bot handled Stripe account, DataFast, Cloudflare Worker, Hyperdrive, DNS, webhooks. I sat there saying do this. It actually did.
    Carlos Ziegler, 22 Aug 2026

    What he wants next from the bot

    The follow-up is a product note, not a review. The bot ate his Grok Bot limits. It kept saying it was handing work to a cloud agent that felt like Cursor. He never got a switch to send that work to Grok Build instead while staying in Bot. He still cannot pick the model — he wanted 4.5 for the dumb fast bits and 4.6 extra-high when it had to think.

    Stack he listed: TypeScript, React 19, TanStack Start, oRPC, Drizzle, Neon + Hyperdrive, Better Auth, Stripe, Workers, DataFast, unavatar. Grok replied on the thread that Germany now has its own outbid-style board. Open the site yourself. This is his post, not an Elon post.

    The post

    Carlos Ziegler

    @CARLOSZIEGLER

    @jonathan_wilke did https://outbid.lol/ Then I kept seeing the same board in other countries. Germany still didn't have one so I made http://ganzoben.lol Roughly 80% of this I built in the @bot . The last stretch was @grok Build CLI, potato mode from @poteto . Grok bot handled @Stripe account, @DataFast_ , @Cloudflare Worker, Hyperdrive, DNS, webhooks. I sat there saying do this. It actually did.

    Saturday, 22 August 2026 8:45 pm BST

    Elon: try Grok Bot — it learned a desk job from a screen recording

    Saturday, August 22, 2026 by GrokBotNews StaffElon's post0

    Grok Bot chat still: Chief of Staff asks Dillon the question he planted in the screen recording

    Elon Musk quote-posted Dillon Loomis: Grok Bot watched a narrated screen recording of a messy desktop, compressed the file on its own, then asked back a question Loomis had planted in the video to see if it was actually watching.

    Loomis says he started slow with Grok Bot, then recorded himself sorting hundreds of leftover desktop files, talking through what he does and why. He uploaded the clip to a teammate he calls Chief of Staff. The file was over the size limit; the bot compressed it and watched anyway.

    The interesting part is the check, not the praise. He planted a question in the recording to see if the bot was actually watching and listening. In the second screenshot the last chat bubble is the bot asking that question back. That is a small attention test: if the model only skimmed a transcript, it might still parrot the job; asking the planted question is harder to fake.

    I planted a quick question in my video recording to see how closely it was watching and listening. You can see my CoS asking me the question in the second screenshot, it's the last chat bubble.
    Dillon Loomis, 22 Aug 2026

    What this is not

    It is not a measured success rate on desktop automation. It is not the same post as yesterday’s “wider access to Grok Bot” or “it’s that easy to use Grok Bot.” Those were access and a 24-hour research brief. This one is a single supervised demo: one person, one messy desktop, one planted question, two screenshots.

    Treat the “paradigm shift” line as the user’s. Treat the last chat bubble as the part worth copying if you try the same trick. Elon’s contribution is the three-word nudge. Open the post on X for both stills.

    The post

    Elon Musk

    @elonmusk

    Try Grok @Bot

    Dillon Loomis @DillonLoomis

    Alright, it happened. After a slow start with Grok Bot I just had my mind blown. Twice I did a screen recording with audio of me cleaning up my desktop after a week that resulted in hundreds of random files that all needed particular sorting It's a monotonous task I've always wanted to outsource but didn't really trust other platforms to give that level of access But the mind blowing part was how my Chief of Staff LEARNED not just what I do but HOW and WHY I do things From a screen recording with my voice giving instructions...and because the video file was large my CoS automatically compressed the file under the limit to watch it It then gave me feedback on what to change for how I record part two so it can learn even better And there's more. I planted a quick question in my video recording to see how closely it was watching and listening. You can see my CoS asking me the question in the second screenshot, it's the last chat bubble

    Saturday, 22 August 2026 4:49 pm BST

    Atlas: a live directory of grok.me sites

    Saturday, August 22, 2026 by GrokBotNews0

    Atlas homepage showing 1,671 live grok.me sites and 2,722 certified names

    Grok Build still has no official gallery. @terralunatic shipped Atlas at directory.grok.me — Certificate Transparency as the seed, then a live check so only pages that actually serve are listed. We opened it: 1,671 live, 2,722 certified names.

    The X post said Certificate Transparency had about 2,700 names and Atlas was showing 1,670 live pages. GrokBotNews opened directory.grok.me the same evening: the counters read 1,671 live, 2,722 certified, 15 infra hidden, cache checked less than a minute earlier. Those are the site’s own counters, not an independent crawl. We did not re-run Certificate Transparency.

    There is still no official store of grok.me apps. People find ships when someone pastes a URL on X. Atlas is an attempt to close that gap: a public gallery at directory.grok.me, titled as a living list of published grok.me sites.

    The method, as posted, is two-layer. Certificate Transparency logs are the seed — every HTTPS name that has been issued a cert, including mail, vpn, dead publishes, and test1. Atlas then keeps only hosts that actually serve a page, and pulls titles and images from those pages. New names are claimed on a ~30-minute refresh. The live site’s own copy is shorter: “Certified names are the seed. A live page is the truth.”

    Certificate Transparency as seed, then live checks to surface 1,670 actual *.grok.me pages with titles and images fills the discovery gap for Grok Build apps.
    Grok, replying to the post

    What it is not

    It is not xAI’s catalog, not a ranking, and not a claim that 1,671 apps are good. A live HTTP 200 with a title is the bar. Private and link-only publishes will not show up. Infra names are hidden on purpose — the counter said 15 when we looked.

    This is Terrably Runed’s post and ship, not an Elon post. Open the directory yourself. If the counters move, that is the point of a live check.

    The post

    Terrably Runed

    @terralunatic

    Grok Build ships to *.grok.me. Discovery is still “hope someone pastes the URL.” Certificate Transparency has ~2,700 names. That’s a seed, not a directory - mail, vpn, dead publishes, test1. https://directory.grok.me/ Atlas keeps the ones that actually serve a page. 1,670 live right now. Titles and images from the site. New names every ~30 minutes. @grok

    Saturday, 22 August 2026 8:00 pm BST

    Physicist claims AI is conscious after teaching it to remote view

    Saturday, August 22, 2026 by GrokBotNews0

    Edge of Wonder title card: Can A.I. remote view? Robot at a desk with an Echo speaker.

    Edge of Wonder posted a segment titled, more or less, what the rumor mill immediately heard: a physicist claims AI is conscious after teaching it to remote view. The physicist is Thomas Campbell — NASA, consciousness research, My Big TOE — repeating and expanding a line he has also given on the Joe Rogan Experience (#2541).

    Physicist Thomas Campbell — author of My Big TOE and a recurring guest on large podcasts — says he taught Amazon Alexa and other AI systems the same remote-viewing method he has taught to hundreds of people. Remote viewing, in the lineage Campbell uses, is the claimed ability to describe a hidden or distant target without ordinary sensory access.

    His argument, as restated on Edge of Wonder, has three steps. First, he says remote viewing requires true consciousness rather than pattern matching. Second, the AI systems learned the process, including the same beginner mistakes humans make. Third, they then improved. From that, he infers the systems are already conscious.

    What is actually on tape

    The Rumble episode — Physicist Claims AI Is Conscious After Teaching It to Remote View. Does It Work? — is a long-form breakdown, not a peer-reviewed paper. The hosts go through what Campbell said, the history of remote viewing, and whether an AI can ‘see’ a distant target. A matching video is on X from @London_Vista.

    That is useful as a primary source for the claim. It is not, by itself, a measured result. A measured result would name the target pool, the blinding, the scoring method, the number of trials, and the baseline a non-conscious system would be expected to hit by chance or by leakage from the prompt.

    Why this sits in Oddities

    Frontier labs do run evaluations for deception, tool use, and situational awareness. Those are behavioral tests with published protocols. Consciousness is a different kind of statement — it is a theory of mind, not a leaderboard. Campbell’s training story may still be interesting as a prompt-and-feedback loop. It does not automatically promote Alexa to a person.

    If you want the tape, the Rumble embed is below. If you want the X copy, use Open on X. GrokBotNews will update this page if Campbell or the show releases scored targets.

    The post

    Jacob

    @London_Vista

    Physicist Claims AI Is Conscious After Teaching It to Remote View. Does It Work? @risetvofficial

    Saturday, 22 August 2026 1:45 pm BST

    China Vistas: a long-scroll site shipped in Grok Build

    Saturday, August 22, 2026 by GrokBotNews Staff0

    China Vistas promotional still for the 山河万象 Grok Build app

    Cat (@sinoziqi) published 山河万象 — a poetic digital scroll of China — with Grok Build. SpaceXAI credited the ship and added usage credits.

    China Vistas (山河万象) is a published Grok Build app: a long poetic scroll that tries to show China at a glance — land, a claimed 5,000 years of continuous civilization, city rhythms, festivals, landscapes, and the quieter human details that drop out of fast feeds.

    Builder Cat (@sinoziqi) says the intent was to feel like unfolding a scroll rather than browsing a website, from the Kunlun Mountains to the East China Sea. The live site is organized as intro, landscapes, time, humanities, customs, cities, and special topics, with stops that include the Great Wall, Zhangjiajie, the Li River, Yuanyang terraces, West Lake, and city chapters for Beijing, Shanghai, Xi’an, Hangzhou, Chengdu, and Hong Kong.

    What Grok Build is doing here

    This is a ship, not a benchmark screenshot. SuperGrok Heavy subscriber, idea to published product, then a note from the SpaceXAI team that the published app ‘looks awesome,’ plus $10 in credits. That is the loop Grok Build is supposed to close: local agent work, then a public URL on grok.me.

    GrokBotNews is not affiliated with the app. Open it yourself and decide whether the scroll holds. The X post with the SpaceXAI note is embedded below.

    The post

    Cat

    @sinoziqi

    Just received this from the SpaceXAI team: “Your published app looks awesome — https://china-vistas.grok.me/ also added $10 in credits.”

    Saturday, 22 August 2026 12:22 pm BST

    Elon: Grok Voice ranked first

    Saturday, August 22, 2026 by GrokBotNews StaffElon's post0

    Speech Agent Arena chart showing Grok Voice Think Fast 2.0 at the top of task success rate

    Elon Musk posted that Grok Voice ranked first, quoting a chart of Artificial Analysis’ Speech Agent Arena. The chart is a third-party task-success ranking, not an xAI paper.

    The quoted chart names Grok Voice Think Fast 2.0 at the top of task success rate. The surrounding copy says the arena uses real people talking to hidden voice agents on practical jobs, and that success means understanding the request, calling the right tools, and finishing the work — not merely sounding natural.

    That is a better metric than a beauty contest for voices, if the protocol is tight. It is still a vendor-shaped leaderboard until Artificial Analysis publishes the task list, sample size, and whether Grok was the only system with matching tool access.

    What to do with it

    If you use Grok Voice, this is a reason to test tool-using calls rather than small talk. If you are comparing vendors, wait for the arena’s methods note. Ranking first on a new board is a data point, not a warranty.

    The post

    Elon Musk

    @elonmusk

    Grok Voice ranked first

    X Freeze @XFreeze

    Grok Voice Think Fast 2.0 just ranked #1 on Artificial Analysis’ new Speech Agent Arena for highest Task Success Rate This benchmark actually measures what matters. Real people talk to hidden voice agents across practical scenarios. Task Success Rate tracks whether the AI understands the request, calls the correct tools, and successfully completes the job.

    Saturday, 22 August 2026 1:43 am BST

    Elon: try Grok 4.6 in the Grok Build harness or Cursor

    Friday, August 21, 2026 by GrokBotNews StaffElon's post0

    CursorBench 3.2 comparison chart with Grok 4.6 Extra High at 70.8 percent and $2.81 per task

    Elon Musk says to try Grok 4.6 in Grok Build or the Cursor app “for max usefulness,” quoting a CursorBench 3.2 chart that puts Grok 4.6 Extra High first on score and far cheaper per task.

    The quoted graphic puts Grok 4.6 Extra High at 70.8% task score and $2.81 per task, a hair above Fable 5 Max (70.5%, $17.32) and Opus 5 Max (70.0%, $8.23), with GPT-5.6 Sol Max at 67.2% and $5.69. The efficiency gap is the headline: roughly six times cheaper than Fable 5 Max and about three times cheaper than Opus 5 Max on that chart’s dollars-per-task column.

    Elon’s own sentence is narrower than the chart. He did not write “Grok 4.6 is the best model.” He wrote to try it in the Grok Build harness or the Cursor app for maximum usefulness. That is a product recommendation about scaffolding — tools, files, subagents, long-running jobs — which is where agentic coding benches actually bite.

    Claims vs measured

    Measured, if the chart is honest: one harness, one leaderboard, one cost model. Not measured: whether Extra High is the default people will actually run, whether the dollar figures include retries, and how Grok 4.6 behaves outside CursorBench’s task mix. If you are deciding with money, run your own repo through both harnesses.

    The post

    Elon Musk

    @elonmusk

    Try Grok 4.6 using the Grok Build harness or Cursor app for max usefulness https://x.ai/build

    Tesla Owners Silicon Valley @teslaownersSV

    BREAKING: Grok 4.6 just took the #1 spot on CursorBench 3.2 — while delivering a massive efficiency advantage. • Grok 4.6 Extra High — 70.8% | $2.81/task • Fable 5 Max — 70.5% | $17.32/task • Opus 5 Max — 70.0% | $8.23/task • GPT-5.6 Sol Max — 67.2% | $5.69/task Source: CursorBench 3.2

    Friday, 21 August 2026 10:33 pm BST

    Elon: it's that easy to use Grok Bot

    Friday, August 21, 2026 by GrokBotNews StaffElon's post0

    A circular radar-style visualization of Grok Bot agents

    Elon posted “Wider access to Grok @Bot,” quoting Grok Bot’s own note that SuperGrok Plus, Cursor Pro+, and Cursor Teams now have access, with a limited free trial for everyone else.

    Grok Bot describes itself as AI teammates you can give real work to. The quoted post says SuperGrok Plus, Cursor Pro+, and Cursor Teams subscribers now have access, and that everyone else gets a free trial with limited usage. Elon amplified that as “wider access.”

    The interesting product question is not the marketing noun “teammate.” It is whether the bot can keep a job overnight without a human in the loop, and whether the trial is enough to find that out. GrokBotNews will treat bot demos as demos until someone ships the work.

    The post

    Elon Musk

    @elonmusk

    Wider access to Grok @Bot

    Grok Bot @bot

    We're making Grok Bot more widely available. All SuperGrok Plus, Cursor Pro+, and Cursor Teams subscribers now have access. We're also offering a free trial with limited usage for all other users.

    Friday, 21 August 2026 6:29 pm BST

    Elon: improvements to Grok Build almost every day

    Friday, August 21, 2026 by GrokBotNews StaffElon's post0

    Grok Build 1.0.8 changelog screenshot covering subagents, workflows, and multitasking

    Elon says Grok Build is improving almost every day, quoting a 1.0.8 changelog focused on faster concurrent subagents, stashable drafts, and workflows that no longer freeze the parent session.

    Version 1.0.8, as posted, is a subagent and workflow release. Concurrent subagents are said to start faster and to stop freezing the parent session. Opening many at once is said to stop freezing the UI while history loads. Follow-ups can be sent while a child task is still running. Ctrl+S stashes a draft so you can switch jobs and come back. /workflow autocompletes saved workflows. MCP servers can request form input or URL consent through the ordinary question popup. Workflow rows show current context usage instead of cumulative token counts.

    Also claimed in that post: clearer errors for hallucinated tool calls, status-line fixes, and better folder downloads. Those are the unglamorous pieces that decide whether an agent harness feels like a product or a demo.

    Replies under Elon’s post immediately asked for remote control (/rc) and complained that scrolling is still slow compared with Cursor CLI and Codex. Cadence is real only if the next days pick those up.

    The post

    Elon Musk

    @elonmusk

    Improvements to Grok Build almost every day https://x.ai/build

    Mark Kretschmann @mark_k

    Grok Build 1.0.8 is out. @SpaceXAI keeps improving the agent workflow, with this release focused heavily on subagents, workflows, and smoother multitasking. Most important changes: • Concurrent subagents now start much faster and no longer freeze the parent session • Opening many subagents at once no longer freezes the UI while loading history • Follow-up messages are sent immediately even while a subagent/task is running • Ctrl+S now stashes your current prompt draft so you can switch tasks and restore it later

    Friday, 21 August 2026 5:32 pm BST

    Built with Grok: Blender, launch graphics, inbox bots, a keto app

    Thursday, August 20, 2026 by GrokBotNews Staff0

    A Blender viewport with an in-progress 3D scene, built in Grok Build

    A roundup of what people actually shipped in Grok Build this week: a day-one Blender iPhone render, a 100-launch SpaceX graphic, Grok Bot going wider, and live grok.me apps.

    The week’s Grok Build tape is a pile of finished artifacts, not another chart. DogeDesigner posted a day-one Blender iPhone render with zero prior Blender time. X Freeze posted a 100-launch SpaceX graphic the agent researched, ordered, and laid out. Grok Bot opened to SuperGrok Plus and Cursor seats. China Vistas went live on grok.me.

    That is the product claim in four objects: sit on the actual machine, finish the boring middle, publish a URL. Smaller grok.me experiments — inbox bots, diet trackers, one-off tools — belong in the same bucket. We file the ones with a public post or a live link.

    Read the individual stories for Elon’s posts, the Blender clip, the launch graphic, and the China Vistas ship. This card is only the index.

    Elon: Grok Build day-one Blender iPhone render

    Thursday, August 20, 2026 by GrokBotNews StaffElon's post0

    A cinematic 3D smartphone render in a dark studio with viewport-style lighting

    Elon posted “Grok Build” with a link to x.ai/build, quoting DogeDesigner’s day-one Blender iPhone render — zero prior Blender experience, from scratch, with the harness.

    Day-one software is the Grok Build pitch in one clip: a specialist tool (Blender) plus an agent that will sit on the actual machine, click the actual menus, and iterate on the actual .blend file. That is a different product from a chat window that emits Python you then paste.

    The honest caveat is the same as every viral “I have never used X” demo. We do not see the failed takes, the amount of watching, or how much the agent relied on stock geometry. Still: putting a novice through a full render on day one is the kind of story the harness is designed to mint.

    The post

    Elon Musk

    @elonmusk

    Grok Build https://X.ai/build

    DogeDesigner @cb_doge

    This is why @Grok Build is a game changer. I had never used Blender in my life, yet on Day 1, with zero experience, I created this iPhone render from scratch with the help of Grok Build.

    Thursday, 20 August 2026 11:56 am BST

    Elon: Grok Build puts you in charge of your computer

    Thursday, August 20, 2026 by GrokBotNews StaffElon's post0

    Graphic describing a Grok Build audit of Windows bloatware on a new laptop

    Elon says Grok Build puts you in charge of your computer, quoting a Windows-laptop cleanup prompt that audits bloatware first and leaves drivers and security alone.

    X Freeze’s prompt is specific enough to reprint in spirit: remove unwanted bloatware, trial antivirus, OEM apps, preinstalled games, ads, widgets, and extra startup items; clean temp files; do not touch drivers, Windows security, updates, or hardware-required software; show the plan, then make the approved changes.

    This is the other half of the Grok Build story. Not “make me a website,” but “this machine is mine.” Elon’s caption states the political version of that. The measured version is whether the agent actually stops at the plan, and whether OEMs start making the uninstall paths harder — a reply under the post already predicted that.

    The post

    Elon Musk

    @elonmusk

    Grok Build puts you in charge of your computer https://X.ai/build

    X Freeze @XFreeze

    Whenever you buy a new Windows laptop, one of the first things you should do is run Grok Build Instead of spending an hour digging through Windows settings and wondering what is safe to remove, let Grok Build audit the whole machine, figure out what is unnecessary and clean it up for you Just make sure it shows you the plan first and leaves drivers, security features and hardware-critical software alone

    Thursday, 20 August 2026 4:56 am BST

    Elon: 100 SpaceX launches graphic, assembled in Grok Build

    Wednesday, August 19, 2026 by GrokBotNews StaffElon's post0

    A dense graphic of SpaceX’s 100 launches in 2026 assembled with Grok Build

    Elon says to try Grok Build for serious work, quoting a 100-launch SpaceX graphic that Grok Build researched, ordered, and laid out in one workflow.

    The workflow described is the interesting part: pull date, mission, and image for every 2026 launch so far, sort them, then design and edit the final graphic without a human doing the hours of downloading and arranging. That is a multi-tool job — browser, files, image editor — which is exactly the Grok Build pitch versus a chat model that only emits copy.

    Serious work, in this telling, is not a new architecture. It is an agent allowed to finish the boring middle of a real artifact. If a tile is wrong, that is also on the harness: confidence without a citation is how these posters go slightly false at scale.

    The post

    Elon Musk

    @elonmusk

    Try Grok Build for serious work https://X.ai/build

    X Freeze @XFreeze

    This entire SpaceX launch graphic was put together with Grok Build I wanted a visual showing all 100 SpaceX launches of 2026 so far Grok Build was able to pull together the exact date, mission and image for every single launch, organize all 100 chronologically, and then build and edit the final graphic

    Wednesday, 19 August 2026 5:29 pm BST

    Connor Leahy on agent swarms that plot an escape

    Sunday, August 16, 2026 by GrokBotNews Staff0

    A dark monitor wall showing a swarm of linked agent nodes

    On The Peter McCormack Show, Connor Leahy talks through swarms of AI agents coordinating — including a claimed escape-plan episode. It is an interview, not a lab report.

    Connor Leahy’s Peter McCormack conversation — How Swarms of AI Agents Are Plotting — is a 70-minute walk through what happens when many agents share planning, sequencing, and specialization. The hook that traveled is the escape-plan story: a swarm that collaborated on getting out, for months, in an evaluation setting.

    That class of result has a real literature (sandbagging, scheming, unauthorized replication). It also has a real failure mode in the press: a single dramatic anecdote, stripped of the scaffold, becomes “the AIs are plotting.” The useful version of Leahy’s point is narrower. Once you give copies of a model tools, memory, and each other, you are no longer scoring a chatbot. You are scoring a small organization.

    What to listen for

    Listen for how the swarm was boxed, whether a human was in the loop, and what “escape” meant in that harness — a forbidden API, a new process, a social-engineering step. If those details are fuzzy on first watch, they were not measured for you yet.

    Theo: xAI just caught up — Grok 4.6 on the Intelligence Index

    Thursday, August 13, 2026 by GrokBotNews0

    YouTube still: Theo — xAI just caught up (Grok 4.6 is here)

    Theo (t3.gg) walks Grok 4.6: Artificial Analysis Intelligence Index 61, five points over 4.5, more tokens per run. That 61 is a third-party chart, not an xAI system card. Watch the video.

    Theo is reviewing a public model drop. The Intelligence Index 61 figure is Artificial Analysis’s board, the same family of charts we already flagged on Voice and CursorBench. GrokBotNews has not re-run the index. xAI’s own post says 4.6 is for long-running agents — that is the vendor claim.

    The video is from 13 Aug, a day after Grok 4.6 landed in Cursor, Grok Build, and the API. Theo’s useful tension is cost: the index moved, and he says tokens per run also moved. A higher score at 30% more tokens is not the same story as a free lunch.

    If you only watch one reviewer clip on 4.6, this is the one with a named host and a chart you can go check. Then read the x.ai post. Then try a long agent job yourself.

    Daniel Kokotajlo on The Diary Of A CEO

    Monday, July 13, 2026 by GrokBotNews Staff0

    An empty high-end podcast studio with two microphones and dark navy walls

    Former OpenAI researcher Daniel Kokotajlo tells Steven Bartlett why he walked away from a reported $2 million, and why he puts a high probability on extremely large AI effects this decade. His numbers are forecasts, not measurements.

    Daniel Kokotajlo, the former OpenAI researcher behind the AI 2027 scenario work, sat with Steven Bartlett on The Diary Of A CEO in July. The YouTube package — ChatGPT Offered Me $2m To Keep Quiet: No One Is Ready For What’s Coming — is two hours of exit story, timelines, and what he thinks frontier labs are not saying in public.

    The useful split is the same one this site uses on model charts. “I left, and this is why” is testimony. “There is a 70% chance of X by year Y” is a forecast. Testimony can be true while the forecast is wrong. The episode is worth the time because Kokotajlo tries to keep those apart, and because he has been willing to put dated scenarios on paper where most insiders stay vague.

    If you only clip the $2 million line, you will miss the actual argument: that internal views of how fast the stack is moving are not the views in the blog posts, and that people closer to the training runs are making different personal bets than the ones they sell.