You know the part of onboarding a new hire where nothing can happen until IT gives them a laptop, a badge, and the WiFi password? Grok Bot skips it. You name a bot, and it already has a computer with a browser open and your logins sitting in it.
That is the pitch, and I want to be fair to it: it works. Early users describe it as the first agent product that felt like adding a coworker instead of configuring one.
But the sentence doing the most work on the product page is wrong about the product.
The verdict, up front. If you already pay for SuperGrok Heavy or Cursor Ultra, open it tonight and give one bot one small job you can inspect afterward. If you don’t already pay for those, don’t upgrade for this. It’s an early beta with no published task-completion rate, no audit log, and a security model that the marketing page and the documentation describe differently. The gap between those two descriptions is the story.
The mental model: you hired a team, you rented a desk
Here’s the frame I’d use if you’re explaining this to your boss.
Picture an office with one desk. One chair, one monitor, one browser, one drawer of sticky notes with passwords on them. You hire eight people. They all sit at that desk, one at a time, and whatever the last person left signed in stays signed in for the next one.
That is Grok Bot’s architecture. Not eight computers. One.
Call it the Shared Desk, because once you have that image, every strange detail in the docs stops being strange.
Why can your research bot hand a file to your writing bot without you copy-pasting between chats? Same desk. Why does xAI’s own security page warn you not to treat separate bots as a security boundary? Same desk. Why does deleting a bot leave its logins alive? The person left. The desk kept the sticky notes.
The shared desk is simultaneously the best feature and the biggest risk, and those are not two facts. They’re one fact wearing two hats.
Claims vs. receipts
Claim: “Bots have their own computer.”
That line appears in the hero copy and the FAQ on x.ai/bot, verified live August 19, 2026. Now open xAI’s own launch post from August 11 and read the section titled “A computer of its own.” The sentence inside it says bots share a computer of their own in the cloud.
Two different claims, from the same company, on the same day. The documentation is unambiguous about which one is true: every bot on your account uses one persistent cloud computer, assigned per user, not per bot. And then it adds the sentence that should be on a poster in every security team’s office: do not use separate bots as a security boundary.
A good chunk of launch coverage picked up the wrong version. At least two outlets published that each bot gets a dedicated cloud computer. That is not a small copyedit. It’s the difference between “I can contain the blast radius by splitting bots” and “I cannot.”
Claim: “Get started for free.”
That’s the text on the final call-to-action button on the product page, verified today. It links to a .dmg download. Every plan listed above it on the same page costs money: Cursor Ultra at $200/month, SuperGrok Heavy at $300/month, Cursor Premium Teams at $120 per seat per month. The docs say eligible plans are required. One secondary write-up mentions a one-time individual trial, but nothing on the pricing cards describes it, so I’d label that access path unannounced rather than free.
Claim: it runs on Linux.
The docs, dated August 11 and unchanged since, list macOS, Windows, and iPhone on iOS 18 or later, then state plainly that Linux desktop, Android, and iPad are not supported at initial launch. Several outlets reported Linux builds shipped anyway, and at least one reviewer screenshotted a “Download for Linux” button on the marketing page. I could not resolve this one. The live page today shows a macOS button and a link labeled “other platforms and devices.” I’m reporting the conflict rather than picking a side.
Claim: it’s frontier-grade.
This one mostly checks out, with a version-number trap. Grok 4.6 shipped August 12 and scores 60.9 on the Artificial Analysis Intelligence Index, which rounds to the 61 xAI claimed. Independent measurement matched the vendor’s number, which is worth saying out loud because it often doesn’t. It sits behind Claude Opus 5 at 63.0 and Claude Fable 5 at 62.1, per the index snapshot dated August 18, 2026.
The trap is Terminal-Bench. xAI reports 26% on v3.0. Artificial Analysis reports 88.4% on v2.1. Same model. Different benchmark versions. Anyone putting those two numbers in the same sentence is producing nonsense.
And the number that matters most for Grok Bot does not exist. There is no published task-success rate, no failure rate, no human-intervention rate for the product itself. The model has receipts. The workforce does not.
What it actually costs to run
The sticker price is the least interesting part of the bill.
Every eligible plan includes a weekly usage allowance whose size xAI has not published. Past that, you buy on-demand usage billed from raw model and token cost. There is no Grok Bot spend cap yet, and there’s no model picker, so you can’t steer toward something cheaper.
Now layer in how Grok 4.6 is priced: $2 per million input tokens and $6 per million output, which flips to $4 and $12 once a prompt crosses 200,000 tokens. That higher rate re-bills the entire request, not just the overage.
Long-running agent work is a context-accumulation machine. An agent that keeps working after you close the laptop is, structurally, the ideal device for generating tokens you were not watching. Some users have reported multi-bot sessions burning through weekly limits inside a couple of hours of real work, though that’s anecdotal and worth treating as such.
None of this is open source. There’s no license to read, no self-hosting path, and no way to audit what the thing did beyond your own chat transcript. The teams documentation says an audit view of bot actions is coming.
So What
Do this: If you’re already paying for Cursor Ultra or SuperGrok Heavy, create exactly one bot with one narrow read-only job on a system you can verify by hand. Give it a task you already know the right answer to. That’s how you get a completion rate when the vendor hasn’t published one.
Skip this: Do not upgrade to a $200 or $300 tier to try a beta whose reliability nobody has measured. That is buying a number that does not exist yet.
Wait on this: Anything involving payments, production systems, customer records, or credentials you cannot rotate quickly. The approval system stops a proposed action, and xAI says so directly: an approval does not reverse work already completed.
Steal this line: “The isolation boundary is the account, not the bot.” Say it in your next architecture review and watch how fast the room re-scopes what it was about to connect.
The part that outlives this product
Six months from now the beta labels come off and the completion rates get published. Grok Bot will either be good, or it won’t.
What survives is the question the shared desk forces you to answer: which of your workflows should ever be handed to something that logs in as you? That’s not a Grok question. Claude Cowork, ChatGPT Work, and Microsoft’s Copilot are all converging on the same shape. Every one of them will eventually ask you for the keys.
Figure out your answer now, while the stakes are one newsletter unsubscribe task instead of your accounts payable.
Sources
Introducing Grok Bot — SpaceXAI launch post, August 11, 2026
SpaceXAI’s Grok Bot turns agents into persistent digital coworkers — VentureBeat
Grok Bot review: what actually ships in the early beta — eesel AI
Grok 4.6 Benchmarks Explained: Why 26% and 88% Are the Same Model — Codersera
SpaceX completes record $60 billion acquisition of Cursor — Investing.com via Yahoo Finance
Grok Bot: xAI’s AI Teammates Get Their Own Computer — Digital Applied
Pricing and benchmark figures verified August 19, 2026. Both are volatile. Re-check the product page and the Artificial Analysis index before citing these numbers after September 1.


