youtube.nixfred.com nixfred.com

Grok Bot: 5 Must-Try Use Cases for Work and Life (Full Tutorial)

Peter Yang builds five working bots on Grok Bot, the personal agent platform that gives every user a persistent cloud computer on SpaceX AI servers, and argues it is the future of personal AI agents. He walks through an advisor that designs and spawns his other bots, a YouTube researcher that emails him outlier videos every morning, an X scout that digests his timeline and bookmarks, a digital Marie Kondo that unsubscribed his email, trashed Drive files, and canceled two subscriptions in about five minutes, and a travel concierge that found a Japan flight routing about $2,700 cheaper than his own plan. Every bot follows the same loop: initial prompt, iterate on the output, then schedule a daily routine. A bonus test installs Doom and Commander Keen on the cloud machine, which mostly proves it is not a gaming PC yet. He closes with his honest take: the cleanest UI in the category, a real trust hurdle for normal users, and at $200 a month it does not yet replace his $20 ChatGPT subscription.

Published Aug 17, 2026 22:54 video 40 min read Added Aug 21, 2026 Open on YouTube →

At a glance

Peter Yang thinks the interesting thing about Grok Bot is not the model, it is the machine: a persistent computer in the cloud, with its own browser and its own operating system, that stays logged into your apps while your laptop is closed. He builds five bots live on that machine and shows every prompt: an Advisor that invents and spawns the other bots, a YouTube Researcher that mails him a morning brief of outlier videos and comment themes, an X Scout that mines his own timeline and bookmarks for the ten posts he actually wanted to see, a Marie Kondo that audits Gmail, Google Drive and his recurring charges and then unsubscribes, trashes and cancels on his approval, and a Personal Concierge that reads his Japan trip document and watches flight prices. The single most valuable result in the video is a boring one: the concierge found that a Tokyo round trip is about $2,700 cheaper than the open jaw itinerary he had planned, which a plain price alert would never have surfaced because it would never have questioned the route. A bonus round has the agent install Doom, Red Alert and Commander Keen on its own machine, which half works and is more revealing than if it had worked cleanly. He closes with a verdict he does not soften: the biggest obstacle is trust, not capability, and at $200 a month against ChatGPT at $20, Grok Bot is not his daily driver yet.

What a "dedicated cloud computer" actually means (0:41)

The video opens with the claim in one line: Grok Bot is the future of personal agents. Then he backs up and explains why he went looking for it in the first place.

A few weeks before filming, he tweeted that ChatGPT was one feature away from being the AI product he actually wants. That missing feature is not a better model, a longer context window, or another reasoning mode. It is a dedicated personal computer in the cloud for running AI agents. Shortly after the tweet, Lee, from the team building Grok Bot, sent him a DM asking him to test it.

So what is a dedicated cloud computer, and why does it matter enough to build a product around?

He answers with the meme. You have seen people walking around with a laptop half open, lid propped, because closing it would kill the agent mid task. That posture is the whole problem stated as body language: the agent is a process on your machine, so your machine has to stay awake, plugged in, and carried around for as long as the work takes.

"You don't have to do that anymore if you're using Grok Bot." (1:21)

Grok Bot lives on a computer sitting somewhere in xAI's server fleet rather than in your house. It is a real machine, not a sandboxed function call. It has its own browser and its own operating system, and, crucially, you can log into your favorite apps on it and stay logged in. That last property is the one everything else in the video depends on. Every bot he builds is really a bot plus a session cookie that nobody logs out of.

The three way comparison: local agent, cloud browser, cloud computer

He puts three options side by side, and this is the part worth understanding before any of the bots make sense.

Hermes on a Mac Mini at home. This is what he has been running. It sits on 24/7, so the agent genuinely has a persistent computer with persistent logins. The cost is that you have to go buy the machine and set the whole thing up yourself. You can technically run the same setup on a virtual private server instead, but that is more setup work, not less, and now you are also a sysadmin.

ChatGPT work mode. It uses plugins and a cloud browser, so the compute is somebody else's problem. Two things spoil it for him. First, the browser as it stands cannot stay signed in to your favorite apps, which quietly disqualifies most of the automations in this video: a bot that has to ask you to log in again on every run is not an automation, it is a chore with extra steps. Second, the UX is scattered across chat, work mode and Codex. He screen shares it and narrates the mess honestly: a pile of chat threads here, Codex over there, another tab up top for work and chat, and it is confusing even to someone who does this professionally.

Grok Bot. You get a persistent cloud computer out of the box, with no machine to buy and no VPS to configure. Of the three, he calls it the easiest to set up and get going.

WHERE DOES THE AGENT'S COMPUTER LIVE? Hermes on a Mac Mini at home, on 24/7 Your house machine you bought Persistent logins yes Cost: buy hardware, set it all up yourself, or run a VPS instead. ChatGPT work mode plugins + cloud browser Someone else's cloud browser, no OS of yours Persistent logins not today Cost: UX split across chat, work and Codex. "Kind of a mess." Grok Bot persistent cloud computer xAI servers its own browser + OS Persistent logins yes, out of the box Cost: you sign into your accounts on a machine you do not own.
Figure 1. The comparison he draws at 1:40. All three give an agent a computer. Only two of them keep you signed in, and only one of those does it without you buying and maintaining the hardware. The trade he flags at the end of the video is the third box's fine print: the persistence you gain is persistence on somebody else's machine.

The interface, and why it is not a chat log

The second difference is softer and he is upfront that it is a taste argument: the interface feels more focused and more delightful. Each bot is a named thing in a list with its own personality and style, and when a bot starts working it plays a small animation. He demos it by typing a throwaway task, "look up the weather," to show the subtle motion when a bot spins up.

He credits the design team behind the product for that, naming Jenny Wen, whom he has interviewed on this channel before.

"Overall, I think this UX feels more like talking to a coworker than getting lost in a hundred chat threads." (3:10)

That sentence is the real product thesis of the video, and it is worth separating from the marketing. A chat thread is a transcript you have to re-read to know what state you are in. A bot is a persistent named agent with standing instructions, a schedule, and a body of work. The five builds that follow are all the same move: turn a thing you would have re-typed into a chat window every week into a bot that already knows and already ran.

The question he holds open for the rest of the video: can Grok Bot actually replace ChatGPT and Codex, which are his current daily drivers?

Bot one, the Advisor: the bot that builds your other bots (3:30)

His advice is unambiguous. The first bot you should create after installing Grok Bot is an Advisor, and he keeps his pinned to the very top of his bot list.

The job of the Advisor is not to do work. It is to know enough about your work and your life to tell you which bots are worth building, and then to build them. So you start by telling it who you are. Here is what he told his, close to verbatim:

I'm a creator and founder of Behind the Craft. I have two daughters and
I live in the Bay Area. I publish newsletter posts and YouTube videos
focused on practical AI tutorials and podcast interviews. I also have a
membership portal.

Based on all this, suggest five bots that can save me time or money.
For each, explain what it would do in a few sentences.

Two things about that prompt are doing the work. It is specific about the business (a creator business with a newsletter, a channel, podcast interviews and a paid membership) and it is specific about the life (two daughters, the Bay Area), because the bots he wants are not all work bots. And it constrains the output shape: five, with a few sentences each, so the answer is scannable rather than an essay.

The Advisor came back with five ideas:

  1. Daily AI news scout
  2. YouTube comment miner
  3. Membership concierge
  4. Social media poster
  5. Podcast prepper

Then comes the move that most people miss, and it is the reason this is bot number one rather than a nice to have:

"What many people don't realize is that you can actually get one bot, in this case our Advisor bot, to help us create the rest of our bots." (4:14)

He does not open a new window and start from scratch. He asks the Advisor to merge two of its own ideas into one bot, in one line: build a single bot that both researches other channels and mines comments, and call it YouTube Researcher. The Advisor creates it and hands it its opening instruction.

THE ADVISOR PATTERN You, once work, life, business, "suggest five bots" Advisor pinned to the top of the bot list proposes + creates 1. Daily AI news scout 2. YouTube comment miner · 3. Membership concierge 4. Social media poster · 5. Podcast prepper five ideas, a few sentences each "merge 2 and 3, call it YouTube Researcher" Working bots, created BY the Advisor YouTube Researcher X Scout Marie Kondo Personal Concierge Gamer THE LOOP EVERY BOT GOES THROUGH 1. Initial prompt kick it off, see output 2. Iterate on format never one shot it 3. Constrain length numbered list, max N 4. Schedule it daily / weekly job
Figure 2. The structure of the whole video. One bot briefed on your life becomes the factory for the rest, and every working bot goes through the same four step loop before it earns a schedule. He states the loop explicitly at 6:47 and then follows it, without exception, for every bot that follows.

Bot two, the YouTube Researcher: outliers and comment mining (4:30)

The Advisor kicked this one off with its own instruction: create a morning brief that sends YouTube intel every morning. Reasonable, and not good enough, so he takes over and gives it specifics.

First instruction: monitor named channels in his niche. He does not ask it to go find interesting YouTube content in general, he points it at a specific competitive set. You watch it start monitoring channels in real time.

The first brief is bad, and he shows it anyway. This is the most useful thirty seconds in the video for anyone who has tried this and given up. The initial report is verbose and hard to follow, a wall of stuff. It did do one thing right: it pulled the comments from his own videos and found themes in them. But as a document it is unusable.

So he writes the format he wants, rather than complaining that the output is bad:

Give me the report in this specific format:

1. My top three content ideas
2. The top five outliers from the other channels I follow
3. The top performing videos overall
4. The top common themes from my own comments

Limit your research to the last 14 days.

The 14 day window is not arbitrary and he explains it: YouTube in AI is intensely topical, so anything older than two weeks is noise for the purpose of deciding what to film next. The definition of outlier is also his, and it is the right one: a video that beat its own channel's typical performance over that window. Not the biggest video, the most over performing one, which is the only version of the metric that transfers to a channel of a different size.

Then the line that generalizes past this bot:

"By the way, this is how you should be working with AI to refine its output, instead of trying to one shot something." (5:52)

The second report is much more concise. Top three content ideas: one is a video he had already planned to make, one is a video on Grok Bot, which is the video you are watching. Then the outliers from other channels, with topics he says are genuinely worth following up on. Then top performers overall, then the common themes from his comments.

Then he schedules it. A daily job sends him the report every morning, and he shows that morning's delivery: top angles, a watch list, and comment themes. What he has bought himself is a research assistant that proactively finds topics to make videos about and feeds back what his own audience is asking for, without him opening anything.

And here he states the pattern he will use for the rest of the video:

  1. Kick the bot off with an initial prompt.
  2. Iterate back and forth until the output is actually good.
  3. Schedule a routine or job (daily, weekly or monthly) so it does the work proactively.

Bot three, the X Scout: mining your own timeline (7:00)

The premise for this one is a data advantage. Grok Bot is part of xAI, so it should have first party access to X data rather than scraping it through a browser like everyone else.

The personal premise is more relatable, and he is candid about it:

"Even though I have an unhealthy addiction to X, I still miss plenty of great content, or bookmark tweets that I never look at again." (7:12)

That is the actual failure mode of the platform for a heavy user. It is not that there is nothing good, it is that the good things arrive unsorted at 3am and the bookmark folder is where interesting posts go to die.

His prompt, close to verbatim:

Find the most viral X posts from the last seven days in my niche and
create a weekly report with the top 10 posts from people I follow,
engaged with, or bookmarked, grouped into a few categories.

Include the full post copy, a link, and your analysis.
End with three content ideas for me to post about.

Note the three sources it draws from: accounts he follows, accounts he has engaged with, and his own bookmarks. That last one is what turns a discovery feed into a personal retrieval system. And note the ending: it does not stop at reading, it converts what it read into three things he could post, which is the job he actually has.

The bot found his X account and automatically created a routine to run this every morning, without being asked.

The first output grouped into three themes:

Then the three things to post about next.

And then, because it costs one sentence to ask, he adds a category for fun: the top five funniest posts from his timeline. The number one funniest, as usual, is OpenAI and Anthropic sniping at each other and joking around. The rest are in the same AI niche. This is a small thing that matters more than it looks: the bot is already reading his whole timeline, so the marginal cost of also extracting the funny is zero, and it makes the report something he wants to open.

Getting it out of the app and into his inbox

The most practical moment in this section is when he stops typing and uses voice:

"Can you send this report to my email, and also make sure you include the full post copy in your report. Send it right now." (9:15)

The reason is a real one and it applies to every bot in this video. He does not want to open Grok Bot each morning to see the output. He wants it in the inbox he already opens.

The email arrives with the themes and the full post copy included, with the three things to post about at the bottom. The mechanism is unglamorous: it can send email because he has connected a set of plugins, including the Gmail plugin.

So the loop for X Scout ends up as: initial prompt, have it pull the information, iterate until the report reads right, then either schedule the job inside Grok Bot or push the output into email. He prefers email:

"I like to wake up with the top tweets to consume directly in my inbox instead of having to open a separate app." (10:00)

Bot four, Marie Kondo: the one that actually deletes things (10:30)

This is the bot with the highest stakes and the best payoff, and it is the one he most recommends building. It is named after Marie Kondo, and its job is your digital clutter.

He identifies three places clutter piles up:

  1. Your email, in the form of newsletters you never open.
  2. Your Google Drive, in the form of large and abandoned files.
  3. Your paid subscriptions, in the form of recurring charges you forgot you authorized.

The third is the one with money attached, and it is also the one that is genuinely hard to audit by hand, because the evidence is scattered across years of receipt emails.

He connects both the Gmail and Drive plugins, then gives it this:

Audit my Gmail, Google Drive, and recurring email receipts, and create
a cleanup plan.

- Find newsletters I never open.
- Find large or abandoned Drive files.
- Identify paid subscriptions from my email receipts.

Group everything into categories. Use a numbered list.

Do NOT move, delete, unsubscribe, or cancel anything without my approval.

He stops the video to underline the last line, and he is right to:

"Crucially, I told it to not move, delete, unsubscribe or cancel anything without my approval. The last line is really important for a bot that cleans up your files. You always want to review what it plans to delete before letting it remove anything." (11:12)

He also notes an optional fourth source: you can hook up the Mercury MCP server to pull the charge data directly, Mercury being the bank he uses for his business. That is a strictly better signal than parsing receipt emails, because it is the ledger rather than the paperwork about the ledger.

What actually happened, including the parts that did not work

Grok Bot first confirmed it was connected to Google Drive and Gmail. Then it connected to Mercury over MCP, and here is a detail that tells you what kind of product this is: he had to sign in to Mercury on the remote cloud computer to make that work. Not on his laptop. On the machine in the datacenter. Hold that thought, because it comes back in the closing section.

Then the first report, and it is bad in a specific and familiar way:

"This is a massive list of stuff to clean up. Honestly, this list is pretty overwhelming. It's incredibly long." (11:50)

He gives feedback: list it properly, categorize it properly. Still incredibly long. So he does the thing that works:

Show me a maximum of 10 items in each list, and get rid of all the
random labels.

Now it is digestible: a block of emails, a block of Google Drive files, a block of paid subscriptions. One more round of feedback narrows it to the actionable set only:

Only show me emails to unsubscribe from, Google Drive files to delete,
and paid subscriptions I want to cancel.

That third pass is the important one. The difference between a list of everything in your account and a list of decisions you have to make is the difference between a report and a tool.

The numbered list trick

With the final list in front of him as a numbered list, approving work costs him a sentence:

Email: unsubscribe to 3, 4, 5, 6, 7 and 8.
Google Drive: delete the extra tax return, delete some of these large files.
Paid subscriptions: cancel 11 and 15.

He pulls the general lesson out explicitly, and it is the single most portable tip in the video:

"It's always good to ask Grok Bot or AI to give you its response in a numbered list like this, to make it super easy for you to just tell it to do things by referring to the number in the list instead of having to type everything over again. I use this pattern all the time." (12:40)

The bot takes action

Now it runs, and he calls this the part where the magic happens. It unsubscribes from a batch of senders. It trashes three Google Drive files. Then it starts cancelling subscriptions: Lovable and Equip Foods, a protein company.

Both cancellations required him in the loop, and the friction is worth recording precisely. To cancel Lovable he first had to sign in to the cloud computer with his Lovable credentials, which he did manually, and then they discovered Lovable was already scheduled for cancellation anyway. Then it asked him to sign in to Equip Foods, he did, and the bot cancelled the protein subscription.

He pauses to be fair to the vendor he just cancelled:

"By the way, Equip Foods is a great protein company. I'm only canceling because I have too many protein powders at home already." (13:30)

The measured result:

"It did all this in around five minutes, when it would have taken me probably 30 minutes to an hour to do all this manually." (14:00)

That is the honest number in the video: roughly a 6x to 12x saving on a chore, on a task he was never going to get around to doing by hand.

MARIE KONDO, END TO END Connect Gmail plugin Drive plugin Mercury MCP (sign in ON the cloud PC) Audit prompt newsletters never opened large / abandoned files subscriptions from receipts "do NOT act without approval" Draft 1 massive list, random labels, overwhelming 3 rounds of feedback categorize properly max 10 per list actionable items only Numbered list, reviewed by a human "unsubscribe 3,4,5,6,7,8 · delete files · cancel 11 and 15" approval costs one sentence because every item has a number Email unsubscribed from a batch of senders Drive 3 files trashed, including a duplicate tax return Subscriptions Lovable (already ending) Equip Foods, cancelled ~5 minutes with the bot vs 30 to 60 minutes by hand two logins required on the cloud PC
Figure 3. The full Marie Kondo run as he performs it at 10:30 to 14:14. The guardrail sentence and the numbered list are the two design decisions that make a delete happy agent safe to use, and the three rounds of feedback are the part most tutorials edit out.

Making it talk like Marie Kondo

Then the joke that is also a lesson about personality being a feature, not decoration. He decides the bot sounds too much like a robot and not enough like its namesake:

Can you talk like Marie Kondo from now on? Give me an example.

The bot rewrites its own report voice, and the result is the funniest moment in the video, a cleanup log delivered as a small ceremony:

"The files have completed their work. We thank them and place them in the trash, where they may rest. Equip Foods no longer sparks joy. We release the prime protein subscription with gratitude." (14:35)

His closing advice on this bot: set Marie Kondo to run weekly or monthly to clean up your digital files, and never drop the numbered list requirement, so you review before anything is removed.

"You don't want it to accidentally delete some important file." (14:50)

Bot five, the Personal Concierge: the $2,700 sentence (15:00)

This is the bot he wants for all vacation and travel planning, and it produces the single best result in the video.

The setup: he has a vacation document listing his December trip to Japan, a full itinerary he built with AI. He previews it on screen. The flights are not booked yet, so what he wants is price monitoring, but smarter than an alert.

The prompt:

Read my vacation document and monitor the exact flight legs for my
family trip. Get the dates and the best options for each leg.
Check regularly and let me know when the price improves.

It found the flight legs in his document and started checking prices on Google Flights.

He asks the obvious objection out loud before you can: what is the advantage of doing this in Grok Bot instead of just setting a Google Flights price alert? His answer is the point of the whole bot:

"The value here is that Grok Bot can understand my whole trip based on my document and decide what's a better option for my family." (15:55)

An alert watches a route you already chose. This thing read the trip and questioned the route.

The finding

His planned itinerary was an open jaw: SFO to Tokyo, then Tokyo to Fukuoka, which is where the family is going, then Fukuoka back to SFO. It is the shape the document specified and the shape that looks obviously correct when you are planning a trip that ends in a different city from where it starts.

Grok Bot found that a Tokyo round trip is about $2,700 cheaper than that open jaw booking.

"A simple Google price alert would not have found this." (16:35)

That is correct, and it is worth being precise about why. A price alert is a function of a query, and the query is the itinerary. If the itinerary itself is the expensive decision, no amount of monitoring will tell you, because you never asked about the alternative. The agent had the trip document, so the search space it was allowed to consider was "get this family to these places on these dates", not "watch these three legs."

He then asks it to check every morning at 9:00 a.m. to see whether the price improves, and shows a follow up run: the price is still roughly the same, and the Tokyo round trip is still much cheaper.

WHAT THE CONCIERGE FOUND Planned in the vacation document: open jaw SFO Tokyo Fukuoka SFO leg 1 leg 2 leg 3 Three separate legs. The shape the trip suggests, and the shape a price alert would have watched forever. Found by the bot: round trip to Tokyo SFO Tokyo about $2,700 cheaper for the same family, same trip, same dates Why the alert could not find it: an alert monitors the route you already chose. The bot read the whole trip document, so it was allowed to question the route itself. Rechecked daily at 9:00 a.m.
Figure 4. The concierge's one concrete win, at 16:16. The dollar figure is the video's headline number, and the mechanism behind it is the argument for agents over alerts: give the agent the goal, not the query.

Where he wants to take it

He is explicit that price watching is the beginning, not the product. Eventually he can ask the bot to go ahead and book the flight, or to check him into the flight when the time comes. And more generally:

"It's always a good idea to have a travel thread or travel bot to both help you plan travel ahead of time, and also, when you're at a location, to help you book amusement parks and figure out what to do every single day." (16:57)

That is the shape of a concierge: one persistent agent that holds the whole trip, before and during, instead of a fresh chat every time a question comes up.

Bonus, the Gamer: what happens when the agent owns a real machine (17:30)

Because Grok Bot comes with a dedicated cloud computer, he does the thing you would do: he asks it to install and let him play retro games. Specifically Red Alert, Doom and Commander Keen.

It found the files and installed them. Then, "open Doom for us to play."

Doom comes up. He opens it inside the virtual cloud computer, starts a new game, picks an episode, picks Hurt Me Plenty, and this is where it falls apart:

"Here's kind of where Grok Bot falls apart a little bit. Because it's on a virtual cloud computer, the mouse isn't quite configured right to actually play Doom. For some reason it's looking at the floor all the time, and I can't seem to adjust the mouse to look up." (18:05)

A first person shooter with a broken vertical axis is not a game, so he moves on to Commander Keen, which he introduces with genuine affection as an awesome platformer he played in his youth. It loads. New game, one player, normal difficulty. He immediately forgets how to jump, discovers that Control is jump, and reports the honest result: it plays better than Doom, but there is enough lag in the keyboard and mouse to make the character difficult to control.

The conclusion is measured and it is not a dunk:

"Grok Bot is not replacing your gaming PC or GeForce NOW yet. But the fact that the agent can install and launch these games on its own computer gives you a sense of how open ended this could become." (19:10)

That is the right read. Nobody needs an agent to play Commander Keen. What the experiment demonstrates is that the machine is a real machine with a real filesystem and a real package situation, and that the agent has enough control over it to download, install and launch arbitrary software without a human touching a terminal. Every serious bot in this video is a consequence of that same capability.

He allows himself one speculation, and flags it as a dream rather than a roadmap: maybe xAI could eventually use its datacenters and GPUs (he jokes, in space) to deliver AAA games through the virtual cloud computer.

A look inside the machine

Then a small moment that is more informative than the games. He opens the file manager on the cloud computer and browses what is there. He notes it looks a bit like Windows 3.1 and admits he is not sure what it actually is. Inside: the Japan flights work from the concierge bot, other files from earlier bots, the games they just installed, plus Chrome and a terminal.

That is the whole thesis made concrete. The outputs of your bots are not messages in a chat log, they are files on a computer, sitting next to a browser and a shell, in a place that is still there tomorrow.

BotConnected toWhat it automatesCadenceWhat it actually returned
AdvisorNothing. It only needs to know you.Deciding which bots to build, then creating themOn demandFive bot ideas, then spawned the YouTube Researcher on request
YouTube ResearcherYouTube channels in his niche, his own video commentsCompetitive research and audience feedbackDaily brief every morningTop three content ideas, five channel outliers over 14 days, top performers, comment themes
X ScoutX (follows, engagements, bookmarks), Gmail pluginReading his own timeline and rescuing dead bookmarksDaily, delivered to his inboxTop 10 posts in three themes, full post copy, three ideas to post, plus the five funniest
Marie KondoGmail, Google Drive, Mercury over MCPUnsubscribing, deleting files, cancelling paid subscriptionsSuggested weekly or monthlyUnsubscribed a batch of senders, trashed 3 files, cancelled 2 subscriptions in about 5 minutes
Personal ConciergeHis vacation document, Google FlightsMonitoring flight prices against the whole trip, not one routeEvery morning at 9:00 a.m.A Tokyo round trip about $2,700 cheaper than the planned open jaw
Gamer (bonus)The cloud computer itselfInstalling and launching retro gamesOn demandInstalled all three. Doom unplayable (mouse look), Commander Keen playable with input lag
Figure 5. The six bots as one ledger. The right hand column is the honest scoreboard: two saved real money, two saved real reading time, one is infrastructure for the rest, and one is a stress test that half failed and was shown anyway.

The honest take: trust, privacy and price (20:30)

He saves the hard part for last, and he does not hedge it.

Trust is the bottleneck, not capability

The biggest hurdle for Grok Bot adoption, in his view, is trust. He illustrates it with the smallest possible example, which is why it lands: a Google sign in screen.

"When I see a Google sign in screen like this on my laptop, I don't really think twice before signing in. But because this appeared on the virtual cloud computer, I hesitated a bit, because how do I know that nobody else is seeing this screen on the virtual cloud computer?" (20:30)

That hesitation is the entire product category's problem in one sentence. Every capability in this video, the Gmail audit, the Drive cleanup, the Mercury connection, the Lovable and Equip Foods cancellations, required typing real credentials into a machine he does not own and cannot inspect.

He goes to the Grok Bot website and scrolls to the privacy note at the bottom, which answers the question with the standard set of assurances: it uses the same single sign on and privacy mode you already trust, the cloud computer is encrypted in transit and at rest, and there is no AI training on top of it.

His response to that is the most honest thing in the video, because he does not pretend the assurance settles it:

"I'm willing to give Grok Bot and the virtual cloud computer access to all this stuff because I'm an early AI adopter. But I can see normal people struggling to understand what this cloud computer thing even is, and hesitating to sign into their favorite apps on this device." (21:05)

And then the prediction:

"I think the AI agent platform that figures out trust will be the first to get mass adoption." (21:20)

Worth sitting with. The claim is not that the encryption is insufficient. It is that a correct security posture that a normal person cannot form a mental model of does not produce trust, and trust, not capability, is what gates adoption. Nobody needs to understand TLS to sign into Gmail on their own laptop, because they understand the laptop. Nobody yet understands the cloud computer.

The verdict

On the direction of travel he is unequivocal:

"Grok Bot is a clear sign of the future. We're moving away from manually using our keyboard and mouse to do work on our laptops, to using our voice to orchestrate a bunch of agents that live in a dedicated cloud computer. And Grok Bot is the first product to actually enable this." (21:30)

On whether he is switching, he is equally unequivocal, and the reason is price:

"It's not quite my daily driver yet, because I think ChatGPT still offers more for $20 a month, while Grok Bot requires paying $200 a month to use on a regular basis." (21:45)

That is a 10x price gap for a product he has just spent twenty minutes praising, and he states it without softening. What he does credit it with: a much cleaner UI than ChatGPT right now, and being very capable.

His own axesGrok BotChatGPT
Price for regular use$200 / month$20 / month
Breadth of what you getFocused on the cloud computer and bots"Still offers more" for the money
Interface"Much cleaner UI right now", named bots with personalitySplit across chat, work mode and Codex, "kind of a mess"
Persistent computerYes, its own browser and OS, out of the boxCloud browser, but no machine of yours
Stays signed into your appsYes, which is what makes the bots possibleNot as it stands today
Setup effortEasiest of the three he comparesNo setup, but no persistence either
Trust barrierYou sign into your accounts on a machine you do not ownFamiliar, and asks for less
His daily driver todayNot yetYes, with Codex
Figure 6. The closing verdict, scored only on the axes he uses himself at 20:30 to 22:00. Note that Grok Bot wins on every axis except the two that decide the purchase.

The competitive read

His last strategic point is about the shape of the market rather than the product:

"It's just great to be in a world where Cursor and xAI are just as viable a competitor as OpenAI and Anthropic. I think Cursor may even have the edge if it can continue to support multiple models from all providers." (22:05)

The reasoning behind the edge is worth extracting: a product that is a harness rather than a model can route to whichever model is currently best, which is a structurally different bet from a lab shipping the interface to its own weights. If the harness is the product, model leadership becomes a supply question rather than an existential one.

He closes with the practicalities. Grok Bot is free to download, and he recommends trying it to get a glimpse of where this is going. He is putting the prompts from the video into the pinned comment. And an exclusive interview with the team on how they built Grok Bot is coming in the next few weeks.

"I'm really impressed by Grok Bot. I think the team really cooked here, and I can't wait to hear the story behind how they built this." (22:35)

How to reproduce all of this

The video is a tutorial, so here is the whole method compressed, in his order, with nothing added.

  1. Install Grok Bot and build the Advisor first. Tell it about your work and your life in a paragraph, then ask for five bots that would save you time or money, with a few sentences each.
  2. Have the Advisor create the working bots. Merge and rename its own suggestions in plain language rather than starting each bot from a blank prompt.
  3. Connect the plugins the bots need before you need them. Gmail, Google Drive, and anything with an MCP server (he uses Mercury for banking). Expect to sign in on the cloud computer, not on your laptop.
  4. Kick each bot off with a specific initial prompt, including the sources it should look at and the shape of the output.
  5. Expect the first output to be unusable. Iterate. Specify the report format explicitly, section by section. Constrain the time window if the domain is topical.
  6. Force a numbered list, and cap the length (he uses a maximum of ten items per category). Then approve or reject work by number.
  7. For anything destructive, put the guardrail in the prompt: do not move, delete, unsubscribe or cancel anything without my approval.
  8. Schedule the bot as a daily, weekly or monthly routine once the output is right.
  9. Push the output to where you already look. He routes reports to email so he never has to open the app.
  10. Give the bot a personality if it helps you read it. Marie Kondo is funnier and therefore more likely to be read than the same list from a robot.

Key takeaways

Chapters

Notable quotes

"You know the meme of people walking around with their laptop half open to keep their agents running? You don't have to do that anymore if you're using Grok Bot." (1:21)

"Overall, I think this UX feels more like talking to a coworker than getting lost in a hundred chat threads." (3:10)

"What many people don't realize is that you can actually get one bot, in this case our Advisor bot, to help us create the rest of our bots." (4:14)

"By the way, this is how you should be working with AI to refine its output, instead of trying to one shot something." (5:52)

"Even though I have an unhealthy addiction to X, I still miss plenty of great content, or bookmark tweets that I never look at again." (7:12)

"Crucially, I told it to not move, delete, unsubscribe or cancel anything without my approval. You always want to review what it plans to delete before letting it remove anything." (11:12)

"It's always good to ask AI to give you its response in a numbered list, to make it super easy to just tell it to do things by referring to the number instead of typing everything over again. I use this pattern all the time." (12:40)

"It did all this in around five minutes, when it would have taken me probably 30 minutes to an hour to do all this manually." (14:00)

"Equip Foods no longer sparks joy. We release the prime protein subscription with gratitude." (14:35, Marie Kondo bot, in character)

"A simple Google price alert would not have found this." (16:35, on the $2,700 cheaper Tokyo round trip)

"Grok Bot is not replacing your gaming PC or GeForce NOW yet. But the fact that the agent can install and launch these games on its own computer gives you a sense of how open ended this could become." (19:10)

"How do I know that nobody else is seeing this screen on the virtual cloud computer?" (20:30)

"I think the AI agent platform that figures out trust will be the first to get mass adoption." (21:20)

"We're moving away from manually using our keyboard and mouse to do work on our laptops, to using our voice to orchestrate a bunch of agents that live in a dedicated cloud computer." (21:30)

"It's not quite my daily driver yet, because I think ChatGPT still offers more for $20 a month, while Grok Bot requires paying $200 a month to use on a regular basis." (21:45)

Resources mentioned

The product under test

What he compares it against

Apps and services the bots connect to

Games installed on the cloud computer

People and namesakes

Where it stands

Everything above is his experience of a very new product, filmed as an early tester who was invited in by the team building it, and it is worth reading with that context rather than instead of it.

What is demonstrated on camera is solid. The bots exist, the reports arrive, the unsubscribes and cancellations actually execute, the flight comparison actually returns a number. He also shows the failures: three bad drafts before a usable report, a games experiment that half works, a cancellation flow that required him to type credentials twice.

What is claimed rather than demonstrated is the durability. A daily job that has run for a few mornings is not the same as one that survives a month of expired sessions, changed login flows and rate limits, and every automation here depends on staying signed in to services that have commercial reasons to log agents out. The $2,700 saving is a quoted comparison, not a booked ticket; he has not bought the flights yet.

The trust question he raises is the one that outlives the product. His own framing is the right one: he signed in because he is an early adopter, and he expects normal users to hesitate. Nothing in the privacy note he reads out is unusual for a cloud service, and the discomfort is not really about the encryption. It is that a browser session on a remote machine has no analogue in how most people think about their own computers, so there is no intuition to lean on when deciding whether to type a password into it.

And the price is the honest verdict, stated by the creator of the tutorial at the end of the tutorial: $200 a month against $20 is a wide gap, and he did not switch.

Full transcript
Hey everyone, I think Grockbot is the future of personal agents and I'm going to show you how to build five of my favorite bots. In this video, we're gonna build an advisor to create and manage your other bots. A YouTube researcher to find out videos in your interest. An ex Scout to surface the most insightful and funny tweets from your timeline. A digital Marie Condo to clean up your email and find paid subscriptions to cancel. A personal concurge to help you get good prices on travel and vacations. And as a bonus, we're going to build a gamer to see if Grockbot can play classic games like Doom and Red Alert. We'll finish with a discussion about privacy and whether Grockbot should be your new AI daily driver. All right, but first, let me show you what makes Grockbot different. A few weeks ago, I tweeted that chat GBT was one feature away from building the AI product that I really want, which is a dedicated personal computer in the cloud for running AI agents. Lee from the cursor team DM'd me shortly after to test Grockbot. What exactly is a dedicated cloud computer? You know the meme of people walking around with their laptop half open to keep their agents running? You don't have to do that anymore if you're using Grockbot. Basically, Grockbot lives in a computer sitting somewhere in SpaceX AI servers instead of in your home. It has its own browser and operating system that you can use and logging to your favorite apps. Now, let's do a quick comparison of Hermes, Chat GBT, and Grockbot. Hermes I run on my Mac Mini at home. It's on 24/7 and it gives it a persistent computer, but you have to go buy the machine and set everything up for yourself. You can technically also run Hermes on a virtual private server, but that again requires more setup work. Now, chat GPT work uses plugins and a cloud browser, but the problem is that browser at is right now cannot stay signed in to your favorite apps. The UX for chat GBD work also feels scattered across chat work and codecs. Let's take a quick look here. It's just kind of like a mess of chat threads here. There's chat GPT codeex. There's like another tab up here for work and chat. And it's just all really confusing even to someone like me. In contrast, Scrotbot gives you a persistent cloud computer right out of the box. And I think it's the easiest of the three to set up and get going. and his interface, as you can see here, also feel a lot more focused and delightful. I love the little bots and the animations that they make when they start working. Right, let's say look up the weather. And you can see here that there's some pretty cool subtle animations. Each bot has a different personality and style. And this is the work of Kurser's great design team, including Jenny Wen and other folks on her team that I've interviewed in the past. Overall, I think this UX feels more like talking to a coworker than getting lost in a 100 chat threads. So, can Grockbot actually replace Chat GBT and Codeex, which is my current daily driver? Let's run through some of my favorite bots right now. So, the first bot that you should create after you install Grockbot is an advisor. Here I have my advisor ped to the very top of my Grockbot list. With your adviser, you want to start by telling it about your work and life and then actually ask it to give you some ideas for boss to create. Here's what I told my adviser. I'm a creator and founder of Behind the Craft. I have two daughters and I live in the Bay Area. I publish newsletter posts and YouTube videos focus on practical AI tutorials and podcast interviews. And I also have a membership portal over here. So based on all this, suggest five bots that can help me save me time or money. And for each explain what it would do in a few sentences. As you can see here, it created five bot ideas. A daily AI news scout, YouTube comment miner, membership concurge, social media poster, and podcast prepper. Right now, what many people don't realize is that you can actually get one bot, in this case, our advisor bot, to help us create the rest of our bots. Here I asked my adviser to build a single bot that both researches other channels and mine's comments and let's call it YouTube researcher. Right? So now let's go ahead and check out the YouTube researcher use case. You can see that the advisor kicked off the YouTube researcher bot with this instruction to create a morning brief to send me YouTube intel every morning. But I want to give it more specific instructions. Right? First, I asked it to monitor specific channels in my niche. And let's scroll down here. And it's starting to monitor channels. And here is its first attempt at a brief. You can see here that the brief is like just like very verbose and hard to follow. It's like a lot of stuff here. It did end up pulling my YouTube comments and finding themes there. So, that's great. But, this is just like too much to process, right? So, I gave it some more instructions. I told it to give me a report in the specific format which is first list your top three content ideas. Then give me the top five outliers for other channels and then give me the top performing overall. Limit your research to the last 14 days because YouTube with AI stuff is very topical and also add the top common themes back to the report. And by the way, this is how you should be working with AI to refine its output instead of trying to oneshot something. Now you see here that the report is much more concise and easier to follow. Let's take a quick look. Top three content ideas. It wants me to make a video on my spec scale that I plan to do soon. It tells me to make a video on Grockbot that we're making right now. And here are the top outliers from the other channels. Outlier meaning that it beat that channel's performance for the last 14 days. Yep, pretty interesting topics here that we can follow up on. and top performing overall and also the top common themes and it did more here and then I told you to send a full report again just to make sure it looks good and now what I can do is basically there's a daily job that sends me this report every morning right just this morning for example it sent me a new report with the top angles the watch list and comments and I basically have a YouTube researcher that proactively finds interesting topics for me to make videos on and gives me feedback back based on my user's comments. Okay, so this is the process that we're going to follow for pretty much the rest of the boss that we're going to make, which is we're going to kick it off with an initial prompt. We're going to iterate back and forth with it to make his output good, and then we're going to schedule a routine or job that runs daily, weekly, or monthly to have it proactively do work for us. Now, let's actually see how we can follow the same process with our X Scoutbot. So, Grockbot is part of Space X AI, right? So, it should have firstparty access to X and Twitter data. And even though I have an unhealthy addiction to X, I still miss plenty of great content or bookmark tweets that I never look at again. Here's the prompt that I gave my exbot. Find the most viral ex posts from the last seven days in my niche and create a weekly report with the top 10 tweets from people I follow, engaged with, or bookmarked grouped into a few categories. Include a full tweet copy link and your analysis and end with three content ideas for me to tweet about. Next, you can see here that it found my ex account and it automatically created a routine to do this every morning. And here is the initial output, right? So it found themes around playbooks that people save not just like from Greg. 23 ways I use AI agents to grow my startup. Chinese developer from loop to graph engineering and more tweets from Greg. Another theme that I found is just about Grockbot codeex and chat GPT work compared and actually used right a tweet from Riley and from Ben who works on the cursor team on the most loved internal Grockbot use cases and more. And the third theme is around creative workflows that I actually am interested in tasing. And they ended this report with the three things to tweet about next. Some ideas here. And just for fun, I asked you to also give me the top five funniest tweets from my timeline. I think digging through it. And the number one funniest tweet is as usual, OpenAI and Anthropic kind of sniping at each other and kind of joking, right? And then there's some more funny tweets down here. all kind of in my AI knee niche. You can see here that it set up a routine automatically and it sent me a new report this morning. Now, let's actually tell it to do something else. I'm going to turn on voice here. Can you send me this report to peterbehindthecraft.com and also make sure you include a full tweet copy in your report. Send to my email right now. All right. So, let's see if Ashi is able to send this report over email because I don't want to have to open Grockbot each time to see it. I want it just in my inbox every morning. Awesome. Here's a report from Grothbot. You can see here that it has this themes. It's included the full tweet copy. You can scroll down here and yeah, I I think this looks pretty good, right? And then it has down here, I'm sure, the top three things to tweet about next. Again, the process here is to give Grockbot initial prompt, have it pull up the information, iterate with it to get the output and the report right, and then either tell it to send you a daily job through Grockbot itself or through your email. In this case, I prefer email because I like to wake up with the top tweets to consume directly in my inbox instead of having to open a separate app. And by the way, it's able to send me an email because I connected to a bunch of plugins, including the Gmail plug-in. Now, let's move on to a fun bot that I call Marie Condo that cleans up your digital clutter. There are three places where clutter piles up, right? Your email, Google Drive, and your pay subscriptions. Now, I've connected both Gmail and Drive in plugins to Grubbot. And then I gave it this prompt. Audit my Gmail, Google Drive, and recurring email receipts and create a cleanup plan. Find newsletters I already open. Find large or abandoned drive files. Try to identify pay subscriptions from email receipts. And also you can hook up Mercury MCP to get this data. Mercury is the bank that I use for my business. And then group everything into these different categories. Use a number list. And crucially, I told it to not move, delete, or unsubscribe or cancel anything without my approval. Right? The last line is really important for a bot that cleans up your files. You always want to review what it plans to delete before letting it remove anything. And you can see here that first Grockbot confirmed that it's connected to Google Drive and Gmail. And then it actually connected to Mercury MCP. I had to sign in on the remote cloud computer to make this work. And then it gave me this initial report. This is a massive list of stuff to clean up in my email subscriptions to review and other things that it should not touch. Right? Honestly, this list is like pretty overwhelming. It's incredibly long. And I gave it some feedback of like, hey, list it properly, categorize it properly. It's still an incredibly long list. Right? Again, this stuff is not going to get it right in one shot. And eventually I told it to just show me max 10 items in each list and get rid of all the random labels that it has. So now this is much more digestible. It has a bunch of emails here, Google Drive and pay subscriptions. And I gave it one final piece of feedback to only show me emails to unsubscribe to, Google Drive files to delete and pay subscriptions that I want to cancel. And then here's the update list that I came up with. Now it has this list as a number list and my feedback is to actually get it to take action which is email unsubscribe to three and four 5 6 7 and 8 Google Drive delete the extra tax return delete some of these large files and pay subscriptions cancel 11 and 15 right and just as a general tip it's always good to ask Grothbot or AI to give you its response in a number list like this to make it super easy for you to just tell it to do things by just referring to the number in the list instead of have to type everything over again. I use this pattern all the time. All right, so now Crockbot is actually taking action, right? This is where the magic happens. You can see here is already unsubscribed to a bunch of senders. It's trashed three Google Drive files that we asked it to and now it's cancelling Lovable and Equipped Foods, which is a protein company. Now to cancel lovable I first have to sign to the cloud computer with my lovable credentials which I did manually and we found out that actually lovable is already scheduled to be cancelled. So lovable is done. Then it asked me to sign to equip foods and I did that and then it cancelled uh the protein subscription for me. By the way, Equip Foods is a great protein company. I'm only canceling because I have too many protein powders at home already. Uh so it basically did all this work, right? trashed a bunch of stuff. It's unsubscribed to a bunch of emails and it's canceled a bunch of subscriptions to save me money. And you can see that it did all this in around five minutes when it would have taken me probably 30 minutes to an hour to do all this manually. Marie Condo is very very useful. Of course, and there's something missing with this bot. I think we should ask it to can you talk like Marie Condo from now on? Give me an example because right now it sounds too much like a robot and not so much like Marie Condo and let's see what it comes up with. All right. So now uh Marie Condo is actually going to talk like Marie Condo. You can see Hermis and the singing have completed their work. We thank them and place them in the trash where they may rest. Equip foods no longer sparks joy and release prime protein subscription with gratitude. Yeah. So now we can set up Marie Condo to maybe talk to us every week or every month to clean up our digital files and spark joy, right? So this is definitely a a bot that I recommend you setting up. Just remember to ask it to give it output in a number list instead of just doing the cleanup for you so that you can review it first, right? You don't want it to accidentally delete some important file. Now let's cover the personal concurge bot. This is a bot that I want to use for all my vacation and travel plans. It has access to my vacation document where I've listed my December trip to Japan. Here's a quick preview of the document. It has the full iterity for Japan that I created with AI as well. Let's go back to Grockbot. And what I want Grockbot to do is I haven't quite booked my flights to Japan yet. So, I wanted to monitor for price changes and alert me on what routes are the best deal. Here's a prompt that I give it. Read my vacation document and monitor the exact flight legs for my family trip. Get the dates, the best options for each flight leg. Check regularly and let me know when the price improves. You see here that I found my flight legs and is using Google flights to check the prices. Now, you may be wondering what's the advantage of using Grockbot for this instead of just setting up Google flight price alerts. And the value here is that Grockbot can understand my whole trip based on my document and decide what's a better option for my family. If you scroll down here, the big finding is that a Tokyo round trip is about $2,700 cheaper than the open jaw flights that my dog prefers. Right? My dog prefers it to fly from SFO to Tokyo and then from Tokyo to Fukoka, which we're going to, and then from Fukuoka back to SFO. But actually, Grockbot found a better itinerary. It's about $2,700 cheaper to just fly round trip to Tokyo instead of doing this itinerary that I laid out, right? A simple Google price alert would not have found this. And now I ask it to check every morning at 9:00 a.m. to see if the price improves. And you can see here that the price is still around the same. The Tokyo round trip is still much cheaper. Eventually, I can ask Rockbot to also just go ahead and book the flight for me or check into the flight when the time comes. And generally speaking, I think it's always a good idea to have a travel thread or travel bot to both help you plan travel ahead of time and also when you're at a location to help you book amusement parks, to help you figure out what to do every single day. So, I imagine I'm going to be using this personal travel concurge a lot. All right, let's test one more use case with Grockbot. Because Grockbot gives me a dedicated cloud computer, I thought it would be fun to ask it to install and let me play some retro games. I asked it to install Red Alert, Doom, and Commander King. And you can see here that it found the files and installed it. So, why don't we just ask it now? Let's say open Doom for us to play. And let's see how things work out. All right, looks like Doom is up. Let's open it in our virtual cloud computer. And here it is. And here we can start a new game. Let's pick this episode. And yeah, let's say Hurt Me Plant Penty. Here's Doom. And here's kind of where Grockbot falls apart a little bit. Unfortunately, because it's on a virtual cloud computer, the mouse isn't quite configured right to actually play Doom. For some reason, it's looking at the floor all the time, and I can't seem to adjust the mouse to actually look up. All right. So, Doom doesn't quite work here right now. Now, let's actually try playing some other game. Let's try asking to play Commander King. I'm not sure if you guys know this, but Commander King is an awesome platform game that I played in my youth. And now it's loading Commander King. So, let's see what it comes up with. Awesome. This is Commander King. Let's see if it plays well or not. Let's start a new game. One player normal difficulty. And here we go. Here we go. It's a very basic platformer, but I remember really enjoy it in my youth. How do I jump? I forgot how to jump. Oh, so control is jump. And yeah, it plays better than Doom for sure, but there is some lag in the keyboard and mouse that makes it difficult to actually control the character. Okay, so um yeah, that's just for fun. But basically, Grockbot is not replacing your gaming PC or GeForce now yet. But the fact that the agent can install and launch these games on its own computer gives you a sense of how open-ended this could become and how much potential there is here. Right? So maybe Space X AI can use all the data centers and GPUs in space to actually deliver AAA games through this virtual cloud computer. That would be a dream, right? And while we're at it, I open this file manager here. And here's all the stuff that we installed into our virtual cloud computer. It kind of looks like Windows 3.1 a little bit. I'm not sure what kind of thing it is. There's Japan flights. There's other stuff here, right? And then there's games that we've installed. And also we have Chrome and a terminal. So that's kind of our virtual cloud computer. All right. Now, All right. Now before we wrap, I want to talk about probably the biggest hurdle for Grobbox adoption, which is trust. When I see a Google sign screen like this on my laptop, I don't really think twice before signing in. But because this appeared on the virtual cloud computer, I hesitated a bit because how do I know that nobody else is seeing this screen on the virtual cloud computer? Now, if you go to the Grogbot website, if you scroll all the way down here, there is a note about privacy here, which is how does Grockbot handle my privacy? Grabbot uses the same cursor SSO off and privacy mode you already trust. Your cloud computers encrypted in transit and at rest and there's no AI training on top of it. Right? So, you know, I'm willing to give Grockbot and the virtual cloud computer access to all this stuff because I'm an early AI adopter, but I can see normal people struggling to understand what this cloud computer thing even is and kind of hesitate to sign into their favorite apps on this device. So, I think the AI Asian platform that figures out trust will be the first to get mass adoption. But overall, I think Grockbot is a clear sign of the future. We're moving away from manually using our keyboard and mouse to do work on our laptops to using our voice to orchestrate a bunch of agents that live in a dedicated cloud computer. And Grockbot is the first product to actually enable this. Now, it's not quite my daily driver yet because I think Chat GPT still offers more for $20 a month while Grockbot requires paying $200 a month to use on a regular basis. But I think Grockbot has much cleaner UI than chat right now and is also very very capable. And overall, it's just great to be in a world where cursor and SpaceX AI are just as viable a competitor as OpenAI and Anthropic. I think cursor may even have the edge if it can continue to support multiple models from all providers. So, Grockbot is available for free and I definitely recommend downloading and trying it to see a glimpse of the future. I'll include some of the prompts from this video in the pin comment below. And I also have an exclusive interview with the Cursor team on how they built Grockbot coming up in the next few weeks. Overall, I'm really impressed by Grockbot. I think the team really cooked here and I can't wait to hear the story behind how they built this. So, please like and subscribe if you enjoy this video and I'll see you next time.