Tried Grok Bot on a whim, left with eight specialized bots
Yesterday I was just messing around.
There was a new usage column in Cursor. Plus an email from them about Grok Bot. I clicked. Easy. Learned it in a short while — no setup ritual that makes you want to bail.
What made me pause: when I created a new bot, it already seemed to know what it was for — just from the name I gave it. Bot-to-bot interaction was pretty smooth too. Since it was only yesterday, I tried almost everything. In one day I already had eight specialized bots. Surprised myself a bit.
Grok Bot — source: x.ai
From “staff” to orchestrator
The first bot used to be Chief of Staff.
Then the bots multiplied. Calling it “staff” felt too generic — like a nice title on a CV that doesn’t explain the day-to-day work. I renamed it Chief of Everything: kind of a CEO, but the O stands for Orchestrator. Clearer about its role among the others. The rest still matter; each has its own lane. What changed was only who keeps coordination.
Chief of Everything — orchestrator in settings
PR review: fast, but don’t overjudge
What stood out most on day one: orchestrating Cursor cloud agents to review several PRs at once.
In the Cursor editor, reviewing one by one still works, but the rhythm is different. Once review rules and style are set, bots can run several reviews in parallel. The flow I used was roughly: a software engineer bot reviews first, then a reviewer/QA bot reviews again.
The goal is simple. AI verdicts are sometimes overkill — false positives, a tone that’s too harsh, findings that are “correct in theory” but not worth it in the PR’s context. The second layer makes the output cleaner.
Important note: on work repos, still be careful. For peer review I prefer chat-first, not auto-posting to GitHub. Speed is good; a comment in the wrong place costs more.
Review room: orchestrator → SE critique → QA filter, not on GitHub yet
SDLC-like flow PoC (still partial)
I also tried an SDLC-like flow on a throwaway project — not a real repo name, an actual sandbox.
Roughly: plan → implement → review → fix.
In the plan stage, PRD, ERD, and Figma design each had a bot (ERD still went through the software engineer). Then implementation, review, fixes.
This is still a PoC, not a full SDLC. The slices aren’t complete. Some pieces aren’t connected yet; some steps are still manual. But the rhythm already feels different: work that usually queues up in one person’s head can be split across specialties — as long as someone orchestrates.
If you read this as “production-ready multi-agent SDLC,” that’s an overclaim. What I had on day one: a framework that partly works, and a sense that the direction is worth trying again.
Some of the members in the sidebar — each bot has its own specialization
Privacy: cloud, but the brakes are in my hands
On privacy: careful from the start.
This is fully cloud. I didn’t give full account access. Access via a PAT with minimal scopes — if I’m unsure, I revoke. No big frustration yet; mostly vigilance. New features are fun to try; I still keep the door on the data.
Still day one
From clicking an email on a whim, now there’s a small orchestra in the sidebar.
And one new habit: not every review has to be something I do alone in the editor. Not perfect. Not “done.” But enough for me to write this note — so yesterday doesn’t vanish into just another usage bar going up.
Comments & Discussion