Ai innovation
Grok Bot Roles: Own Outcomes Soft Pyramid Operators Can Trust
Grok Bot works for Soft Pyramid operators when each Bot owns a repeatable outcome, not a pile of open questions. Start with read-and-prepare work, review the result, then add approved actions. Eight starter roles show how.
The Slack thread was polite and useless. Someone had spun up a “general helper” Bot the night before and pointed half the Frisco leadership questions at it: research this account, draft that outreach, check the ad spend, summarize the stand-up, figure out why checkout felt slow. By morning the conversation looked busy. Nothing was finished enough to act on.
I’m Fakhar Khan, founder of Soft Pyramid in Frisco, Texas. Soft Pyramid already runs delivery with US scope in Frisco and engineering in Lahore, and we treat bots the way we treat people on a roster: Writer, SEO, SDR, EA. Each one owns a job with a repeatable finish line. A catch-all helper owns none of that. It owns a mood.
If you already run AI tools and you manage operators, you know the feeling. The model answers. The work still sits on your desk. That gap is the whole point of Grok Bot, and it is also where most teams get the first week wrong.
Outcome first, questions later
Cursor’s own guidance is blunt: the best Bot roles own a repeatable outcome, not a loose category of questions. That matches how Soft Pyramid staffs work. You do not hire “someone who knows marketing.” You hire for a weekly deliverable someone can review without you rewriting the brief every Tuesday.
Grok Bot is Cursor’s durable teammate product: named Bots on a persistent cloud computer with a browser, filesystem, and terminal, so tasks finish in your real tools instead of dying as chat drafts. The get-started path is simple on purpose. Install, sign in with your Cursor account, create a Bot with one primary job, give it a first task that stops at review, then correct until the process sticks.
Would you trust a new hire who could touch CRM, ads, and production on day one with no written stop line? Same rule here.
Start every role with read-and-prepare work. Review the artifact. Only then add approved actions or a routine. Sending, publishing, purchasing, and production changes stay behind approval. That is not paranoia. It is the same gate I argue for when production AI systems need more than a demo: identity, a failure story, and a human who can say no.
Ask for results you can review without babysitting. Cursor’s day-to-day guidance on working with Grok Bot is clear on this. Name the artifact, keep source links and screenshots with the claim, separate facts from hypotheses, and put durable files in the shared workspace so the next run does not start from a screenshot alone. Files and results are part of the job, not a sidebar trick.
Eight roles that earn a seat on the roster
I am not going to dump vendor prompts at you. Below is how I would brief Soft Pyramid operators if we were standing a Grok Bot roster next to the humans we already trust. Each role ends at a review point on purpose.
Sales Outbound. Owns account research, contact prioritization, and outreach drafts that wait for you. Connect CRM, product-intent sources, and email. Give it a view of twenty-five accounts, your ICP, and style examples. Ask it to score fit and intent, name up to three contacts per account, draft email and LinkedIn copy, skip anyone already in an active sequence, and return a review list. Do not let it send or enroll. When the list is boringly consistent, a nightly research routine that still stops at review is earned, not assumed.
Talent Scout. Owns sourcing, candidate research, and outreach drafts. Connect ATS, approved sourcing tools, email, and calendar. Hand it a role description with must-haves. Ask for twenty candidates, exclude people already in the ATS, show the evidence for each match, and draft personalized outreach in your voice. Do not contact anyone. Candidate privacy and each source’s terms are standing boundaries, not footnotes.
Paid Media. Owns campaign monitoring and budget recommendations. Connect ad platforms, analytics, the budget sheet, and Slack. Ask for current spend and performance by campaign, a comparison against monthly budget and target CAC, reallocations with the numbers attached, and a draft Slack update for growth. Do not change budgets or send the message. Even after the analysis becomes a routine, campaign writes stay behind approval.
Expense Manager. Owns weekly reconciliation and missing-information follow-up. Connect the expense system, finance inbox, and policy spreadsheet. Ask for this week’s summary, receipt matching, flagged categories or policy exceptions with a policy citation, and one draft follow-up per owner. Totals should reconcile back to the source. Do not send messages or change reimbursements.
Product Performance. Owns targeted investigations with evidence. Connect observability, analytics, and incident tooling. Point it at a concrete symptom (“checkout latency since yesterday’s release”), ask it to review dashboards and traces, pick the highest-confidence hotspot, and return a short write-up with screenshots and direct links. Separate facts from hypotheses. Do not change alerts or production settings. A recurring health report can be a routine. Unsupervised production changes cannot.
Bug Reproduction. Owns turning reports into reliable reproduction packs. Connect the issue tracker and staging. Ask it to reproduce in staging with a fresh test account, then return exact steps, expected versus actual behavior, screenshots, browser and OS details, and a minimal test case when possible. Do not use production customer data. Hand test credentials through a secure path, not chat.
Account Health. Owns risk and expansion signals across a portfolio. Connect CRM, product usage, support, and billing. Ask for a ranked watch list that combines usage, escalations, renewal timing, and stakeholder activity, with evidence, why it matters, and a suggested next step for each account. Do not contact customers or edit the CRM. Put risk thresholds in the Bot description so the weekly result stays consistent instead of drifting with whoever messaged last.
Chief of Staff. Owns a source-linked digest of what changed and what needs your attention. Connect Slack, email, calendar, and your planning doc. Ask it to review activity since yesterday across approved channels and return only items that map to the priorities in that doc. For each item: source, why it matters, proposed next step, and whether you owe a decision. Do not send messages or change meetings. Mark what was useful and what was noise, then schedule the digest for a time you will actually open it.
Notice the pattern. Research, draft, pack evidence, stop. That is the Soft Pyramid instinct we already use on n8n and agent delivery: what survives after the demo is retries, credentials, and a human gate, not another clever prompt.
Approvals are the operating system
Before any Bot changes an external system, read Cursor’s approvals and Auto Review notes. The strongest boundary is still the one in the request. State the current value, the proposed value, and the expected impact, then ask the Bot to stop. Approvals control the proposed action, not work already done. Do not approve what you cannot identify.
Auto Review sits behind those prompts and can let an action proceed, require approval, or deny it. Treat it as a complement to least privilege, not a substitute. Credentials stay with you. The Bot hands over the computer for passwords and two-factor. Passwords and one-time codes never belong in ordinary chat.
This is the same maturity move as climbing from Tab completion toward agents with review capacity, which I wrote about in Cursor 101. Speed without a stop line is just a faster way to ship the wrong write.
Turn an example into a durable Bot
When one of those roles starts producing work you would hand a junior operator, make it durable. Cursor’s closing pattern is the one I would put on Soft Pyramid’s wall:
- Put the job, source systems, output format, and standing boundaries in the Bot description.
- Run one real task with a safe scope.
- Correct until the result is reviewable without you rewriting the brief.
- Save the working process as a skill.
- Test that skill on a second input before you trust it.
- Create a routine only after retries and failure cases are defined.
- Keep consequential external actions (sends, purchases, publishes, production writes) behind an approval card you can read in one glance.
Skills capture how. Routines capture when. If a source is missing, the Bot should report the failure instead of inventing yesterday’s data. That single rule saves more trust than any clever schedule.
How I would start this week
Pick one outcome that already eats an hour of operator time every week. Account research for outbound. Expense exceptions. A staging reproduction pack. Give the Bot a name, a job sentence, and a stop line. Attach the sources. Ask for a reviewable artifact. Sit with the first two runs the way you would sit with a new hire’s first two tickets.
If you want a guided path instead of improvising the roster alone, Soft Pyramid’s Grok Bot training walks operators through that same sequence: role design, safe-scope tasks, skills, and approval boundaries before anyone turns on a routine. Use it when the team needs a shared language more than another Slack thread of half-finished asks.
The general helper will keep looking busy. The Bot that owns one outcome will start looking like a teammate. That is the turn Soft Pyramid already made with Writer, SEO, SDR, and EA. Grok Bot is just the next place to apply the same staffing rule with a clearer stop line.
Fakhar Khan
If this is the problem you are staring at, let's talk about it.
Architecture, AI operations, and delivery for US small and mid-size companies — outcomes first.