Meet Caddi in personADVISE AIOct 20–22AI for Mid-Sized LawNov 5Legal InnovatorsNov 17–18
BlogAlways-on AI agents

Grok Bot vs. Muse vs. Dots for Professional Services Firms

xAI, Meta, and OpenAI each shipped an always-on agent this fall. Here is how they compare on the questions a law firm, RIA, or accounting firm has to answer first.

Grok Bot, Muse, and dots are all always-on agents with their own cloud computer, and all three plan each task with a model. Grok Bot has the most developed firm controls on its Enterprise plan. Muse is a consumer agent with a careful permission system and no business admin. dots have the widest app reach and layered action rules. None of them claims repeatable runs or a firm-level books-and-records archive, which is what repeated client work needs.

In about seven weeks, the three largest consumer AI companies shipped the same product. xAI launched Grok Bot in beta on August 11. Meta launched Muse on September 8. OpenAI announced dots at DevDay on September 29. Each one gives you an agent with its own computer and browser that signs into your apps and works toward goals in the background, asking before it does anything sensitive.

For professional services firms, the interesting question is not which one is smartest. It is which one a firm could supervise, and for which work. This comparison sticks to what the vendors themselves publish, as of October 1, 2026.

At a glance

Grok BotMusedots
VendorxAI, built and run with CursorMetaOpenAI
LaunchedAugust 11, 2026 (beta)September 8, 2026September 29, 2026
ModelChosen by Cursor; no model pickerMuse SparkGPT-6 Astra
PlansSuperGrok, Cursor Pro, Cursor Teams; Enterprise via salesFree, Power ($20/mo), Maximum ($100/mo)ChatGPT Pro (from $100/mo), Business Premium; Enterprise beta
Firm adminEnterprise: SSO, SCIM, enforced review rules, network policyNone describedEnterprise admin toggle; specialist-dot identity in pilot
Saved workSkills from a recorded demo, run by routinesGoals and plans; schedulesGoals, Custom Rules, scheduled tasks
Action controlsAllow once / Always allow / Deny; Auto ReviewSentinel: allow, deny, or ask; grants can be perpetualCustom Rules; Auto-review; hard stops on deletes and transfers
Record of activityEnterprise audit log; Action Recording off by default, 90 daysPer-user activity logPer-user Activity View
Training defaultNot used with Privacy Mode onUsed, sanitized; opt-outNot used on Business and Enterprise
Repeatable runs claimed?NoNoNo
From each vendor's own announcement, help, pricing, and security pages, as of October 1, 2026.

Will they do the work the same way every time?

None of the three says so. Grok Bot comes closest: you can teach a Bot a task by recording up to ten minutes of yourself doing it, and it turns that into a draft skill that a routine can run on a schedule. But a skill is a set of instructions the model carries out each time, and xAI's docs tell you to re-test skills when a website, plugin, or file format changes. Muse is goal-driven and describes only its approval cards and credential storage as deterministic. dots work from a goal plus Custom Rules, and OpenAI notes a dot can make mistakes even when following them.

That is fine for a person's own work. For a procedure a firm runs two hundred times a month, like opening accounts or matters, it means each run is decided again, so testing last month's cases does not tell you how the next one will go.

Who approves what?

All three vendors put real work into action controls, and all three say they deliberately avoid asking about everything, because people stop reading prompts. Grok Bot offers Allow once, Always allow, or Deny, with an Auto Review model checking risky actions; on Enterprise, admins can enforce review rules for the whole team. Muse routes every connector action and network request through Sentinel, a separate agent that allows, denies, or asks, with grants that can last one time or forever. dots let you set each kind of action to proceed, proceed if pre-approved, ask first, or hand off, with Auto-review checking planned actions and hard stops on deletes, security access, password changes, and money transfers.

These controls supervise individual actions. A law firm under ABA Model Rule 5.3, or an RIA under the SEC compliance rule, also has to supervise the procedure: what the agent is supposed to do, and whether it did that. Approval prompts do not answer that question on their own.

Can the firm show what the agent did?

Grok Bot is the only one with a firm-level record today, on its Enterprise plan: audit logs that can stream to a SIEM, plus Action Recording of what Bots did, which is off by default and kept for 90 days. Self-serve Teams do not get the audit log. Muse gives each user an activity log of what it has done and plans to do. dots give each user an Activity View of tasks in progress, scheduled, and completed.

For an RIA, SEC Rule 204-2 requires keeping copies of certain written communications, generally for five years. For a law firm, the question comes from clients, auditors, and courts. In both cases a per-user activity feed is a starting point, not a record the firm can retain and produce.

What happens to client data?

Defaults differ more than you might expect. Grok Bot does not train on customer data with Privacy Mode on, runs its computers in the United States, and has no per-organization retention policy yet. Muse trains on sanitized activity by default with an opt-out, and Meta says its current architecture does not prevent Meta from accessing data to operate the service, with a user-keyed Confidential VM planned. dots on Business, Enterprise, and Edu plans do not train on content by default, and disconnecting an app does not delete what a dot already pulled from it.

All three vendors are candid that prompt injection is not solved. For agents holding a lawyer's or adviser's inbox and file access, that sits squarely inside confidentiality duties.

Which one fits which work?

For a person's own varied work, all three are impressive, and the choice mostly follows which ecosystem your firm already lives in. For a firm rollout today, Grok Bot for Enterprise has the most complete admin layer. dots have the widest app reach. Muse is built for individuals and small businesses rather than firms.

For repeated client work, the answer is the same for all three: pair them with an agent the firm can verify. That means a hybrid agent that runs the workflow as code, uses AI for the judgment steps inside it, logs each one, and leaves a record for every run. That is what Caddi builds for law firms, RIAs, and professional services firms: your team shows the workflow on a screen share or describes it in chat, and Caddi runs it across your tools, changing it in plain English when the work changes. Across Caddi customers, 99% of agent runs complete successfully.

Grok Bot, Muse, and dots show what an agent can do inside your apps. A professional services firm should judge them on what it can check. Use them as personal assistants, and put repeated client work on a verifiable agent with a procedure you can review and a record for every run.

Guides by firm type

Keep reading

Caddi

See how Caddi AI Agents can run your firm's repeated client work the same way every time and log every run for review

Or sign up free

Frequently asked questions

What is the difference between Grok Bot, Muse, and dots?

All three are always-on agents with their own cloud computer that work toward your goals in the background. Grok Bot (xAI, built with Cursor) is organized around named Bots, recorded skills, and routines, with the most developed admin controls on its Enterprise plan. Muse (Meta) is a consumer personal agent running in an isolated VM, governed by a separate permission agent called Sentinel. dots (OpenAI) run on GPT-6 Astra, connect to over 4,000 apps through ChatGPT, and are shaped by Custom Rules and an Auto-review system. Details as of October 1, 2026.

Which always-on agent is best for a law firm or an RIA?

For firm use today, Grok Bot for Enterprise has the most complete admin layer: SSO, SCIM, enforced review rules, audit logs, and Action Recording. dots are available to Business Premium and as an admin-enabled Enterprise beta. Muse has no business admin plan. None of the three claims repeatable runs or offers a books-and-records archive, so for repeated client work most firms will pair a personal agent with a verifiable agent that runs the procedure as code.

Do Grok Bot, Muse, or dots run workflows the same way every time?

None of the vendors claims that. All three plan each task with a model. Grok Bot's skills are the closest to saved workflows, but a skill is a set of instructions the model carries out on each run, and xAI's docs say to re-test skills when a site or format changes. Muse is goal-driven and calls only its approval cards and credential storage deterministic. dots work from goals plus Custom Rules.

Do these agents keep an audit trail?

Each keeps some record. Grok Bot's audit logs and Action Recording are Enterprise features, and Action Recording is off by default with 90-day retention. Muse shows each user an activity log with no stated admin export. dots show each user an Activity View, and OpenAI's materials do not describe an exportable audit log. For SEC- or bar-regulated work, check whether the record can be retained and produced at the firm level.

Do Grok Bot, Muse, or dots train on my data?

It depends on the product and plan. Grok Bot with Privacy Mode on does not use customer data for training. Muse uses sanitized activity for training by default, with an opt-out. dots on Business, Enterprise, and Edu plans do not train on content by default; on personal plans it follows the Improve the model setting.

What is a verifiable AI agent?

An agent whose work someone other than the agent can check: the procedure can be read before it runs, the same input takes the same path, model judgment sits in named and logged steps, every run leaves a record tied to source, and a named person owns exceptions and changes. Hybrid agents, which run the workflow as code and use AI for the judgment steps, are how you build one.