Compare / OpenAI dots vs Grok Bot

OpenAI dots vs Grok Bot

OpenAI dots is a persistent personal agent from OpenAI (United States); Grok Bot is an Individual/Team Agent from xAI (via Cursor) (United States). On Assistant Benchmark (1–10, updated 8 Oct 2026), OpenAI dots scores 8.4 over 7 dimensions and Grok Bot scores 7.3 over 7; both were scored on 4 of the same dimensions, shown below. PAEval holds 31 capability records for OpenAI dots and 33 for Grok Bot (as of 9 Oct 2026).

Scores stay on each evaluator's own scale and are never combined. PAEval does not pick a winner.

OpenAI dotsGrok Bot
VendorOpenAIxAI (via Cursor)
HQ regionUnited StatesUnited States
TypePersistent personal agentIndividual/Team Agent
Pricing (as documented)The first dot is included in Pro, with no independent surcharge; Pro is currently 100/200/500 US dollars per month, and 500 includes…Cursor Pro/Pro+/Ultra are priced at US$20/60/200 per month respectively; Teams Standard/Premium are US$40/120 per user/month respectively,…
Assistant Benchmark1–10 mean of scored dimensions8.4coverage 7/157.3coverage 7/15
micro1 PersonalAgentBench% of 10 workflows with trusted completionNot scored30%95% CI 11–60%
SuperCLUE-XClaw0–100 weighted total of 5 dimensionsNot scoredNot scored
Autonomy & background workDocumentedLimited or conditional
Execution environmentsDocumentedDocumented
Calls, messages & identityLimited or conditionalLimited or conditional
Connectors & toolsLimited or conditionalLimited or conditional
Memory & personalizationDocumentedDocumented
Channels & interfacesDocumentedDocumented
Research & deliverablesLimited or conditionalDocumented
Transactions & bookingsLimited or conditionalLimited or conditional
Collaboration & parallelismLimited or conditionalLimited or conditional
Reliability & recoveryDocumentedLimited or conditional
Permissions, security & dataDocumentedLimited or conditional
Capability records31 (0 independent)33 (0 independent)
Capability cells show the strongest documented or observed status in each domain; availability is not task success.Add more products →

Assistant Benchmark, dimension by dimension

OpenAI dotsGrok Bot
Carrying out an online task9test8observed in use
Travel bookingNot testedNot tested
Recommendation quality9test7observed in use
Purchasing a product6observed in use8observed in use
Responding to emails10test9observed in use
Proactive behavior8observed in useNot tested
Running a routineNot testedNot tested
Third-party integrationsNot tested3observed in use
Permissions & privacy8observed in useNot tested
Memory9observed in useNot tested
Phone callsn/aNot tested
Multiplayer / groupsNot testedNot tested
Chained tasksNot tested8observed in use
Proactive restraintNot testedNot tested
Content creation / gamesNot tested8observed in use
Source scores on the 1–10 anchor scale. "Not tested" is missing evidence, not zero.Source ↗