Compare / OpenAI dots vs Shuffle

OpenAI dots vs Shuffle

OpenAI dots is a persistent personal agent from OpenAI (United States); Shuffle is a creation/entertainment companion assistant, including website execution from DNFT (United States). On Assistant Benchmark (1–10, updated 8 Oct 2026), OpenAI dots scores 8.4 over 7 dimensions and Shuffle scores 7.8 over 10; both were scored on 6 of the same dimensions, shown below. PAEval holds 31 capability records for OpenAI dots and 14 for Shuffle (as of 9 Oct 2026).

Scores stay on each evaluator's own scale and are never combined. PAEval does not pick a winner.

OpenAI dotsShuffle
VendorOpenAIDNFT
HQ regionUnited StatesUnited States
TypePersistent personal agentCreation/entertainment companion assistant, including website execution
Pricing (as documented)The first dot is included in Pro, with no independent surcharge; Pro is currently 100/200/500 US dollars per month, and 500 includes…Pro $8/month: Unlimited chat + 880 monthly picture/video/music credits; $1 daily pass and $4.99 weekly pass are unlimited chat.
Assistant Benchmark1–10 mean of scored dimensions8.4coverage 7/157.8coverage 10/15
micro1 PersonalAgentBench% of 10 workflows with trusted completionNot scoredNot scored
SuperCLUE-XClaw0–100 weighted total of 5 dimensionsNot scoredNot scored
Autonomy & background workDocumentedPartly disclosed / partly tested
Execution environmentsDocumentedDocumented
Calls, messages & identityLimited or conditionalPartly disclosed / partly tested
Connectors & toolsLimited or conditionalNot disclosed
Memory & personalizationDocumentedDocumented
Channels & interfacesDocumentedDocumented
Research & deliverablesLimited or conditionalDocumented
Transactions & bookingsLimited or conditionalNo record
Collaboration & parallelismLimited or conditionalDocumented
Reliability & recoveryDocumentedIndependently observed
Permissions, security & dataDocumentedDocumented
Capability records31 (0 independent)14 (1 independent)
Capability cells show the strongest documented or observed status in each domain; availability is not task success.Add more products →

Assistant Benchmark, dimension by dimension

OpenAI dotsShuffle
Carrying out an online task9test7.5test
Travel bookingNot tested9test
Recommendation quality9test9test
Purchasing a product6observed in use8test
Responding to emails10test8test
Proactive behavior8observed in useNot tested
Running a routineNot testedNot tested
Third-party integrationsNot tested7test
Permissions & privacy8observed in use7test
Memory9observed in use6test
Phone callsn/an/a
Multiplayer / groupsNot testedn/a
Chained tasksNot tested7test
Proactive restraintNot testedNot tested
Content creation / gamesNot tested9test
Source scores on the 1–10 anchor scale. "Not tested" is missing evidence, not zero.Source ↗