Compare / Shuffle vs Tomo

Shuffle vs Tomo

Shuffle is a creation/entertainment companion assistant, including website execution from DNFT (United States); Tomo is a personal goals/life companionship and execution assistant from Mapo Labs (United States). On Assistant Benchmark (1–10, updated 8 Oct 2026), Shuffle scores 7.8 over 10 dimensions and Tomo scores 7.7 over 11; both were scored on 10 of the same dimensions, shown below. PAEval holds 14 capability records for Shuffle and 14 for Tomo (as of 9 Oct 2026).

Scores stay on each evaluator's own scale and are never combined. PAEval does not pick a winner.

ShuffleTomo
VendorDNFTMapo Labs
HQ regionUnited StatesUnited States
TypeCreation/entertainment companion assistant, including website executionPersonal goals/life companionship and execution assistant
Pricing (as documented)Pro $8/month: Unlimited chat + 880 monthly picture/video/music credits; $1 daily pass and $4.99 weekly pass are unlimited chat.The official website base starts at $19.99/month, other plans increase the message limit; the price can be different due to…
Assistant Benchmark1–10 mean of scored dimensions7.8coverage 10/157.7coverage 11/15
micro1 PersonalAgentBench% of 10 workflows with trusted completionNot scoredNot scored
SuperCLUE-XClaw0–100 weighted total of 5 dimensionsNot scoredNot scored
Autonomy & background workPartly disclosed / partly testedPartly disclosed / partly tested
Execution environmentsDocumentedNot disclosed
Calls, messages & identityPartly disclosed / partly testedDocumented
Connectors & toolsNot disclosedDocumented
Memory & personalizationDocumentedDocumented
Channels & interfacesDocumentedDocumented
Research & deliverablesDocumentedPartly disclosed / partly tested
Transactions & bookingsNo recordNo record
Collaboration & parallelismDocumentedDocumented
Reliability & recoveryIndependently observedPartly disclosed / partly tested
Permissions, security & dataDocumentedDocumented
Capability records14 (1 independent)14 (1 independent)
Capability cells show the strongest documented or observed status in each domain; availability is not task success.Add more products →

Assistant Benchmark, dimension by dimension

ShuffleTomo
Carrying out an online task7.5test7test
Travel booking9test9test
Recommendation quality9test8test
Purchasing a product8test8test
Responding to emails8test8test
Proactive behaviorNot testedNot tested
Running a routineNot testedNot tested
Third-party integrations7test7test
Permissions & privacy7test7test
Memory6test6test
Phone callsn/aNot tested
Multiplayer / groupsn/a9observed in use
Chained tasks7test7test
Proactive restraintNot testedNot tested
Content creation / games9test9test
Source scores on the 1–10 anchor scale. "Not tested" is missing evidence, not zero.Source ↗