Compare / Meta Muse vs Asaply

Meta Muse vs Asaply

Meta Muse is a persistent personal agent from Meta (United States); Asaply is a local consumption/errand transaction type personal agent from Asaply Inc. (United States). On Assistant Benchmark (1–10, updated 8 Oct 2026), Meta Muse scores 9.3 over 8 dimensions and Asaply scores 7.5 over 2; both were scored on 2 of the same dimensions, shown below. PAEval holds 30 capability records for Meta Muse and 16 for Asaply (as of 9 Oct 2026).

Scores stay on each evaluator's own scale and are never combined. PAEval does not pick a winner.

Meta MuseAsaply
VendorMetaAsaply Inc.
HQ regionUnited StatesUnited States
TypePersistent personal agentLocal consumption/errand transaction type personal agent
Pricing (as documented)Free has a limit; Power is US$20/month, 500M Muse tokens/week; Maximum is US$100/month, 3B/week; automatic monthly renewal.Not compiled
Assistant Benchmark1–10 mean of scored dimensions9.3coverage 8/157.5coverage 2/15
micro1 PersonalAgentBench% of 10 workflows with trusted completion20%95% CI 6–51%Not scored
SuperCLUE-XClaw0–100 weighted total of 5 dimensionsNot scoredNot scored
Autonomy & background workDocumentedDocumented
Execution environmentsDocumentedNot disclosed
Calls, messages & identityLimited or conditionalLimited or conditional
Connectors & toolsDocumentedPartly disclosed / partly tested
Memory & personalizationDocumentedPartly disclosed / partly tested
Channels & interfacesDocumentedDocumented
Research & deliverablesDocumentedNot disclosed
Transactions & bookingsLimited or conditionalNo record
Collaboration & parallelismDocumentedLimited or conditional
Reliability & recoveryLimited or conditionalNot disclosed
Permissions, security & dataLimited or conditionalDocumented
Capability records30 (0 independent)16 (1 independent)
Capability cells show the strongest documented or observed status in each domain; availability is not task success.Add more products →

Assistant Benchmark, dimension by dimension

Meta MuseAsaply
Carrying out an online task9testNot tested
Travel booking7observed in usen/a
Recommendation quality9test7test
Purchasing a product10observed in use8test
Responding to emails10observed in usen/a
Proactive behaviorNot testedNot tested
Running a routine10observed in useNot tested
Third-party integrations10observed in useNot tested
Permissions & privacy9observed in useNot tested
MemoryNot testedNot tested
Phone callsNot testedn/a
Multiplayer / groupsNot testedNot tested
Chained tasksNot testedNot tested
Proactive restraintNot testedNot tested
Content creation / gamesNot testedn/a
Source scores on the 1–10 anchor scale. "Not tested" is missing evidence, not zero.Source ↗