Compare / OpenAI dots vs Ollie
OpenAI dots vs Ollie
OpenAI dots is a persistent personal agent from OpenAI (United States); Ollie is a home personal assistant from Confabulation Corporation (United States). On Assistant Benchmark (1–10, updated 8 Oct 2026), OpenAI dots scores 8.4 over 7 dimensions and Ollie scores 8.3 over 13; both were scored on 6 of the same dimensions, shown below. PAEval holds 31 capability records for OpenAI dots and 16 for Ollie (as of 9 Oct 2026).
Scores stay on each evaluator's own scale and are never combined. PAEval does not pick a winner.
| OpenAI dots | Ollie | |
|---|---|---|
| Vendor | OpenAI | Confabulation Corporation |
| HQ region | United States | United States |
| Type | Persistent personal agent | Home personal assistant |
| Pricing (as documented) | The first dot is included in Pro, with no independent surcharge; Pro is currently 100/200/500 US dollars per month, and 500 includes… | Public index price: Free 6 items per day + 50 items per month; Everyday $25/month, 10 per day + 150 per month; Always-On $100/month, 15… |
| Assistant Benchmark1–10 mean of scored dimensions | 8.4coverage 7/15 | 8.3coverage 13/15 |
| micro1 PersonalAgentBench% of 10 workflows with trusted completion | Not scored | Not scored |
| SuperCLUE-XClaw0–100 weighted total of 5 dimensions | Not scored | Not scored |
| Autonomy & background work | Documented | Independently observed |
| Execution environments | Documented | Partly disclosed / partly tested |
| Calls, messages & identity | Limited or conditional | Partly disclosed / partly tested |
| Connectors & tools | Limited or conditional | Documented |
| Memory & personalization | Documented | Documented |
| Channels & interfaces | Documented | Documented |
| Research & deliverables | Limited or conditional | Partly disclosed / partly tested |
| Transactions & bookings | Limited or conditional | No record |
| Collaboration & parallelism | Limited or conditional | Documented |
| Reliability & recovery | Documented | Not disclosed |
| Permissions, security & data | Documented | Documented |
| Capability records | 31 (0 independent) | 16 (2 independent) |
Capability cells show the strongest documented or observed status in each domain; availability is not task success.Add more products →
Assistant Benchmark, dimension by dimension
| OpenAI dots | Ollie | |
|---|---|---|
| Carrying out an online task | 9test | 9test |
| Travel booking | Not tested | 7test |
| Recommendation quality | 9test | 7test |
| Purchasing a product | 6observed in use | 8test |
| Responding to emails | 10test | 8test |
| Proactive behavior | 8observed in use | 8test |
| Running a routine | Not tested | 10test |
| Third-party integrations | Not tested | 7test |
| Permissions & privacy | 8observed in use | 8test |
| Memory | 9observed in use | Not tested |
| Phone calls | n/a | n/a |
| Multiplayer / groups | Not tested | 9test |
| Chained tasks | Not tested | 8test |
| Proactive restraint | Not tested | 10test |
| Content creation / games | Not tested | 9test |
Source scores on the 1–10 anchor scale. "Not tested" is missing evidence, not zero.Source ↗