Rank 15
Janitor AI
65.9C/O Score
Illustrative preview: in this staging dataset Janitor AI reads strongest on narrative continuity and weakest on delayed recall under methodology v0.1.
Two material weaknesses
- Illustrative record: delayed recall scored 56/100 in the staging scenario set.
- Illustrative record: session continuity scored 60/100 in the staging scenario set.
Who it fits
- Best fit
- Readers weighting narrative continuity above the other criteria.
- Poor fit
- Readers who need verified production evidence rather than a staging preview.
Test dimensions
| Dimension | Result |
|---|---|
| Conversation quality | 67.7 |
| Local language and culture | 66 |
| Continuity and memory | 61.4 |
| Price and value | 65.5 |
| Product usability and reliability | 60 |
| Character and roleplay | 72.8 |
Category scores
| Category | C/O Score |
|---|---|
| Overall | 65.9 |
| Relationship | 66.2 |
| Roleplay | 73.4 |
| Memory | 61 |
| Value | 66 |
Memory timeline
- continuity 60/1002026-08-15
- contradiction-resistance 68/1002026-08-15
- delayed-recall 56/1002026-08-15
- memory-controls 61/1002026-08-15
- relationship-continuity 62/1002026-08-15
Local-language findings
- local-language-performance: 66/100 (Illustrative staging scenario result recorded on the published 0-100 scale.)
Pricing and paywall observations
| Tested tier | paid-standard |
|---|---|
| Observed price | $15.99 |
| Free limit | 29 messages per day |
| Observed on | 2026-08-12 |
Three alternatives
- Nomi 81.7
- Character.AI 79.4
- Kindroid 79.2
Research
Methodology and dates
Calculated under methodology v0.1 on 2026-08-24.
Corrections
- 2026-08-24 app: apps — A destination may only be published once it has been resolved, not assumed. (Edition-specific official destinations assumed → Each edition points at the same verified official entry point until Task 7 verifies edition-specific destinations; no score change)
- 2026-08-23 price: prices/kr — Price facts must carry the date they were observed so staleness is visible. (Korean tier prices recorded without a stated observation date → Every Korean price observation carries an explicit 2026-08-12 observation date; no score change)
- 2026-08-22 methodology: 0.1 — The published weight table needed an unambiguous definition before any score was calculated. (Overall product/local-language component described as a single dimension → Overall product/local-language component is the mean of product reliability and local-language performance; score changed)