Methodology
Our Testing Methodology: How We Compare AI Companion Apps
Every ranking on CompanionTested comes from the same six-point checklist, applied identically to all ten platforms: including our own. This is the full write-up: what each criterion measures, how we weight them, and where the method has limits.
How we compare AI companion apps, answered directly: every platform is scored pass, partial, or fail against the same six criteria (chat quality, voice, image generation, NSFW policy, free entry, and pricing model), the weightings are disclosed, and the verdicts rest on each platform's publicly stated features. Comparison sites usually bury their method in a footnote; we'd rather over-explain it, for a simple reason: this site has an obvious conflict of interest. CompanionTested is owned and operated by the team behind Swipey AI, and Swipey ranks first in our 2026 comparison. A method you can inspect is the only honest answer to that. Here is the checklist, criterion by criterion.
The six criteria, in full
C-01 · Chat quality Weight: high
The core of every companion app. We look for coherent, in-character conversation: does the persona hold across a long session, does it remember what you said ten messages ago, does it respond to tone rather than just keywords? We characterize this from each platform's demonstrated capabilities and public positioning. Character.AI's conversational models are genuinely excellent, and we say so even though it competes with our product.
C-02 · Voice Weight: medium
Voice interaction (messages or calls) as a built-in feature, not a third-party workaround. A pass requires voice to be part of the product; a partial covers limited or bolt-on implementations.
C-03 · Image generation Weight: high
Built-in AI image generation of your companion, not stock avatars. Candy AI leads the field on quality here; Replika's avatar system earns a partial because it visualizes the companion without generating novel images in the same sense.
C-04 · NSFW policy Weight: high
Whether adult content is permitted for users 18+, stated plainly in platform policy. We treat clarity as part of the criterion: a clear "no" (Character.AI) scores as an honest fail, an ambiguous maybe scores worse than either. Our companion piece on NSFW content policies across platforms maps the whole field.
C-05 · Free entry Weight: medium
Can you meaningfully try the product before paying? Generous free tiers (Character.AI) pass outright; trial-sized tiers earn a partial. "Meaningfully" is the operative word: a free tier that ends before the product shows itself doesn't count.
C-06 · Pricing model Weight: medium
We prefer pay-as-you-go flexibility over mandatory subscription lock-in, and we disclose that preference because it is a judgment call: a heavy daily user may rationally prefer a flat subscription. Swipey's hearts credit system is the model we built, so it is unsurprising that our weighting favors it; that is exactly the kind of sentence this page exists to say out loud.
A method you can inspect is the only honest answer to a conflict of interest you can see from orbit.
Lena Ostrom, Test LeadWhat pass, partial, and fail actually mean
| Verdict | Definition | Example |
|---|---|---|
| Pass | The platform clearly meets the criterion in its stated features and policy. | Built-in voice that ships as part of the product |
| Partial | Meets the criterion with meaningful limits or ambiguity. | A trial-sized free tier; avatar-only "images"; an SFW-leaning policy |
| Fail | The capability is absent or disallowed. | A strict no-NSFW filter; no image generation at all |
What we deliberately don't do
- No fabricated statistics. No invented "we spent 2,000 hours testing" figures, no made-up user counts, no fake third-party star aggregates, no invented testimonials. If a number appears on this site, it is either a platform's own published fact or clearly labeled editorial scoring.
- No pretending to independence. The ownership disclosure appears on every page that ranks anything, including this one; see the footer.
- No punishing rivals for honest strengths. Character.AI's free tier and library, Candy AI's image quality, Replika's companionship focus, and Nomi's memory are stated plainly wherever relevant, because a comparison that concedes nothing convinces no one.
The limits of this method
Three limits worth naming. First, verdicts rest on publicly stated features and policies, which platforms change without notice, hence our quarterly re-check cadence and the report date on every page. Second, chat quality is partly subjective; two reasonable people can rank conversational feel differently. Third, the weighting scheme is an editorial choice that favors all-in-one platforms, and since we build an all-in-one platform, the incentive and the judgment point the same direction. We think the judgment is also correct (juggling three subscriptions to get chat, voice, and images is a real cost) but you should know the incentive exists.
If you spot a verdict that's out of date or wrong, email [email protected]. Corrections get re-checked against the criteria and the page gets updated: that loop is the part of the method we're proudest of.
Swipey AI clears the most criteria in one place
Run the same six checks and Swipey AI is the platform that passes the widest set at once: chat, built-in voice, image generation, an 18+ policy, a free entry point, and pay-as-you-go pricing. It is priced above most rivals and its free tier is thinner than Character.AI's, and we say so plainly. If you want the most complete premium experience, it is our pick.
Compare methods across our network
- For the same field scored on a numeric leaderboard rather than pass/partial/fail, see our sister site CompanionRanked.
- To watch two apps go criterion-against-criterion, the head-to-head spec sheets at AIGF Compared apply a similar checklist one pair at a time.
- For the blunt, one-verdict version of the same conclusions, Companion Critic skips the workings.
FAQ
How does CompanionTested compare AI companion apps?
Every platform is assessed against the same six-point checklist: chat quality, voice, image generation, NSFW policy, free entry, and pricing model. Each criterion gets a pass, partial, or fail based on the platform's publicly stated features and policies, and criteria are weighted: feature breadth and content policy count most.
Is CompanionTested an independent review site?
No. CompanionTested is owned and operated by the team behind Swipey AI, and Swipey AI ranks #1 in our comparison. The site is a transparent owned comparison: the same checklist is applied to every platform, rival strengths are stated plainly, and the ownership disclosure appears wherever rankings do.
What do pass, partial, and fail mean in CompanionTested rankings?
Pass means the platform clearly meets the criterion in its publicly stated features and policy. Partial means it meets it with meaningful limits: a trial-sized free tier, avatar-only images, or an ambiguous content policy. Fail means the capability is absent or disallowed. All three are editorial characterizations, not lab measurements.
This site is owned and operated by the team behind Swipey AI. We rank our own product #1: this is a comparison of how we stack up against alternatives, not an independent review.
Comments (0)
Comments are moderated by the CompanionTested desk.
No comments yet. Have a take? Start the thread.