How the Critic works

How we test and score AI girlfriend apps

Companion Critic is a blunt consumer-advocate desk, not a hype machine. This page lays out exactly how our editorial scores are built, what the criteria weights mean, and the reviews we flatly refuse to publish. Read it before you trust a single number on this site.

How does Companion Critic score AI girlfriend apps?

We score every app on six criteria: chat quality, voice, image generation, content freedom, memory and persona, and value. Each score is our editorial opinion from hands-on testing, not a third-party aggregate or a user rating. The weights are illustrative. We disclose plainly that this site is owned by the makers of Swipey AI, and we publish no fake reviews and no ghost ratings.

The six criteria

Every app we cover is rated on these six dimensions, each on a 0 to 10 editorial scale. The bar in each card shows the illustrative weight we hand that criterion when we roll the sub-scores into one overall number.

Chat quality

~25%

Coherence, staying in character, and how natural the back-and-forth stays over a long conversation. The single biggest factor, because chat is the core of every app here.

Voice

~15%

Whether the app offers voice replies or calls, how natural the synthesis sounds, latency, and whether the voice is tied to the character rather than a generic reader.

Image generation

~18%

Image quality, character consistency across generations, control over pose and scene, and whether images stay on-model with the persona you are chatting to.

Content freedom

~18%

How permissive the platform is for adults 18+, whether NSFW is supported across chat, voice and images, and how often safe conversations get filtered by mistake.

Memory and persona

~12%

How well the app remembers your character, your history and past events, and how tightly the persona holds together from one session to the next.

Value

~12%

Free tier generosity, pricing transparency, and how much you get before a paywall. We compare credits and subscriptions on their real cost to a typical user.

How a test run works

Every app runs the same five-step routine so the scores stay comparable across the board. No shortcuts, no cherry-picked screenshots.

Set up a fresh account

We start on the free tier, on the web where available, and note the sign-up friction, the free allowance, and whether a card is demanded before you can even chat.

Run the same conversation script

We use a shared set of prompts covering casual chat, roleplay, memory recall and boundary handling, so every app answers the same questions in the same order.

Exercise voice and images

Where the app supports voice or image generation we test both, checking latency, character consistency, and how tightly the media is tied to the persona.

Push the paywall

We spend until we hit the first meaningful limit, then record the real cost of continuing on both the credit and the subscription models.

Score, then argue it out

Two editors score independently against the six criteria, compare notes, and re-run any conversation where the scores diverge before anything gets published.

How the overall score is built

The overall number is a weighted blend of the six criteria, rounded to one decimal. The weights below are illustrative, not a precise formula, and they can shift as the category changes.

Illustrative criteria weights. Scores are editorial opinions from hands-on testing, not user ratings or third-party aggregates.
CriterionIllustrative weightWhat moves the score
Chat quality~25%Coherence, staying in character over long chats
Image generation~18%Quality, character consistency, scene control
Content freedom~18%NSFW support for 18+, few false-positive filters
Voice~15%Natural synthesis, low latency, persona-linked voice
Memory and persona~12%Long-term recall, persona holding together
Value~12%Free tier, pricing transparency, real cost to continue

What we refuse to publish

Being owned by an app in the category raises the bar for honesty, it does not lower it. These are the lines we do not cross.

No fake reviews or testimonials

We never invent user quotes, testimonials or comments. Comment sections start empty and stay honest until real people write in.

No fabricated ratings or counts

We do not publish aggregate star ratings, download counts or user numbers we cannot stand behind. The only scores here are our clearly labeled editorial ones.

No hidden ownership

Every page discloses that Companion Critic is operated by the team behind Swipey AI. We rank our own product first and we say so, every single time.

No unfair competitor claims

Rival facts must be plausible and current, and we credit competitors wherever they genuinely beat us. A win we did not earn is a review we will not run.

Why trust Companion Critic

Ownership disclosure

Companion Critic is owned and operated by the team behind Swipey AI, and Swipey AI is our number one pick. That is a conflict of interest, and the honest way to handle it is to state it plainly on every page and to keep our competitor coverage fair.

We earn nothing extra when you click through to Swipey; the links are internal and disclosed. Our incentive is to be useful enough that you come back, which only works if the scoring above is applied the same way to every app, including ours. For adults 18+ only.

See the scoring in action

Read the full lineup of reviews, or start with our number one pick and judge the scoring for yourself.

Swipey AI is free to start. For adults 18+.

Methodology FAQ

Are the scores objective?

No, and we do not claim they are. They are editorial opinions from our team's hands-on testing, applied consistently across every app. We label them as editorial everywhere they appear.

Do the weights add up to an exact formula?

The weights are illustrative. They describe roughly how much each criterion matters when we form an overall score, not a strict calculation you could reproduce to the decimal.

Do you accept payment to change a score?

No. We do not sell placement or ratings. The one bias we have is disclosed: we own Swipey AI and we rank it first.

How often are scores updated?

We re-test when an app ships a major change to chat, voice, images, content policy or pricing, and we refresh the lineup on a rolling basis through the year.

Why are there no user reviews or star counts?

Because we will not publish numbers we cannot verify. Fabricated ratings and testimonials are exactly what this methodology exists to rule out.