How the desk works
How we test and score AI girlfriend apps
AI Romance Report is a comparison publication, not a wire service. This page sets out exactly how our editorial scores are formed, what the criteria weights mean, and the reviews we refuse to print. Read it before you trust a single number on this site.
Updated July 2026
How does AI Romance Report score AI companion apps?
We score every platform on six criteria: chat quality, image generation, content freedom, memory and persona, voice, and value. Each figure is our editorial opinion from hands-on testing, not a third-party aggregate or a reader rating. The weights are illustrative. We disclose that this publication is owned by the team behind Swipey AI, and we print no fabricated reviews.
The six criteria
Every platform is rated on these six dimensions, each on a 0 to 10 editorial scale. The bar in each card shows the illustrative weight we give that criterion when combining sub-scores into an overall figure.
Chat quality
25%Coherence, memory across a session, staying in character, and how natural a long back and forth feels. Chat is the core of every platform here, so it carries the most weight.
Image generation
18%Image quality, character consistency across generations, scene control, and whether the picture stays on model with the persona you are talking to.
Content freedom
17%How permissive a platform is for adults 18+, whether NSFW is supported across chat, voice and images, and how often ordinary messages get filtered by mistake.
Memory and persona
15%Whether a companion remembers earlier conversations, keeps a stable personality, and carries continuity across days rather than resetting each session.
Voice
13%Whether voice replies or calls exist, how natural the synthesis sounds, latency, and whether the voice is tied to the character rather than a generic reader.
Value
12%Free tier generosity, pricing transparency, and how much you get before a paywall. We compare credits and subscriptions on their real cost to a typical reader.
How a test run works
Each platform goes through the same five-step routine, so the scores stay comparable across the whole field.
Open a fresh account
We start on the free tier, in the browser where possible, and note the sign-up friction, the free allowance, and whether a card is required before you can send a first message.
Run the same script
Every platform answers the same set of prompts covering casual chat, roleplay, memory recall and boundary handling, in the same order, so the responses are comparable.
Exercise voice and images
Where voice or image generation exist we test both, checking latency, character consistency, and how tightly the media is tied to the persona rather than bolted on.
Push the paywall
We keep going until we hit the first real limit, then record the actual cost of continuing on both the credit and the subscription models.
Score, then re-check
Two editors score independently against the six criteria, compare notes, and re-run any conversation where the two scores disagree before anything is published.
How the overall score is built
The overall figure is a weighted blend of the six criteria, rounded to one decimal. The weights below are illustrative, not a precise formula, and they shift as the category changes.
| Criterion | Illustrative weight | What moves the score |
|---|---|---|
| Chat quality | 25% | Coherence, memory, staying in character over long chats |
| Image generation | 18% | Quality, character consistency, scene control |
| Content freedom | 17% | NSFW support for 18+, few false-positive filters |
| Memory and persona | 15% | Continuity across sessions, a stable personality |
| Voice | 13% | Natural synthesis, low latency, persona-linked voice |
| Value | 12% | Free tier, pricing transparency, real cost to continue |
What we refuse to print
Being owned by a platform in the category raises the bar for honesty rather than lowering it. These are the lines the desk does not cross.
No invented readers or letters
We never fabricate user quotes, testimonials or comments. Comment sections start empty and stay honest until real readers write to the desk.
No fabricated ratings or counts
We do not print aggregate star ratings, download totals or user numbers we cannot stand behind. The only scores here are our clearly labelled editorial ones.
No hidden ownership
Every page states that AI Romance Report is operated by the team behind Swipey AI. We rank our own product first and we say so, in plain sentences, every time.
No unfair competitor claims
Rival facts must be plausible and current, and we credit competitors wherever they genuinely beat us. A win we did not earn is a verdict we will not print.
Why trust AI Romance Report
Ownership disclosure
AI Romance Report is owned and operated by the team behind Swipey AI, and Swipey AI is our number one pick. That is a conflict of interest, and the honest way to handle it is to state it plainly on every page and to keep our competitor coverage fair.
We earn nothing extra when you click through to Swipey; the links are internal and disclosed. Our incentive is to be useful enough that you come back, which only works if the scoring above is applied the same way to every platform, ours included. For adults 18+ only.
See the method in action
Read the full field, or start with our number one pick and judge the scoring for yourself.
Swipey AI is free to start. For adults 18+.
Methodology FAQ
Are the scores objective?
No, and we do not claim they are. They are editorial opinions from our hands-on testing, applied the same way to every platform, including our own. We label them editorial wherever they appear.
Do the weights add up to an exact formula?
The weights are illustrative. They describe roughly how much each criterion matters when we form an overall figure, not a strict calculation you could reproduce to the decimal.
Do you accept payment to change a score?
No. We do not sell placement or ratings. The one bias we carry is disclosed at the top and bottom of every page: we own Swipey AI and rank it first.
How often are the scores updated?
We re-test when a platform ships a major change to chat, voice, images, content policy or pricing, and refresh the field on a rolling basis through the year. The dateline shows the last full pass.
Why are there no reader ratings or star counts?
Because we will not print numbers we cannot verify. Fabricated ratings and testimonials are exactly what this methodology exists to rule out.