How we test and score AI girlfriend apps
The AIGF List is an editorial ranking, not a spreadsheet of borrowed numbers. This page sets out exactly how our scores are formed, what the criteria weights mean, and the kind of reviewing we refuse to publish. Read it before you trust a single figure on this site.
Updated July 2026
How does The AIGF List score AI girlfriend apps?
We judge every app on five criteria: conversation quality, voice, image generation, content policy, and value. Each score is our editorial opinion from hands-on use, not a third-party aggregate or a user rating. The weights are illustrative. We disclose plainly that this site is made by the team behind Swipey AI, and we publish no fabricated ratings or testimonials.
The five criteria
Every app is read on these five dimensions, each on a 0 to 10 editorial scale. The bar in each card shows the illustrative weight we give that criterion when we settle on an overall figure.
Conversation
~30%Coherence, memory within a session, staying in character, and how natural a long exchange feels. The single heaviest factor, because conversation is the product.
Image generation
~20%Image quality, character consistency between generations, control over pose and scene, and whether pictures stay on-model with the persona you are talking to.
Content policy
~20%How permissive the platform is for adults 18 and over, whether NSFW is supported across chat, voice and images, and how often ordinary conversation gets filtered by mistake.
Voice
~15%Whether the app offers spoken replies or calls, how natural the synthesis sounds, latency, and whether the voice is tied to the character rather than a generic reader.
Value
~15%Free tier generosity, how transparent the pricing is, and how much you get before a paywall. We weigh credits and subscriptions on their real cost to a typical reader.
How a test run works
Each app goes through the same routine so that the scores stay comparable across the whole list.
Set up a fresh account
We begin on the free tier, in the browser where possible, and note the sign-up friction, the free allowance, and whether a card is required before you can talk at all.
Run the same conversation script
We use a shared set of prompts covering casual talk, roleplay, memory recall and boundary handling, so every app answers the same questions in the same order.
Exercise voice and images
Where an app supports voice or image generation we test both, checking latency, character consistency, and how tightly the media is bound to the persona.
Push the paywall
We spend until we reach the first real limit, then record the actual cost of continuing on both the credit and the subscription models.
Score, then re-check
Two editors read independently against the five criteria, compare notes, and re-run any conversation where their scores diverge before anything is published.
How the overall figure is built
The overall number is a weighted blend of the five criteria, rounded to one decimal. The weights below are illustrative rather than a precise formula, and they shift as the category changes.
| Criterion | Illustrative weight | What moves the score |
|---|---|---|
| Conversation | ~30% | Coherence, memory, staying in character over long chats |
| Image generation | ~20% | Quality, character consistency, scene control |
| Content policy | ~20% | NSFW support for adults 18+, few false-positive filters |
| Voice | ~15% | Natural synthesis, low latency, persona-linked voice |
| Value | ~15% | Free tier, pricing transparency, real cost to continue |
Tip Read our scores by your own weights, not ours. If budget is your constraint, a strong Value score should outrank a high Image score; if you want adult content, Content policy matters most. The overall figure is a starting point, not the whole answer.
What we refuse to publish
Being made by an app in the category raises the bar for honesty rather than lowering it. These are the lines we will not cross.
No fabricated reviews or testimonials
We never invent user quotes, testimonials or comments. Comment sections start empty and stay honest until real readers write in.
No fabricated ratings or counts
We do not publish aggregate star ratings, download totals or user numbers we cannot verify. The only scores here are our clearly labeled editorial ones.
No hidden ownership
Every page states that The AIGF List is made by the team behind Swipey AI. We rank our own product first, and we say so, every time.
No unfair competitor claims
Rival facts must be plausible and current, and we credit competitors wherever they genuinely beat us. A win we did not earn is a review we will not run.
Why trust The AIGF List
Ownership disclosure
The AIGF List is owned and operated by the team behind Swipey AI, and Swipey AI is our number one pick. That is a conflict of interest, and the honest way to handle it is to state it plainly on every page and to keep our competitor coverage fair.
We earn nothing extra when you click through to Swipey; the links are internal and disclosed. Our incentive is to be useful enough that you come back, which only works if the scoring above is applied the same way to every app, including our own. For adults 18+ only.
See the method in action
Read the full list, or start with our number one pick and judge the scoring for yourself.
Swipey AI is free to start. For adults 18+.
Methodology questions
Are the scores objective?
No, and we do not claim they are. They are editorial opinions from our own hands-on use of each product, applied consistently across every app. We label them as editorial wherever they appear.
Do the weights add up to an exact formula?
The weights are illustrative. They describe roughly how much each criterion counts toward an overall figure, not a strict calculation you could reproduce to the decimal.
Do you accept payment to change a score?
No. We do not sell placement or ratings. The one bias we carry is disclosed on every page: we make Swipey AI, and we rank it first.
How often are the scores updated?
We re-test when an app ships a meaningful change to chat, voice, images, content policy or pricing, and we refresh the list on a rolling basis through the year.
Why are there no user reviews or star counts?
Because we will not publish numbers we cannot stand behind. Fabricated ratings and testimonials are exactly what this methodology exists to rule out.