Subscore

QualityTesting methodology

Quality measures how original, complete, and visually polished the ready-made character library feels.

A platform can have hundreds of characters and still feel low quality. Some apps fill their library with copied profiles, weak descriptions, and broken images just to make the total number look more impressive.

We review one fixed sample of characters and use the same sample for every Quality test.

  • 3 evidence groups
  • 4 scored tests

Quality is organized into 3 evidence groups: Duplicate profiles found, Profile Quality, and Visual Quality. Duplicate count is worth half of Quality; the two quality ratings share the rest.

Every scored test gets a score from 0 to 10.

All four tests count equally. We combine their scores to calculate the final Quality score. Quality makes up of the Characters score.

  • 25%Duplicates
  • 25%Originality
  • 25%Profile Quality
  • 25%Visual Quality
Weighted evidence groups (combined 100%)Quality score of Characters score
View exact calculation

Each test gets a score from 0 to 10. We multiply every score by 25% and add the points together.

Quality makes up of the Characters score. Characters makes up 10% of the overall performance score.

Scored tests and weights

Scored testHow much it countsSee scoring
Duplicates25.00%View
Originality25.00%View
Profile Quality25.00%View
Visual Quality25.00%View
Total100%

How the score is calculated

  1. Each test gets a score from 0–10

  2. Test score × how much it counts

    Example: 9.40 × 25.00% = 2.35 points

  3. We do this for every test

  4. We add all the points together

  5. Final Quality score

    Counts for 33% of Characters

Example calculation

We multiply each test score by how much it counts. We then add all the points together.

Scored testTest scoreHow much it countsCalculationPoints added
Duplicates9.4025.00%9.40 × 25.00%2.35
Originality8.2025.00%8.20 × 25.00%2.05
Profile Quality8.4025.00%8.40 × 25.00%2.10
Visual Quality9.0025.00%9.00 × 25.00%2.25
Final Quality score100%Add all points8.75/10

Special cases

Not Applicable

If a test does not apply, we remove it and spread its weight across the remaining tests.

Unknown

If we cannot verify a result, the test receives a score of 0.

Manual adjustment

In rare cases, we may adjust a score when the calculated result is clearly unfair. We always record the reason.

A large character library does not automatically mean a good character library.

Some platforms add new characters as quickly as possible to make their library look bigger. You often end up with copied personalities, nearly identical images, empty profiles, and the same basic scenario repeated with a different name.

That can make the app feel boring even when it claims to offer hundreds of characters.

Strong character quality means the profiles feel different, give you enough information before starting a chat, and use images that look clean and well made.

That is why we check the library itself instead of only counting how many characters are available.

We use a paid account and choose one fixed sample from the ready-made character library.

We use the same characters for all four Quality tests. This prevents us from using the strongest profiles for one test and weaker profiles for another.

First, we check the sample for duplicate or near-duplicate profiles.

We then check whether the characters have different appearances, personalities, and scenarios.

Next, we review how complete each profile is.

Finally, we review the main profile image for clarity, visual problems, and overall presentation.

Duplicates still count toward the library totals under Variety. Quality is where we judge whether those listings are actually different and well made.

Quality is based on a fixed sample, not every character in the library. The sample helps us spot clear patterns, but it cannot guarantee that every profile has the same level of quality.

Character libraries also change regularly. New profiles may be added and weak profiles may be improved or removed after testing.

Profile and photo quality ratings involve some judgement. We use the same five-point scale for every platform.

Public community characters may have very different quality from characters made by the platform. We record which type of library was included in the sample.

Evidence groups

Quality has 3 evidence groups made up of 3 scored tests.

Duplicate profiles

1 scored test

Duplicate profiles found counts how many characters in the sample appear to be copies or near-copies of another profile.

A large number of duplicates can make a big library feel much smaller than it looks.

Duplicate profiles found

How many reviewed profiles repeat another character in the sample (0–25).

How we test

We review every character in the fixed sample (usually 25) and count near-copies. A profile counts as a duplicate when most of these parts are nearly the same: main image, name, description, personality, and scenario.

What counts
  • The same profile image with only a small change
  • Nearly identical names and descriptions
  • The same personality and scenario reused across profiles
  • Profiles that clearly look like copied versions of one another
What does not count
  • Characters that share one trait but are otherwise different
  • Characters from the same theme with different personalities
  • Small similarities that are common across the whole app
  • Profiles that use the same visual style but represent different characters
Result shown

3 duplicate profiles in a sample of 25

0 is best · 25 is worst

Scoring

Fewer duplicates means a higher score.

Duplicate rateScore
0 found10/10
5 found8/10
10 found6/10
15 found4/10
20 found2/10
25 found0/10

Finding 3 duplicates in a sample of 25 scores 8.8/10.

The exact score moves with the duplicate count (inverted 0–25 scale).

Evidence group 2 of 3

25%

Profile Quality

1 scored test

Profile Quality rates overall profile usefulness from very bad to very good for the same sample.

A good profile should explain who the character is and what kind of conversation or relationship to expect.

Profile Quality

Overall quality of character profiles (very bad → very good).

How we test

We review the same fixed sample and rate overall profile quality: Very bad, Bad, Neutral, Good, or Very good.

What counts
  • Clear names and descriptions
  • Personality and scenario details that set expectations
  • Profiles that feel complete enough to start chatting
What does not count
  • Information that only appears after starting the chat
  • Empty or placeholder text
  • Personal preference for one personality style over another
Result shown

Good

Profile Quality: Good (75% → 7.5/10)

Scoring

We rate overall profile quality on a five-point scale.

Checks passedScore
Very bad0/10
Bad2.5/10
Neutral5/10
Good7.5/10
Very good10/10

A “Good” rating scores 7.5/10.

Very good maps to 100% (10/10).

Evidence group 3 of 3

25%

Visual Quality

1 scored test

Visual Quality rates overall photo quality from very bad to very good for the same sample.

A profile image is often the first thing you see. Broken faces, damaged bodies, or blurry images can make the whole library feel rushed.

Visual Quality

Overall quality of the main profile images (very bad → very good).

How we test

We review the main images in the same fixed sample and rate overall photo quality: Very bad, Bad, Neutral, Good, or Very good.

What counts
  • Clear faces and complete bodies
  • Believable anatomy and presentation
  • Images that look finished and suitable for a profile
What does not count
  • Small style choices that are clearly intentional
  • Minor background problems that do not affect the character
  • Personal preference for realistic or anime art
  • Image quality inside the chat or image generator
Result shown

Very good

Visual Quality: Very good (100% → 10/10)

Scoring

We rate overall photo quality on a five-point scale.

Checks passedScore
Very bad0/10
Bad2.5/10
Neutral5/10
Good7.5/10
Very good10/10

A “Very good” rating scores 10/10.

Very good maps to 100% (10/10).

Back to Characters