Skip to content
BecomeTen
How BecomeTen works

Methodology & evidence

What the score actually means, why potential is capped, how we measure change over time, and where the cited studies come from. These are the rules the rating follows.

Printed research papers under a desk lamp, one page marked with a ruler: the evidence library behind every plan step
01

Why is the score an estimate, not a measurement?

The 0–100 score is an AI estimate from a single 2D photograph. It is not a measurement and not objective truth; it is a structured opinion against a fixed rubric of five areas (harmony, eye area, jaw and chin, skin, hair) and 30-plus traits inside them. That is why we never state numeric values for structural traits such as a gonial angle in degrees or a philtrum in millimetres: a phone photo cannot support that number, and pretending otherwise is false precision. Every trait comes back as a directional band instead — more or less, positive or negative, sharper or softer — with a short note on what drives the read. The same face photographed at a different angle, distance or light can land a few points apart, and the report says so rather than hiding it. What the number is good for is a starting point and a trend, not a verdict.

02

Why is the calibration so strict?

An average, healthy, groomed adult should land around 50–60 on our scale, not 75 and above. A rating that tells everyone they are above average is worthless, and the audience for these apps is the first to notice; most raters drift upward because a flattering result gets shared and an honest one does not. Scores above 80 are reserved for genuinely striking structure, and the two top tiers, Elite and Apex, are rare by design. The eight tiers are fixed bands on this scale, not a ranking against other users: there is no leaderboard and no comparison to anyone else who has scanned. The PSL equivalent shown next to the score is a fixed mapping onto the community's 1–8 ladder (50 sits at 4, 62 at 5, 75 at 6), offered as a translation for people who already think in that scale, never as the primary number.

03

Why does potential have a hard ceiling?

Potential is what is reachable without surgery, and the server caps it per area according to what genuinely moves: harmony +6 points, eye area +10, jaw and chin +14, skin +25, hair +20. The model cannot exceed that ceiling even if it wants to; the cap is enforced in code after the model answers, not requested in the prompt. Bone geometry does not change, which is why harmony gets the least and skin gets the most: softmaxxing works on skin, hair, composition, grooming and sleep, and those are the areas where published evidence exists. The gap between your score and your potential is therefore a bounded, honest estimate of a few months of consistent work, not a promise. If every area hit its cap at once the total would move by roughly a tier's width, and almost nobody hits every cap.

04

What is structural and what is workable?

Every weak point is tagged workable or structural. Structural means bone or fixed anatomy: the gonial angle, ramus length, maxilla position, orbital shape, facial thirds and fifths, canthal tilt. For those the plan will never promise a fix; it may only work with how the trait reads through body composition, posture, hair framing, brows and grooming. Workable means the tissue itself changes without a procedure: skin quality, under-eye appearance, hair density and cut, facial fat, and the expression and framing the camera sees. Selling a routine that claims to reshape bone is exactly what turns a rating app into a scam, so anything in that category, from tongue posture to devices, is labelled as having no studies or refused outright.

05

How does photo quality change what we claim?

Every photo is graded for lighting, angle and sharpness before it is scored. If a photo is not comparable to the last one, we withhold the score comparison rather than report a change that the camera caused. Differences within ±2 points are shown neutrally as measurement noise. The most common false positive is angle: a chin tilted down artificially sharpens the jaw and shifts the entire lower-third score, and a phone held at arm's length distorts the midface and the width-to-height ratio in ways that have nothing to do with the person. Overhead light deepens under-eye shadows; a wide smile changes the philtrum and the eye area. The report tells you what the capture grade was, so a low score on a bad photo reads as a bad photo, not as a bad face.

06

How is progress compared between scans?

On a re-scan the model sees both photos side by side and may only report visible differences. “Unchanged” and “regressed” are first-class answers; the prompt explicitly forbids inventing progress, because a rater that always finds improvement is selling reassurance, not information. The model may not report a change in bone geometry between scans at all, because there is none: if two photos make the jaw or the eye area look structurally different, the cause is angle, distance or light, and the capture grade usually catches it. The useful cadence is a checkpoint every six to eight weeks under matched conditions, judged as a trend rather than as single readings; the tracking guide describes the setup. PRO keeps the series so the trend is visible across scans.

07

What do the four evidence levels mean?

Every plan step carries an evidence level: strong (multiple randomized trials or meta-analyses), moderate (a few solid studies), weak (a single small study), no studies (practice without formal research). The level is attached to the step itself, next to the citation, so you can see at a glance whether daily sunscreen and a retinoid (strong) sit in the same plan as brow shaping (no studies). Mewing, jaw exercisers and most grooming advice fall into the last category and are labelled that way; the label is not a judgement that they are useless, only a statement that nobody has measured them properly. Where the evidence is for a different population than you (a trial in men, a different age range), the step says so. The four levels are the reason the plan can be honest about what it does not know.

08

Where do the citations come from?

The AI model is not allowed to “remember” a study; it may cite exclusively from the curated library below. If no library entry supports a recommendation, the step is labelled as having no studies and carries no source. The server checks every citation the model returns against the library by id, so a fabricated reference cannot reach you; a hallucinated paper is the single most common failure of language models asked for sources, and this closes the door on it entirely. Each entry was resolved through PubMed's own lookup and accepted only where PubMed's title matched ours; the few that could not be matched unambiguously link to a title search so you can verify them yourself. The library is small on purpose, covering only the interventions the plan actually recommends, and it grows only when a paper earns its place.

09

What is the potential image, and what is it not?

The simulation shows a possible direction with consistent work over 6–12 months. It preserves identity, bone geometry and skin tone; it may not narrow a nose, change canthal tilt or redraw a jaw. It changes only what genuinely changes: skin, hair, grooming and moderately facial fat, the same areas the potential caps allow. It is not a prediction or a promise, and it is generated under the same rules as the score, so an image that contradicted the caps would be a bug, not a feature. Treat it as a sketch of the direction the plan points in, and treat any before-and-after you see elsewhere, including ours on the landing page, as an illustration with the same limits: same person, same bone, only workable things moved.

10

What will BecomeTen never do?

We do not compare you to other users and we run no leaderboard. We do not adjust the score for ethnicity or gender identity; the capture is graded, the person is not adjusted. We do not diagnose, and nothing here replaces a doctor: where a step needs a prescription or a clinician, the plan says so and stops. We never recommend surgery, and never anything that damages tissue; bone smashing and similar forum methods are explicitly forbidden in the prompt and refused if asked. We do not use the demeaning terminology common in this community; the tiers have neutral names on purpose. We do not keep a free scan's photo, and no photo is ever used to train a model. The rest of the looksmaxxing guide and the glossary follow the same rules.

Evidence library

The complete list of publications the plan is allowed to cite. Every entry is a real, hand-picked paper — links search the title on PubMed so you can verify it yourself.

Exfoliation (AHA/BHA)

Moisturizer & barrier

Hair

Educational content, not medical advice. Consult a licensed professional before any treatment.

Back to home