The Cyborg Index · August 2026 snapshot

How to read a humanoid ranking

The Cyborg Index reads the public evidence record of humanoid robots, not the marketing. This lesson walks through one real robot so you can see what the numbers mean — and what they do not mean.

01

One sentence to remember

A high rank means a robot has the strongest publicly evidenced capability record right now. It does not mean it is the "best" robot, the most advanced in secret, or the one that will win in five years.

02

Meet Tesla Optimus

Tesla Optimus is ranked #10 in the August 2026 edition. It is famous, but fame is not evidence. Here is how its scores break down.

Capability score
0.72 / 5

Average across 40 capability checks. This is what ranks the robot.

Proof score
0.16 / 5

How much of that score is backed by public evidence.

Coverage
20 / 40

How many checks have any public evidence at all.

Proof gap
0.56 points

Tesla Optimus's Capability score is 0.72, but its Proof score is only 0.16. That gap means most of its recorded capability is supported by weaker sources — claims and demos — rather than independent or operational evidence.

03

The whole field is still early

The current leader, UBTech Walker S, has a Capability score of 1.48 out of 5. That is the top of the table — and it is still below 1.5. The ranking is not saying humanoids are almost ready. It is saying the public record is thin across the whole field.

04

Evidence is capped, not ignored

Not all sources are equal. A vendor claim or a staged demo supports less than an independent report, a customer trial, or a measured result. Weaker sources are capped, not ignored.

Unknown

No public evidence yet. Counts as 0 in the average.

Vendor claimed

The maker says the robot can do it. We have not seen it.

Demo shown

Shown in a video or staged setting. Useful signal, but capped.

Validated

Independently observed, measured, or replayable.

05

Commercial proof moves in stages

An axis cannot score above the stage its commercial evidence supports. Here is the ladder, from weakest to strongest.

01

Claim

The vendor says it can do something.

02

Demo

It can be shown under chosen conditions.

03

Pilot

It is being tested in a real setting.

04

Named deployment

A named customer is using it for real work.

05

Measured operation

Its performance is tracked and reported.

06

Scale

Many sites, many shifts, many tasks.

06

What to remember

  • 01Rank reflects the public evidence record, not fame, funding, or future promise.
  • 02Scores are out of 5. A top score around 1.5 means the whole field is still early.
  • 03Missing evidence counts as 0, not a verdict that the robot cannot do the thing.
  • 04A robot can rank high while most of its score rests on thin proof. That is visible as a proof gap.
07

Key terms

Capability score
The Index score. An average of 40 capability checks, each scored 0–5. It is the ranking authority.
Proof score
How much of the Capability score is backed by public evidence. It is a reference read, never the rank.
Coverage
How many of the 40 checks have any public evidence at all.
Proof gap
The distance between Capability score and Proof score. A large gap means the public record is thinner than the score implies.
Evidence ceiling
A weaker source can only lift a score so far. A vendor demo cannot push a check as high as an independent measured result.

Try it on the live ranking

Open the current Cyborg Index, pick any robot, and look for the same four numbers: Capability score, Proof score, Coverage, and Proof gap. The story is always in the gap.