Evidence-first physical AI intelligence
Know what humanoid robots can actually do.
Most robot databases repeat specifications. We report what independent evidence supports — how capable a robot has been shown to be, how commercially real it is, and how much you should trust either number. The three are never merged.
Three answers, never one score
A single rating would hide the thing you need to know.
A well-evidenced prototype and a poorly-evidenced product are not the same, and one number cannot say both. So we publish the dimensions separately and show our working.
Capability evidence
What the robot has been shown to do, weighted by how independent and how durable the demonstration was. Manufacturer video earns less than an examined third-party test.
How evidence is tiered →Commercial reality
Whether anyone is actually paying for it and running it. Reservations, pilots, contracts and verified production are kept structurally separate — a 2035 agreement is not a robot on a line today.
How commercial reality is scored →Evidence confidence
How much to trust the two numbers above. Thin sourcing caps confidence no matter how impressive the underlying claims look. It is a statement about our evidence, not about the machine.
What caps confidence →Selected profiles
Four robots, four different evidence positions.
Chosen to show the range of what the evidence supports — not a ranking, and not a recommendation. Every profile carries the full source ledger behind its numbers.
Tiangong Ultra
Beijing Humanoid Robot Innovation Center · ChinaWalker S2
UBTECH Robotics · ChinaAGIBOT G2
AGIBOT · ChinaUnitree R1
Unitree Robotics · ChinaDemo is not deployment
Where a robot actually sits on the commercial ladder.
These four positions are kept structurally separate in the data model. This is an explanation of the ladder, not a score, and no robot is placed on it here.
One claim, fully traced
Every number on this site opens into its source.
“Longcheer Technology (Apr 2026): AGIBOT G2 robots performed pre-shipping tablet testing on Nanchang production line. PoC acceptance confirmed early 2026.”
Latest evidence changes
What changed, and why.
Drawn from recorded score history. Nothing on this page is generated or inferred.
| Date | Robot | Change | Engine |
|---|---|---|---|
| 2026-09-06 | Booster T1 | Batch B — Unitree R1 EVIDENCE READY via EdUHK Academic | Scoring Methodology v1.2.0 (ENGINE v1) |
| 2026-09-06 | Booster T2 | Batch B — Unitree R1 EVIDENCE READY via EdUHK Academic | Scoring Methodology v1.2.0 (ENGINE v1) |
| 2026-09-06 | Walker S2 | Batch B — Unitree R1 EVIDENCE READY via EdUHK Academic | Scoring Methodology v1.2.0 (ENGINE v1) |
| 2026-09-05 | LimX Luna | Batch C/D — Tiangong Ultra EVIDENCE READY | Scoring Methodology v1.2.0 (ENGINE v1) |
| 2026-09-05 | AGIBOT A2 Ultra | Batch C/D — Tiangong Ultra EVIDENCE READY | Scoring Methodology v1.2.0 (ENGINE v1) |
| 2026-09-05 | Persona Humanoid | Batch C/D — Tiangong Ultra EVIDENCE READY | Scoring Methodology v1.2.0 (ENGINE v1) |
Read the methodology
Evidence tiers, operating mode, source independence, durability, and why a robot with two publishable claims can still report Very Low confidence.
How scoring worksFound an error? Tell us.
Anyone may submit a correction or new evidence, manufacturers included. Submissions never change a score directly, and scores cannot be bought.
Corrections policy