How BotInfo scores robots

Billions of dollars into the robotics industry, there is still no agreed-upon, buyer-facing way to measure what a robot you can actually purchase will actually do. Lab benchmarks are built for researchers. Manufacturer spec sheets are built for marketing. Procurement happens in the gap.

BotInfo's evaluation standard scores buyable robots on the things buyers ask us about: how long until it runs, how long it runs for, what it can carry, what it legally is in the United States, what it costs to get to operational — with every score carrying an evidence grade and a date.

The evidence-grade system

No score on this site is ever stronger than its evidence. Every published figure carries one of four grades:

GradeMeaning
E0Spec-sheet claim — manufacturer documentation, cross-checked against regulatory filings where possible. A claim, clearly labeled as one.
E1Video-verified — scored from dated public footage, source cited.
E2Hands-on — a BotInfo-operated unit run through our published test protocol.
E3Field report — verified feedback from an institutional deployment.

Grades are never blurred. An E1 score is “video-verified,” never “tested by us.” Where the evidence doesn't support a score, the scorecard says “insufficient evidence” — an honest gap beats an invented number, and the gaps themselves tell you what no one in this industry will currently show you.

The task battery (v0.1)

Each robot is assessed against nine buyer-relevant categories:

  1. Unbox-to-operational time — from crate to first commanded motion.
  2. Real battery runtime under mixed load, versus claimed.
  3. Locomotion — stairs, ramps, carpet, door thresholds.
  4. Manipulation — payload at reach, pick-and-place repeatability.
  5. Safety behavior — e-stop, fall behavior, bystander proximity.
  6. Regulatory & connectivity — US equipment-authorization status (verified against FCC records), offline capability.
  7. SDK & software maturity — documentation quality, API surface, agent-platform compatibility.
  8. Support & parts — parts availability, warranty terms as published.
  9. Total cost to operational — checkout-verified price, plus the accessories real deployments actually need.

Categories 6 and 9 draw on datasets BotInfo already maintains daily: verified checkout pricing with visible dates, and per-model US authorization status — data no vendor publishes honestly about itself.

Regulatory scoring rubric (category 6): 5 = per-model FCC grant verified in the public record · 4 = authorized-generation platform (grantee-level records, model mapping not 1:1 public) corroborated by flowing US distribution · 2 = authorization status unresolved · 0 = not authorized for US import/marketing/sale.

Published scorecards

Each scorecard below is scored against the battery above, with every line graded and dated.

Unitree Go2 — quadruped scorecard

The most accessible proven quadruped platform legally buyable in the US today.

Unitree G1 — humanoid scorecard

The entry point to US-legal humanoids — authorized-generation and checkout-buyable.

Unitree R1 — humanoid scorecard

The cheapest per-model-FCC-verified humanoid a US buyer can put in a lab today.

Unitree A2 — quadruped scorecard

The industrial-class quadruped with a verified pre-deadline FCC grant.

Unitree H2 — humanoid scorecard

The flagship humanoid that beat the FCC deadline — the clearest picture of where the US-legal ceiling sits.

Neutrality

Published scores are never for sale and never vendor-influenced. Vendors may in future pay for evaluation services — expedited hands-on testing — but never for outcomes. If that line ever moves, this page will say so in plain text, dated.

Methodology v0.1 — published August 2026. Sub-score definitions, change history, and the full per-model dataset are part of BotInfo's intelligence tier.

Humanoid price trackerRobot dog price tracker