EXTERNAL POLICY ASSESSMENTS
What the metadata establishes—and what it does not.
Known Robot inspects a pinned upstream revision with the non-executing validator, publishes the generated manifest and missing-information report, and preserves author and license attribution. These records are not evaluations or compatibility results.
Assessment boundaries
- No policy code, checkpoint or robot command is executed.
- Upstream performance statements remain attributed model-card claims.
- An assessment never creates trials, verification status or compatibility graph edges.
- A measured evaluation becomes a separate record only after actual simulation execution or attributable physical evidence.
Published assessments
METADATA ONLY · INCOMPLETE
A non-executing assessment of a LeRobot ACT policy trained on the PushT dataset.
Source: aadarshram/act_pusht · 6d403b142934 · Apache-2.0
12 missing requirements · 2 cautions · no evaluation performed
METADATA ONLY · INCOMPLETE
A non-executing assessment of an upstream policy documented for a physical SO-101 ball-in-cup task.
Source: abdul004/so101_act_policy_v5 · c14c7f4108c0 · Apache-2.0
12 missing requirements · 2 cautions · no evaluation performed
METADATA ONLY · INCOMPLETE
A non-executing assessment showing why a base VLA checkpoint is not itself deployment-ready evidence.
Source: lerobot/smolvla_base · d9f33c94a60f · Apache-2.0
13 missing requirements · 2 cautions · no evaluation performed
Why publish incomplete examples?
A useful validator should reveal uncertainty instead of laundering it into a compatibility claim. These contrasting records show a simulation policy, a physically documented policy whose results remain upstream claims, and a base checkpoint that still needs task-specific fine-tuning and evaluation.
Run the validator on your policy →