RESEARCH INTERFACE / STUDIO 01

Every demonstration.
One measurable scene.

Inspect the complete SO-101 tic-tac-toe corpus frame by frame, align both recorded cameras with MuJoCo, and send the same observations through the published 120K SmolVLA policy.

DATASETtic-tac-toe-so101-block-a-clean-v116d46485
POLICYsmolvla-tic-tac-toe-games-1-15-120kfac0cb54
ENGINEMuJoCo 3.11.0 / browser WASM250 Hz
195recorded episodes
144,723synchronized frames
18language tasks
2camera streams
80.4 minrecorded motion

INTERACTIVE / REVISION-PINNED

SO-101 Replay Studio

Dataset Replay reconstructs recorded joint state. Policy Lab asks the checkpoint for a new 50-step action chunk and executes that output in MuJoCo. The two claims never share a label.

EPISODE 000 · LOADING
MUJOCO 3.11.030 FPS DATA
Loading MuJoCo geometry Preparing the canonical SO-101 scene…
Ready00:00.00 / 00:00.00
CAMERA 01Top / workspace--:--
RECORDED
CAMERA 02Wrist / gripper--:--
RECORDED
BOARD STATEMove —
Loading task…

Board transition pending.

EPISODE METRICSrecorded
6-AXIS TRACEstate solid / action dashed

CORPUS VIEW / 195 EPISODES

Dataset diagnostics

Descriptive measurements of the recorded corpus. They are not a policy benchmark or a physical success rate.

EPISODE DURATIONseconds
TARGET COVERAGE195 moves
INTEGRITYrevision-pinned

MEASUREMENT CONTRACT / READ BEFORE INTERPRETATION

01

Replay is reconstruction.

Recorded state drives the arm. Contact-only kinematic calibration aligns each physical trajectory with the canonical digital scene.

02

Inference is new output.

Policy Lab loads the exact published checkpoint and reports its device, latency, seed, and complete 50-step action chunk.

03

Simulation is not hardware proof.

A simulated rollout can reveal action shape, instability, and collisions. It cannot replace bounded physical evaluation or safety gates.