π0.5 pre-training — landscape, throughput & π-equivalent sizing
estimateSizing π0-equivalent pre-training (match timesteps-seen) on our 8×H200 nodes — a reference-driven VLA landscape table, throughput/memory vs batch, and 1-node vs multi-node time & cost.
π0.5 wrist-only — full-FT baselines
reportWrist-only π0.5 difficulty is mostly a masking artifact: physically removing the third-person camera from attention nearly recovers the both-camera ceiling.
π0.5 — architecture & tensor IO
model cardDual-expert flow-matching VLA, pinned down: per-module tensor IO for pi05_libero, the 3-camera input contract, and every ablation knob mapped to the exact module it touches.