VGGT-Omega 1B-256 OxE+MimicGen+RoboCasa365 β€” step 23000 snapshot

Single checkpoint snapshot from the 3DA_unified VGGT-Omega 1B-256 multi-source pretraining run, taken at training step 23,000.

Provenance

  • Run: [VGGTOMEGA]_oxe_mimicgen_robocasa365_H8_vggtomega256_text_8n_mb2acc2_compile_2315944
  • W&B run id: bags/robot-gld/ohf3lpoa
  • Backbone: VGGT-Omega 1B (encoder_input_size=256, DINOv3 patch=16)
  • Data: OXE 23-dataset mix + MimicGen + RoboCasa365 (Cosmos24 subset)
  • Training: 8 nodes Γ— 4 GH200 = 32 GPUs, CSCS Clariden

Important note: mixed-H training history

  • Steps 0–22,000: trained with history_horizon=4 (canonical run config)
  • Steps 22,001–23,000: trained with history_horizon=8 via a wrong-defaults resume attempt (job 2331403)

If you intend to use this for further H=4 training, the weights at step 22000 are cleaner. See the matching W&B run summary for per-step metrics.

File

  • 0023000.pt β€” full DeepSpeed-style checkpoint (~12.8 GB, contains model_state, optimizer_state, scheduler_state, step, action/proprio normalizer stats)
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support