How Marin Trains 535B: Scaling Ladder, MoE Expert Parallel, Harrier Data and Live W&B
Stanford Marin pre-registers a paloma macro-loss of 2.04 with a 5-rung Scaling Ladder at 1% cost, then trains 535B-A23B on 11×GB200 in public with live W&B telemetry — 847 training buckets already show the most teachable frontier run.