Research

  1. 2026.08 Map the Failure Boundary We froze a Go1 locomotion policy and measured where it fails across 6,400 combinations of floor friction and lateral push, then retrained on the two gaps the map exposed and measured again.
  2. 2026.05 Monte Found a Decorative Channel and a Reward Exploit in One of Our Benchmarks Monte's adversarial training exposed two failure modes in a MARL benchmark: a communication channel that looked load-bearing but wasn't, and a reward function that incentivized trivial policies.
  3. 2026.05 Where V-JEPA 2.1's Dense Features Hold Up (and Where They Don't) A pre-registered V-JEPA 2.1 robustness study across all four model sizes, with deployment lessons.
  4. 2026.01 10,924x: The Instability Bomb at 1.7B Scale Part 2 of the mHC reproduction series. I scaled from 10M to 1.7B parameters and watched Hyper-Connections hit 10,924x signal amplification. That's not a typo.
  5. 2026.01 DeepSeek's mHC: When Residual Connections Explode DeepSeek replaced standard residuals with massive Hyper-Connections, and then watched them explode. I reproduced the 9x signal amplification to understand their fix: the Sinkhorn-Knopp algorithm. Part 1 of 2.