Joshua Saxe
First principles / bitter lesson would suggest the west has less of a moat with Nvidia training hardware than many suspect. There are many ways to update weights outside of the path dependent infra in the big labs. And there's a promising lit showing creative ways of doing big model training on less capable hardware. Coding agents should also weaken the cuda moat