Imagine walking through a funhouse where the floor tilts differently under every step, but the tilt is invisible and changes depending on which direction you're facing. That's what a legged robot experiences inside a particle accelerator or fusion reactor: the magnetic field exerts pose-dependent forces and torques on every ferromagnetic component, and those wrenches shift as the robot moves through the field gradient. Conventional controllers treat these as generic disturbances and react after the fact. This paper says: predict the field, model the wrench, and compensate before you stumble. The committed claim is straightforward but meaningful: by building a magnetic wrench model into model predictive control, a quadruped can reject 2.5× more magnetic force disturbance and 1.6–2× more torque disturbance than a non-compensated controller, expanding the robot's safe operational footprint by roughly 29% of the facility area — including a 10.84% zone that would have caused immediate collapse under standard control. The architecture has three interlocking pieces. First, a custom MuJoCo physics plugin models magnetic forces on rigid-body elements, turning the sim into a magnetic-field-aware training ground. Second, an inverse field-estimation framework infers the latent magnetic field from the robot's dynamic response and whatever sensor readings are available — essentially solving the inverse problem of 'what field would produce these forces on this body in this pose.' Third, a Magnet-Aware MPC plus Whole-Body Control stack uses that estimated field to predict upcoming wrenches along the planned trajectory and pre-compensate in real time. The key insight is that magnetic wrenches are pose-dependent and spatially varying, so you cannot treat them as a static bias — the controller must re-estimate and re-compensate at every step. On the ladder, the baseline is the same MPC/WBC stack without magnetic compensation — essentially the robot's own controller with the magnetic model turned off. The 2.5× force rejection factor and 1.6–2× torque rejection are measured against this internal baseline. The paper does not compare against other disturbance-rejection approaches (robust MPC, learned residual policies, domain randomization) or against any external group's work on robots in magnetic environments. This is reasonable — there is essentially no prior art on legged locomotion in strong spatially varying magnetic fields — but it means the ladder is short: the paper competes against itself. Integrity is solid for an early-stage robotics result. Validation spans both MuJoCo simulation and physical hardware experiments, which is the right combination. The inverse field-estimation framework is tested by checking whether inferred fields produce correct wrench predictions on the real robot. The main caveat is that the 'magnetic field' used in hardware experiments is not a full-scale accelerator environment — the paper demonstrates the principle at lab scale. No pre-registration, no community benchmark (none exists), and code availability is not stated. The milestone question is concrete: today's result works in a controlled magnetic field strong enough to destabilize a quadruped. The next real unlock is deploying in an actual Big Science facility — CERN, ITER, or a synchrotron — where fields are stronger, more complex, and the robot must navigate autonomously for extended periods. That requires scaling the field-estimation framework to handle richer spatial gradients and validating against real facility field maps. The obvious experiment not run: testing in a real accelerator tunnel or tokamak hall with the actual field topology. The honest read is (a) — access to these facilities for robotics experiments is extremely limited and expensive, and the field maps are often proprietary. This is a resource constraint, not a methodological dodge. The second missing experiment is comparison against learned disturbance-rejection policies (e.g., RL with domain randomization over magnetic forces), which would tell us whether the physics-model approach genuinely outperforms a brute-force learned compensator.