Imagine you've trained to run on a treadmill your whole life, and someone drops you on a beach at midnight. Your legs sink differently with every step, sand kicks up into your eyes, and the only light source throws shadows so harsh you can't tell a footprint from a crater. That's roughly what happens when a legged robot designed for structured terrain meets loose lunar regolith under extreme lighting — and this paper is the first systematic field report documenting exactly how badly things go wrong. The core claim: researchers from ETH Zurich's Robotic Systems Lab deployed two legged robots — the quadruped ANYmal-D and the wall-climbing Magnecko — at the European Space Agency's LUNA analogue facility, which replicates lunar surface conditions including loose regolith simulant and controllable lighting. The robots walked, collected visual-inertial navigation data, and encountered a catalogue of failure modes that no simulation had fully predicted. This is not a performance breakthrough paper. It is a lessons-learned paper, and an honest one. The terrain problems are physical and immediate. Foot-regolith interaction produces sinkage that changes the effective leg length mid-stride and slip that defeats standard locomotion controllers tuned for rigid ground. Every footstep launches dust particles that settle on sensors and optics. The paper documents these effects but does not yet offer solutions — it characterizes the problem space. The regolith simulant (EAC-1A) approximates but does not perfectly match real lunar soil, a caveat the authors acknowledge. Perception failures proved equally punishing. The LUNA facility's controllable illumination exposed a brutal truth: current visual-inertial odometry and depth-sensing pipelines break under the lighting conditions that actually exist on the Moon — overexposed sunlit surfaces adjacent to pitch-black shadows, with low-texture regolith providing almost no visual features for SLAM algorithms to lock onto. The paper reports these degradation modes qualitatively and through representative data, though it does not yet quantify failure rates against a standard perception benchmark. Architecturally, ANYmal-D uses a model-predictive-control locomotion stack with visual-inertial state estimation — the standard pipeline for modern quadrupeds. Magnecko adds magnetic adhesion for inclined surfaces. Neither platform was modified with regolith-specific locomotion policies or dust-hardened perception. The paper's contribution is not algorithmic novelty; it is the identification of what the algorithms need to survive. The integrity profile is mixed in a characteristic way for field robotics. There is no pre-registered benchmark, no community-standard evaluation suite for lunar legged locomotion (because one does not yet exist), and the validation is observational rather than comparative. The paper is essentially a structured field report from a single campaign at a single analogue facility. Its honesty about failure modes is a strength, but the absence of quantitative baselines — how much does sinkage degrade stride efficiency? by what factor does dust reduce perception range? — limits what readers can extract. The real value of this paper is its implicit roadmap. It names four integration gaps that must close before lunar legged robots become viable: regolith-aware locomotion policies, illumination-robust perception, repeatable analogue test protocols, and mission-level operational validation. Each of these is a concrete engineering problem, not a vague aspiration. The next milestone the field needs is a quantitative benchmark suite for lunar analogue locomotion — without one, every group will test under different conditions and progress will be unmeasurable.