You know how a cooking student learns from watching a chef botch a flambé? The fire goes wrong, but the student still picked up the knife work, the timing of the pan tilt, and the intent to caramelize — everything before the failure was useful instruction. TRACC works the same way: it watches a failed human attempt, copies the motion trajectory up to the point of failure, then uses the inferred task goal to train a policy that finishes the job. The committed claim: a humanoid robot can learn a complete skill from a single failed human video, with no successful demonstration at all. The system splits the problem into two phases — imitate the usable prefix of the motion, then switch to a task-completion reward that drives the policy toward the intended outcome. The prefix acts as prior knowledge, bootstrapping the robot into the right region of the motion space before the reward signal takes over. Architecturally, TRACC sits in the imitation-learning-from-observation family, but with a critical twist. Standard methods in this space (motion retargeting plus RL fine-tuning, à la PHC, UH-1, or OmniH2O) assume access to a clean, successful reference trajectory. TRACC relaxes that assumption by segmenting the video into a usable prefix and a failure suffix, then stitching an RL completion phase onto the prefix. The pose extraction uses off-the-shelf vision models, the retargeting maps to a simulated humanoid, and the policy trains in IsaacGym — nothing exotic in the stack, which is actually a strength. The evaluation uses six in-the-wild failed human tasks from the Oops! dataset — a publicly available collection of real failure videos. Tasks include things like failed basketball shots, failed jumps, and failed object manipulations. The authors compare against an ablation where the full (including failed) trajectory is imitated without the task-completion switch, demonstrating that naively imitating through the failure point produces a policy that reproduces the failure. This is the right ablation, but the baseline set is thin: there is no comparison against a from-scratch RL agent given only the task reward (no video at all), which would establish how much the video prefix actually helps versus just having a well-shaped reward. Integrity is a mixed bag. The Oops! dataset is public and not curated by the authors, which is good. But all evaluation is in simulation (IsaacGym), the six tasks are chosen post-hoc from a larger dataset without stated selection criteria, and there is no real-robot transfer. The results show the method works on these six tasks, but without quantitative success-rate comparisons against a no-video baseline or a successful-video upper bound, the reader cannot gauge how much of the final performance comes from the video prefix versus the task reward design. The milestone question is sharp: the field needs to go from simulation-only (this paper) to real-robot execution with in-the-wild video input. That means closing the sim-to-real gap for humanoid whole-body skills trained from arbitrary internet video. Current sim-to-real humanoid work (e.g., from 1X, Figure, Unitree) is still heavily domain-restricted. A realistic next number: demonstrate TRACC-style learning on a physical humanoid for at least 3 distinct tasks sourced from unscripted internet video, with >50% first-attempt success. The obvious experiment not run: testing with a real robot, or at minimum, comparing against an RL-only baseline with no video input. The honest read is (a) — compute and hardware access. Sim-to-real for diverse humanoid skills is expensive and requires hardware the authors likely don't have. There's also a conspicuous absence of comparison to methods that learn from successful demonstrations on the same tasks, which would establish an upper bound. That comparison would clarify whether TRACC is closing 30% or 90% of the gap between no-demo and full-demo performance.