Imagine you're designing a TV remote. You could give the user 47 buttons — one for every internal function — or you could give them 6 buttons that combine naturally: volume, channel, power, input, menu, select. The engineering challenge is figuring out which 6 buttons cover the space of things people actually want to do. That is the core mechanism of this paper: using competitive games as a pressure cooker that forces an agent to carve motor behaviors into a small, distinct, combinable button set that a human can actually use. The committed claim: competitive self-play against past selves produces a discrete set of motor skills that are simultaneously semantically distinct, human-interpretable, and expressive enough for a person to compose them into solutions for tasks never seen during training. This is not the first unsupervised skill-discovery paper, but the grounding mechanism — games as the selection pressure — is genuinely new. Previous approaches like DIAYN or DADS optimize information-theoretic objectives that produce skills with poor interpretability or redundancy. GGSD instead lets competition do the pruning. The architecture is a two-level hierarchy trained via multi-agent reinforcement learning. A high-level policy picks from a small discrete skill set (the 'buttons'), while a low-level policy executes continuous motor commands conditioned on the chosen skill. The entire system trains through self-play against historical copies of itself, which creates an automatic curriculum: as the agent improves, its opponents improve, forcing skill differentiation and tactical diversity. After training, you rip out the high-level policy and hand the discrete skill selector to a human. The key structural bet is that competitive pressure naturally produces the properties you want — distinctness, interpretability, composability — without needing explicit reward engineering for each. The paper demonstrates this across three embodiments: an Ant quadruped, a Franka robotic arm, and a Unitree G1 humanoid. After self-play training in combat-style games, humans took over the high-level controls and solved unseen tasks — navigating a Maze with the Ant, pushing a cube to a target with the Franka arm — without any additional training of the low-level skills. The emergent combo behaviors are particularly interesting: transitions between skills produce movement patterns that weren't explicitly trained but arise from the dynamics of skill sequencing, expanding expressivity beyond the individual primitives. Integrity-wise, this is simulation-only work across three environments, which is standard for the subfield but means the gap to real hardware remains unaddressed. The baselines include DIAYN and other unsupervised skill-discovery methods, which are the right comparisons. Human evaluations of playability are included, which is crucial for the specific claim being made — you can't evaluate 'human-playable' without humans. However, the human studies appear small-scale and informal rather than rigorously controlled. The milestone question is about scale and transfer. Three environments with small discrete skill sets (the paper doesn't specify exact numbers but implies single-digit counts) is a proof of concept. The next concrete threshold is demonstrating this on a real robot with human-in-the-loop control, and scaling the skill vocabulary to handle environments with richer action requirements — say 15-20 distinct skills composed across manipulation tasks that matter industrially. The gap is probably 2-3 years given sim-to-real transfer challenges. The obvious experiment not run: sim-to-real transfer on actual hardware. The Unitree G1 is a real robot platform, the Franka arm is standard lab hardware — the ingredients for a hardware demo are right there. The honest read is likely (a) compute and logistics: real-robot experiments are expensive and slow, and the self-play training pipeline may need significant adaptation for real-world latency and noise. They're probably saving hardware validation for a follow-up paper or a robotics venue submission.