Imagine you're blindfolded in a dark room trying to find your keys. You could wave your arms randomly through empty air — or you could run your hands along surfaces, guided by the principle that if you feel something, you're getting warmer. Most robot reinforcement learning is the arm-waving version: random actions in free space, burning compute on motions that teach the agent nothing about manipulation. TacEx is the surface-tracing version, directing the robot's curiosity specifically toward moments of touch. The core claim is clean: by decomposing a world-model's epistemic uncertainty into separate sensory channels and directing exploration toward the tactile channel, you get robots that seek out contact instead of avoiding it. The method builds on ensemble disagreement — a standard curiosity technique — but filters it through touch. Where vanilla model-based curiosity rewards any state the model is confused about (including bizarre free-space gyrations), TacEx rewards uncertainty specifically in predicted tactile readings. The robot learns that touching objects is informative, so it touches more objects, so it collects interaction-dense data, so downstream policies actually work. The results chain through two stages. First, purely reward-free exploration: TacEx generates datasets rich in contact events — grasps, pushes, manipulations — without any task reward or demonstration. Second, those datasets feed offline RL for downstream pick-and-place tasks, and separately post-train vision-language-action (VLA) models. The VLA result is particularly interesting: models pre-trained without any tactile signal improve substantially when post-trained on TacEx-collected data, suggesting tactile curiosity generates training signal that transfers across architectures. Architecturally, TacEx sits in the model-based RL family with ensemble world models and intrinsic motivation. The sensory decomposition is the structural novelty — splitting the disagreement signal across modalities is a design choice, not a new algorithm. The hardware dependency is real: you need touch sensors (the paper uses a GelSight-style tactile sensor on the fingertips). This is both the method's strength and its limitation — tactile sensors aren't standard equipment on most robot platforms, and the specific sensor properties shape what 'tactile uncertainty' means. The validation is simulation-based with some real-world experiments, comparing against standard intrinsic motivation baselines (RND, Plan2Explore, disagreement-based curiosity). The baselines are reasonable but not exhaustive — notably, the comparison to VLA post-training doesn't pit TacEx against other data-collection strategies for post-training (e.g., scripted exploration, human demonstrations of similar density). The offline RL evaluation shows clear improvements, but the absolute success rates and the sensitivity to sensor noise deserve scrutiny. The milestone question is sharp: can tactile curiosity scale to multi-finger dexterous manipulation with richer contact dynamics? Current results are on relatively simple pick-and-place with parallel-jaw grippers. The gap between 'grasp an object on a tabletop' and 'in-hand reorientation of a tool' is enormous in contact complexity. If TacEx-style decomposition holds when tactile signals become high-dimensional (e.g., full-hand tactile gloves with hundreds of taxels), it would move from a neat trick to a foundational exploration primitive. The obvious experiment not run: scaling to dexterous hands with dense tactile arrays. The authors almost certainly know this is the next step — the question is whether the ensemble disagreement signal remains tractable when tactile dimensionality explodes from a single GelSight patch to 500+ taxels. My read: they're saving it for the next paper, and the compute/hardware requirements for dexterous sim-to-real are genuinely steep.