Imagine you're editing a complex spreadsheet that's connected to a dozen live APIs — your bank, your calendar, your email. You make a series of changes, some of which trigger real actions in those remote services. Now imagine hitting Ctrl+Z and having every change — local AND remote — cleanly revert. That's what Planarian does for LLM agents. The committed claim: Planarian is the first agent runtime that provides consistent, efficient state management across both local sandboxed environments and remote services, using a unified abstraction called 'statepoints.' This isn't just local checkpointing — the hard part is undoing things that already happened on external services that don't support undo natively. The system introduces three primitives. Snapshot captures a point-in-time version of everything — local process state, file system, AND a log of compensating actions that can reverse remote API calls. Rollback restores to a prior statepoint by reverting local state and replaying those compensating actions. Fork creates isolated parallel branches from any statepoint, letting the agent explore multiple strategies simultaneously without interference. The local snapshotting is incremental, avoiding the cost of full copies on every checkpoint. The key engineering trick for remote state is compensating actions — essentially, Planarian transparently records the inverse of each remote operation (if you sent an email, the compensating action deletes it; if you created a file in a cloud service, it deletes it). This is a well-known pattern from database saga transactions, but applying it to the heterogeneous, uncoordinated world of arbitrary web APIs an agent might call is genuinely novel systems work. The limitation is obvious: not every remote action has a clean inverse (you can't unsend a physical letter), but for the API-driven world agents inhabit, it covers a lot of ground. The headline result is stark: agents using Planarian's explore-and-rollback capabilities improve task quality by up to 15x, while the runtime overhead for state management is only 3%. That 15x number deserves scrutiny — it likely measures worst-case recovery scenarios where an agent without rollback is stuck with a corrupted environment. But even discounting the extreme, the ability to speculatively execute and revert is a genuine capability unlock for agents that currently must be conservative or accept irreversible mistakes. The paper sits at the intersection of operating systems (process checkpointing, copy-on-write file systems), distributed systems (saga pattern, compensating transactions), and AI agent infrastructure. The closest prior work is sandboxing approaches like Docker checkpoints or VM snapshots for local state, and conversation-level 'undo' in agent frameworks — but none unifies local and remote state under a single transactional abstraction with fork semantics. This is OS research in service of AI, not AI research dressed as systems. The integrity question is whether the 15x improvement and 3% overhead hold across diverse real-world agent workloads or only on curated benchmarks. The abstract doesn't name specific benchmarks or baselines, which is a flag — we'd want to see this tested against agents using ad-hoc Git-based checkpointing or Docker snapshots. The compensating-action approach also has a fundamental limitation: it assumes remote APIs are deterministic and reversible, which breaks down for services with side effects (notifications sent, rate limits consumed, funds transferred). The authors likely acknowledge this in the full paper, but it bounds the applicability.