Imagine you're running a restaurant kitchen with a ticket rail. One cook starts silently dumping orders in the bin and calling them 'customer no-shows.' The immediate loss is those meals. But the real damage is what happens tomorrow: the rail backs up with reattempts, other cooks lose rhythm because they're covering ghost tickets, and the whole line slows down. That's the core mechanism this paper identifies in last-mile delivery — fake remarks don't just waste one trip, they poison the next day's workflow. The committed claim: delivery agents who intentionally mark parcels as undeliverable without attempting delivery (fake remarking) cause a measurable spillover productivity loss that reduces the next day's successful deliveries by 1.60% and first-time-right deliveries by 1.86%. This isn't about the obvious direct cost of the faked delivery itself — it's about the systemic drag created by returned parcels flooding back into the pipeline, reshuffling routes, and degrading operational flow. The authors partnered with a leading Indian last-mile delivery firm and used instrumental variable (IV) regression to isolate the causal effect. The IV approach is the methodological spine here — naive OLS would confuse misconduct with route difficulty, weather, or demand spikes. By instrumenting for misconduct, they claim to separate the behavioral signal from the operational noise. The paper also examines correlates: task complexity and opportunistic circumstances (loosely supervised periods, hard-to-verify addresses) predict higher misconduct rates. On the ladder, this paper is working in a space where most prior LMD research focuses on routing optimization, incentive design, and technology (GPS tracking, delivery apps). The novelty is the behavioral economics angle — treating DA misconduct as an endogenous productivity variable rather than a compliance problem to be engineered away. There isn't a direct SOTA baseline to beat because the question itself is relatively fresh: how much does worker gaming of monitoring systems cost in spillover, not just direct loss? The 1.60% and 1.86% numbers are first-of-kind estimates for this specific mechanism. The integrity picture is mixed. The IV strategy is the right instinct, but the abstract doesn't name the instrument, which is the single most important detail for evaluating whether the causal claim holds. Partnering with a single firm means the data is proprietary and non-replicable by outsiders — a structural limitation. The sample is from one Indian LMD company, so external validity to, say, Amazon Flex drivers in the US or European gig platforms is an open question. No pre-registration is mentioned. The practical takeaway for LMD operators is straightforward: monitoring systems that catch fake remarks after the fact are treating a symptom. The spillover math means every faked delivery costs more than its own reattempt — it taxes tomorrow's entire route. This argues for real-time verification (geofencing, photo proof at door) over retrospective auditing. For operations research, the paper opens a lane: if 1.6-1.9% daily productivity drag compounds across a fleet of thousands of DAs over weeks, the cumulative cost dwarfs the direct loss from individual faked parcels. The obvious next experiment the authors didn't run — or at least didn't report — is an intervention study. You've identified the spillover cost; now test whether a specific countermeasure (real-time geofencing, randomized audits, incentive redesign) actually reduces it and by how much. The honest read is this is being saved for the next paper, because the current contribution is cleanly scoped as 'here's the problem and its magnitude,' and the intervention is the natural sequel that doubles publication output.