Imagine you have an armored truck that can only carry 50 pounds of gold per trip, but you need to move 1,750 pounds to fund a hospital. You could wait for 35 trips — or you could figure out that 90% of what you're shipping is packaging, not gold, and compress the payload. That's what this paper does for information-theoretically secure (ITS) federated learning: the armored truck is physics-based key distribution (think QKD or similar), the gold is model updates, and the compression is a combination of frozen backbones, knowledge distillation, and quantization that shrinks what needs to be encrypted by ~35×. The committed claim: by combining three well-known ML compression techniques — freezing pretrained backbone layers, distilling knowledge from a teacher model, and quantizing the remaining trainable parameters — you can reduce the per-round key material needed for secure aggregation in federated learning enough to sustain real training on a physical key distribution testbed without exhausting the key buffer. This is demonstrated on a chest X-ray classification task (CheXpert dataset) across a simulated multi-institutional FL setup with real key generation hardware. The architectural recipe is straightforward. They take a DenseNet-121 pretrained on ImageNet, freeze most of it, attach a small trainable head, distill from a centrally-trained teacher, and quantize the trainable parameters to 8-bit integers. Only the small head's updates need to be masked during secure aggregation, so only those updates consume key material. The masking scheme is standard additive masking (SecAgg-style), but the key exchange is ITS rather than Diffie-Hellman — meaning an attacker with unlimited compute still can't break it. The constraint is that physics-based key generation produces keys at finite rates (bits per second), and keys expire, so you can't just stockpile indefinitely. Where this sits on the ladder: the individual components — frozen backbones, knowledge distillation, quantization, SecAgg — are all established. The novelty is integrating them specifically to solve the key-budget bottleneck of ITS key distribution, and validating on real hardware rather than simulating the key generation rates. No prior work has benchmarked FL under real physics-based key constraints with these compression techniques combined. The accuracy comparison shows the compressed pipeline maintains classification performance comparable to the uncompressed baseline, though the paper's metrics are on CheXpert (a standard but not ultra-competitive benchmark). Integrity-wise, the validation is mixed. The key distribution testbed is real hardware — a genuine physics-based system, not a simulation of one — which is the paper's strongest claim. But the FL topology is simulated (they don't have multiple hospitals actually training), the dataset is a single standard benchmark, and there's no comparison against alternative secure aggregation approaches (homomorphic encryption, trusted execution environments) or against other compression-for-security pipelines. The accuracy numbers are presented but not deeply ablated across different compression ratios or client counts. The milestone that matters: this paper shows ~35× key reduction is achievable while maintaining accuracy on a medical imaging task. The next number to watch is whether this scales to larger models (current trainable parameters are small by design) and whether the key generation rate can support FL with dozens of real institutional clients, not just the simulated topology used here. The gap between 'works on a testbed with simulated clients' and 'deployed across a hospital network with real key infrastructure' is significant — probably 3-5 years of engineering and standardization. The obvious experiment they didn't run: scaling to more clients and larger models, and benchmarking against homomorphic encryption or trusted execution environments as alternative security approaches. The honest read is (a) — the physics testbed is expensive and rare, and multi-site real deployment requires institutional partnerships they likely don't have yet. The paper is a proof-of-concept bridge between the ML compression community and the quantum/physics-based security community, and it succeeds at that narrow goal.