Imagine you're composing a song in a DAW. Some tracks are locked — the drum pattern is exactly what you laid down, no AI touching it. Other tracks are set to 'suggest': the model can improvise a bass line that fits the harmonic constraints but isn't note-for-note dictated. SpaceFlow does this for 3D generation. Each geometric primitive in a scene gets its own control dial, from 'follow this shape exactly' to 'go wild within this bounding region.' That per-part control is the core mechanism, and it's genuinely new territory for text-to-3D pipelines. The committed claim: SpaceFlow is the first training-free pipeline that provides explicit local control over both geometry fidelity and appearance routing in 3D generation. Users supply text descriptions plus a collection of geometric primitives, each tagged with a control level and a text or image cue. The system enforces spatial constraints during structure generation via a flow-based process, then segments the output and routes appearance conditioning per-part to prevent cross-region leakage. No fine-tuning, no retraining — it plugs into existing generative flow models. On the ladder, the paper positions itself against global-control-strength methods that treat an entire scene uniformly. Regional geometry metrics show that high-control regions preserve input shape faithfully while low-control regions produce plausible generative variation — a tradeoff no prior system let users dial explicitly. For appearance on fixed geometry, their text-conditioned routing achieves state-of-the-art prompt faithfulness and color/material accuracy. But the baselines are somewhat underspecified in the abstract — we don't get named competitors with hard numbers here, which limits confidence in the SOTA claim. Architecturally, SpaceFlow lives in the flow-matching family for 3D generation (think rectified flows, not diffusion score-matching), with a spatial constraint injection mechanism during the denoising trajectory. The segmentation-and-routing step for appearance is a clean design choice: segment the generated structure, match segments to primitives, condition each segment independently. This avoids the cross-part color bleeding that plagues holistic conditioning approaches. The 'training-free' claim means it works as a wrapper around existing flow-based 3D generators — a pipeline, not a new model. Integrity is mixed. The paper includes regional geometry metrics (quantitative) and a user study (qualitative), which is better than vibes-only evaluation. But the user study measures 'competitive overall quality' — a softer bar than 'wins.' The absence of named baselines with side-by-side numbers in the abstract makes it hard to pin down exactly how large the improvement is. No code release is mentioned, though a project page exists. Pre-registration is not applicable here, but reproducibility depends on access to the specific flow-based backbone they used. The milestone question is where this gets interesting for practitioners. Right now, SpaceFlow demonstrates local control with geometric primitives — essentially boxes, spheres, and simple shapes as proxies for object parts. The next concrete threshold is compositional scene control at the level of full multi-object scenes with 10+ interacting parts, each with independent geometry-appearance specifications. That's the gap between 'cool demo' and 'usable creative tool.' Expect 1-2 years if the flow-based 3D generation backbone keeps scaling. The obvious experiment not run: image-conditioned appearance routing at scale. The paper shows qualitative results for image cues but reports quantitative SOTA numbers only for text-conditioned routing. The honest read is probably (a) — image-conditioned metrics are harder to benchmark cleanly, and the text results were stronger, so they led with those. The next paper will almost certainly include quantitative image-appearance benchmarks and possibly video-conditioned appearance transfer.