Imagine you're trying to solve a maze while wearing sunglasses. You spend weeks optimizing your route-finding strategy — left-hand rule vs. right-hand rule vs. Trémaux's algorithm — only to discover that swapping your cheap sunglasses for a clear pair of glasses improved your speed more than any strategy change ever did. That's the core finding of ForVis: the sensor you strap to a forest drone matters more than which state-of-the-art SLAM algorithm you run on it. The committed claim: ForVis is the first purpose-built VI-SLAM benchmark for under-canopy UAV flight in real forests, and its headline empirical finding is that sensor selection (OAK-D Pro Wide vs. Intel RealSense D435i) has a larger effect on trajectory error than the spread across all seven tested algorithms. This is a dataset paper with an empirical finding baked in, not a methods paper. The dataset itself covers twelve flights across three distinct forest conditions — open meadow, above-canopy, and under-canopy — totaling 563.8 seconds of flight over 1,096.8 meters of trajectory. Both cameras recorded simultaneously, meaning sensor comparisons are perfectly paired: same flight, same vibrations, same lighting. The seven benchmarked VI-SLAM systems (the paper doesn't name them in the abstract, but the 504-run count implies roughly 504 / 12 flights / 2 sensors ≈ 3 runs per system-sensor-flight combination) all achieved lower median error on the OAK-D Pro Wide than on the D435i. Every single one. Architecturally, this sits in the visual-inertial odometry family — fusing camera frames with IMU data to estimate drone pose in real-time. The paper doesn't propose a new algorithm; it provides the testing ground. The challenge is forests specifically: repetitive textures (trees look like trees), dappled and shifting illumination, and vibration from rotors. These are the conditions where lab-tested SLAM systems go to die, and until now there was no standard benchmark for it. The integrity picture is mixed but honest for a dataset paper. The simultaneous dual-sensor recording is strong experimental design — it eliminates confounds from different flight conditions. But ground truth methodology is the open question: the abstract doesn't specify how trajectory ground truth was obtained (RTK-GPS? Total station? Motion capture is impossible outdoors under canopy). Without knowing the ground truth accuracy, the error comparisons between sensors have an unknown floor. The 504-run benchmark across seven systems is substantial enough to be credible, not cherry-picked. The milestone question for forest drone navigation is autonomy duration: current under-canopy flights are short (ForVis averages ~47 seconds per flight). The field needs 10+ minute autonomous flights with sub-meter accuracy to enable real forestry applications — timber inventory, ecological monitoring, search and rescue. ForVis establishes a baseline; the next concrete milestone is whether any VI-SLAM system can maintain sub-1m ATE over a 600-second under-canopy flight. That's roughly 12× the current average flight duration. The obvious experiment not run: testing with newer event cameras or LiDAR-visual fusion, which are the two hardware directions most likely to leapfrog stereo cameras in forest conditions. The honest read is (a) — budget and equipment constraints. Event cameras and forest-grade LiDAR units are expensive and not yet standard on research UAVs. The authors built what they could with two commercially available sensors, which is the right first move for a benchmark paper.