Open Questions
Combo maneuver drop — the only long-horizon compounding test falls to 95% for the full model and collapses to 0% for both ablations, with no diagnosis of why.
Subjective success metric — real-world "success" is a single safety pilot's judgment call, with no stated threshold or second observer.
Vision never randomized — IMU bias and thrust-to-weight are randomized ±10%, but scene geometry, lighting, and texture are not — despite flying in one fixed real-world room.
Generalization overreach — the conclusion claims applicability "beyond autonomous flight," from one 1.15 kg platform in one space.