Abstract
This review examines latent policy optimization for urban low-altitude safety through an evidence-centered design lens. The analysis asks how rerouting, detect-and-avoid, and recovery should exchange authority. It treats the relevant unit as a complete pathway from data or physical observations to representation, model output, human interpretation, and accountable action. The cited literature is synthesized without inventing experiments or unreported performance values. Particular attention is given to nominally independent safeguards sharing a failure source. The review argues that credible translation requires explicit evidence boundaries, uncertainty-aware evaluation, author-visible traceability, and a documented route for intervention. The resulting framework supports urban UAV operations while distinguishing component promise from system readiness.
References
Beard, R. W., & McLain, T. W. (2012). Small unmanned aircraft: Theory and practice. Princeton University Press.
Bengio, Y., Louradour, J., Collobert, R., & Weston, J. (2009). Curriculum learning. In Proceedings of the 26th International Conference on Machine Learning (pp. 41-48).
Deng, H., Luo, H., Zhu, Y., Li, L., Chen, Z., Zhao, X., Li, M., Zhang, J., Wang, M., Cao, Y., & Kang, Y. (2026). IIB-LPO: Latent policy optimization via iterative information bottleneck. arXiv preprint arXiv:2601.05870.
Floreano, D., & Wood, R. J. (2015). Science, technology and the future of small autonomous drones. Nature, 521, 460-466. https://doi.org/10.1038/nature14542
Han, Z., Chen, W., Han, Y., Mao, R., & Qin, J. (2026). Fast diversified top-k rule discovery via user-guided embeddings. IEEE Transactions on Knowledge and Data Engineering, 38, 1739–1753.
Kendoul, F. (2012). Survey of advances in guidance, navigation, and control of unmanned rotorcraft systems. Journal of Field Robotics, 29(2), 315-378. https://doi.org/10.1002/rob.20414
M. Xu and M. Jin, "Three-Stage Urban Low-Altitude Safety: Dynamic Geo-Fencing Rerouting + Remote ID/ADS-B Based Detect-and-Avoid + Power Failure/Crash Recovery," 2025 Low-Altitude Economy Forum & International Conference on Low-Altitude Flight Technology and Unmanned Aerial Vehicle Application (LEF & ICLU), Guangzhou, China, 2025, pp. 164-168, doi: 10.1109/LEFICLU65987.2025.11297342.
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., & Klimov, O. (2017). Proximal policy optimization algorithms. arXiv preprint arXiv:1707.06347.
Tao, J., Lyu, R., & Cao, X. (2026). A Deep Learning-Based Automated Content Moderation Framework for Online Platforms. Future-Adaptive Intelligence and Lifelong Systems, 1(1).
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., & Polosukhin, I. (2017). Attention is all you need. Advances in Neural Information Processing Systems, 30.
