{"id":241,"date":"2025-08-04T10:23:22","date_gmt":"2025-08-04T14:23:22","guid":{"rendered":"https:\/\/sites.bu.edu\/tron\/?p=241"},"modified":"2025-08-04T10:23:22","modified_gmt":"2025-08-04T14:23:22","slug":"sketching-for-efficient-and-robust-top-down-autonomous-navigation","status":"publish","type":"post","link":"https:\/\/sites.bu.edu\/tron\/2025\/08\/04\/sketching-for-efficient-and-robust-top-down-autonomous-navigation\/","title":{"rendered":"Sketching for Efficient and Robust Top-Down Autonomous Navigation"},"content":{"rendered":"<h3>Overview<\/h3>\n<div class=\"page\" title=\"Page 17\">\n<div class=\"layoutArea\">\n<div class=\"column\">\n<p><span>Inspired by the human ability to efficiently and robustly navigate in a diversity of environments, our project aims to develop a novel top-down, low-resource robotic navigation approach in unmapped environments. Comparing state-of-the-art solutions in robotics with their natural counterparts, this project focuses on three opportunities:<\/span><\/p>\n<ol>\n<li><span>Combine machine learning and optimization techniques to sketch a high-level representation of the environment composed of semantically meaningful units.<\/span><\/li>\n<li><span>Generate a collection feedback controllers, with rigorous performance guarantees, that can be used to both navigate and improve the representation of the environment.<\/span><\/li>\n<li>C<span>onsider multi-agent settings where robots in a team use collaboration to further reduce sensing requirements, and reuse previous experience to reduce computation requirements.\u00a0<\/span><\/li>\n<\/ol>\n<div class=\"page\" title=\"Page 17\">\n<div class=\"layoutArea\">\n<div class=\"column\">\n<h3>BoxMap: High-Level Mapping and Navigation Supported by Machine Learning<\/h3>\n<p>BoxMap is a novel high-level mapping method that uses machine learning algorithms to exploit the structure of sensed partial environments, and update a topological map representing semantic entities (rooms and doors) and their relations.<\/p>\n<p>BoxMap has two main components: machine learning for extracting and fusing high-level entities from measurements, and graph-based map construction and navigation.<\/p>\n<h4>Machine Learning for Extracting and Fusing High-Level Entities from Measurements<\/h4>\n<p>We developed a new deep learning architecture that combines parametric representations (for semantic entities) with non-parametric partial maps (to fuse measurements). The architecture includes:<\/p>\n<ul>\n<li>\u00a0A Convolutional Neural Network backbone plus a Detection Transformer with gating to extract features from low-level measurements (laser scans) into estimate of parameters of semantic entities.<\/li>\n<li>Hand-crafted ReLU layers that translate, in a differentiable way, our parametric representation into non-parametric Truncated Distance Sign Functions (TSDFs).<\/li>\n<li>Loss functions that compare non-parametric representations. The loss is hierarchical, as it first computes a loss on rooms only, subtracts their footprint, and then focuses on doors.<\/li>\n<\/ul>\n<figure id=\"attachment242\" aria-describedby=\"caption-attachment242\" style=\"width: 1034px\" class=\"wp-caption alignnone\"><img loading=\"lazy\" src=\"\/tron\/files\/2025\/08\/dl_architecture-1024x242.png\" alt=\"Conceptual diagram showing a CNN backbone, a transformer-based encoder-decoder producing a set of room and door predictions, and a map-based loss.\" width=\"1024\" height=\"242\" class=\"wp-image-242 size-large\" srcset=\"https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/dl_architecture-1024x242.png 1024w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/dl_architecture-636x150.png 636w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/dl_architecture-768x181.png 768w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/dl_architecture-1536x363.png 1536w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/dl_architecture-2048x484.png 2048w\" sizes=\"(max-width: 1024px) 100vw, 1024px\" \/><figcaption id=\"caption-attachment242\" class=\"wp-caption-text\">Overview of the machine learning component of BoxMap: a CNN backbone, a transformer-based encoder-decoder, hand-crafted layers producing a set of room and door predictions, and a TSDF-based loss.<\/figcaption><\/figure>\n<h4>Mapping component<\/h4>\n<p>A semantic map module builds a graph of semantic entities (rooms connected by doors). The graph is updated by transforming it into a local TSDF, and then fusing it with new measurements using the neural network above. The fusion produces new candidates for rooms and doors, which are then incorporated in the topological map via simple overlap tests.<\/p>\n<figure id=\"attachment243\" aria-describedby=\"caption-attachment243\" style=\"width: 1034px\" class=\"wp-caption alignnone\"><img loading=\"lazy\" src=\"\/tron\/files\/2025\/08\/pred_steps-1024x716.png\" alt=\"A diagram with multiple rows of images. The first row shows the cumulative occupancy grid generated from the current graph map. The second row shows the current measurements. The third row shows the result of the inference of the machine learning module. The fourth row shows the current, explored, unexplored rooms and a planned path. The fifth row shows the path of the robot.\" width=\"1024\" height=\"716\" class=\"wp-image-243 size-large\" srcset=\"https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/pred_steps-1024x716.png 1024w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/pred_steps-636x445.png 636w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/pred_steps-768x537.png 768w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/pred_steps-1536x1074.png 1536w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/pred_steps-2048x1433.png 2048w\" sizes=\"(max-width: 1024px) 100vw, 1024px\" \/><figcaption id=\"caption-attachment243\" class=\"wp-caption-text\">Map construction and exploration using BoxMap<\/figcaption><\/figure>\n<h4>Testing<\/h4>\n<p>BoxMap has been tested in a Python-only simulation (based on pseudoSLAM), and then in a ROS-Gazebo simulation.<\/p>\n<div class=\"page\" title=\"Page 1\">\n<div class=\"layoutArea\">\n<div class=\"column\">\n<p><span>Our BoxMap representation scales quadratically with the number of rooms (with a small constant), resulting in significant savings over a full geometric map. Moreover, our high-level topological representation results in <\/span><span>23<\/span><span>.<\/span><span>9% <\/span><span>shorter trajectories in the exploration task with respect to standard methods. <\/span><\/p>\n<\/div>\n<\/div>\n<\/div>\n<h3>Lyapunov Control with Monotonic Layers for Navigation in BoxMap<\/h3>\n<figure id=\"attachment245\" aria-describedby=\"caption-attachment245\" style=\"width: 1850px\" class=\"wp-caption alignnone\"><img loading=\"lazy\" src=\"\/tron\/files\/2025\/08\/Lyapunov-neural-networks.png\" alt=\"\" width=\"1840\" height=\"1215\" class=\"size-full wp-image-245\" srcset=\"https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/Lyapunov-neural-networks.png 1840w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/Lyapunov-neural-networks-636x420.png 636w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/Lyapunov-neural-networks-1024x676.png 1024w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/Lyapunov-neural-networks-768x507.png 768w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/Lyapunov-neural-networks-1536x1014.png 1536w\" sizes=\"(max-width: 1840px) 100vw, 1840px\" \/><figcaption id=\"caption-attachment245\" class=\"wp-caption-text\">Lyapunov monotone neural networks have layers that compose three operations: projections of the input along given directions; monotone neurons, which output monotonically increasing piecewise linear functions of the projections; sums of all monotone neurons.<\/figcaption><\/figure>\n<p>We developed Lyapunov monotonic neural networks, a novel deep learning architecture to encode Lyapunov functions. Our architecture ensures, by construction, that Lyapunov functions are unimodal and quasi-convex (i.e., with a unique minimum at the origin and star-convex level sets).<\/p>\n<figure id=\"attachment247\" aria-describedby=\"caption-attachment247\" style=\"width: 625px\" class=\"wp-caption aligncenter\"><img loading=\"lazy\" src=\"\/tron\/files\/2025\/08\/RoA-Cart-Pole-615x636.png\" alt=\"\" width=\"615\" height=\"636\" class=\"wp-image-247 size-medium\" srcset=\"https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/RoA-Cart-Pole-615x636.png 615w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/RoA-Cart-Pole-991x1024.png 991w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/RoA-Cart-Pole-768x794.png 768w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/RoA-Cart-Pole.png 1137w\" sizes=\"(max-width: 615px) 100vw, 615px\" \/><figcaption id=\"caption-attachment247\" class=\"wp-caption-text\">Verified region of attraction for a 4-D cart-pole system before (blue) and after (red) maximization.<\/figcaption><\/figure>\n<p>When paired with a piecewise linear feedback controller and a piecewise linear model for the dynamics, our Lyapunov monotonic neural networks enable rigorous verification of stability over a Region of Attraction (RoA) using Mixed-Integer Linear Programming (MILP). The result of the verification can be used to train the Lyapunov function and the controller together over a fixed RoA, or increase the RoA.<\/p>\n<\/div>\n<div class=\"column\">\n<p>In the context of BoxMap, this methodology has been applied to synthesize the control of a unicycle in rooms connected by doors (i.e., compatible with the high-level representation used by BoxMap).<\/p>\n<figure id=\"attachment250\" aria-describedby=\"caption-attachment250\" style=\"width: 1034px\" class=\"wp-caption alignnone\"><img loading=\"lazy\" src=\"\/tron\/files\/2025\/08\/output-1024x482.png\" alt=\"\" width=\"1024\" height=\"482\" class=\"wp-image-250 size-large\" srcset=\"https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/output-1024x482.png 1024w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/output-636x300.png 636w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/output-768x362.png 768w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/output-1536x723.png 1536w, https:\/\/sites.bu.edu\/tron\/files\/2025\/08\/output-2048x965.png 2048w\" sizes=\"(max-width: 1024px) 100vw, 1024px\" \/><figcaption id=\"caption-attachment250\" class=\"wp-caption-text\">Controller synthesis for a unicycle model navigating in a rectangular room toward a door (right side). The environment is divided in three regions, with a separate controller for each region. Left: control field for a zero-heading angle. Right: resulting trajectories.<\/figcaption><\/figure>\n<h3><strong>Funding and support<\/strong><\/h3>\n<p><img loading=\"lazy\" src=\"\/tron\/files\/2025\/08\/nsf_logo-150x150.png\" alt=\"\" width=\"150\" height=\"150\" class=\"alignright wp-image-251 size-thumbnail\" \/>This project is supported by the National Science Foundation grant <a href=\"https:\/\/www.nsf.gov\/awardsearch\/showAward?AWD_ID=2409733&amp;HistoricalAwards=false\">&#8220;Sketching for Efficient and Robust Top-Down Autonomous Navigation&#8221; (Award number 2409733)\u00a0<\/a><\/p>\n<p>Start date: September 1, 2024<br \/>\nEnd date: July 31, 2027<\/p>\n<div class=\"page\" title=\"Page 1\">\n<div class=\"layoutArea\">\n<div class=\"column\">\n<p><span><em>Disclaimer: Any opinions, findings, and conclusions or recommendations expressed in this material are those of the author(s) and do not necessarily reflect the views of the National Science Foundation.<\/em><\/span><\/p>\n<\/div>\n<\/div>\n<\/div>\n<\/div>\n<\/div>\n<\/div>\n<\/div>\n<\/div>\n<\/div>\n","protected":false},"excerpt":{"rendered":"<p>Overview Inspired by the human ability to efficiently and robustly navigate in a diversity of environments, our project aims to develop a novel top-down, low-resource robotic navigation approach in unmapped environments. Comparing state-of-the-art solutions in robotics with their natural counterparts, this project focuses on three opportunities: Combine machine learning and optimization techniques to sketch a [&hellip;]<\/p>\n","protected":false},"author":11510,"featured_media":0,"comment_status":"closed","ping_status":"open","sticky":false,"template":"","format":"standard","meta":[],"categories":[4],"tags":[],"_links":{"self":[{"href":"https:\/\/sites.bu.edu\/tron\/wp-json\/wp\/v2\/posts\/241"}],"collection":[{"href":"https:\/\/sites.bu.edu\/tron\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/sites.bu.edu\/tron\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/sites.bu.edu\/tron\/wp-json\/wp\/v2\/users\/11510"}],"replies":[{"embeddable":true,"href":"https:\/\/sites.bu.edu\/tron\/wp-json\/wp\/v2\/comments?post=241"}],"version-history":[{"count":2,"href":"https:\/\/sites.bu.edu\/tron\/wp-json\/wp\/v2\/posts\/241\/revisions"}],"predecessor-version":[{"id":252,"href":"https:\/\/sites.bu.edu\/tron\/wp-json\/wp\/v2\/posts\/241\/revisions\/252"}],"wp:attachment":[{"href":"https:\/\/sites.bu.edu\/tron\/wp-json\/wp\/v2\/media?parent=241"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/sites.bu.edu\/tron\/wp-json\/wp\/v2\/categories?post=241"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/sites.bu.edu\/tron\/wp-json\/wp\/v2\/tags?post=241"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}