{"id":{"repo_id":"mit","oai_identifier":"oai:dspace.mit.edu:1721.1/144623"},"canonical_url":"https://search.dev.ndltd.org/etd/mit/oai:dspace.mit.edu:1721.1/144623","repository":{"repo_id":"mit","name":"MIT","base_url":"https://dspace.mit.edu/oai/request"},"display":{"title":"Structuring Optimal Control of Legged Locomotion with Learning-based Methods","abstract":"Both optimal control methods and learning-based methods have been widely used for the control of legged locomotion. While optimal control formulations allow the designer to guarantee constraints on the solutions found, learning-based methods can leverage data and past experiences to globally search for solutions robust to noise and errors in the model parameters. This work explores how optimal control methods can be guided and structured by using data-driven techniques such as supervised learning and Bayesian optimization. Two case studies are presented. The first presents a model predictive controller for quadrupedal landing that reasons about body states, reaction forces, and contact timings in an online fashion. This highly nonlinear problem is made tractable by collecting thousands of feasible trajectories offline from trajectory optimizations and learning to generate them from the initial falling conditions. By initializing the search for a solution with the approximation from a deep neural network, the MIT Mini Cheetah is shown to be able to recover from significant falls in simulation and hardware in real time. The second studies the effect of learning command-dependent weights for a convex model predictive controller. The weights for the running costs are adjusted dynamically as a function of the command input. Using black-box optimization techniques and a defined higher-level reward, a function mapping the command input to the weights can be determined by sampling the trajectories from sweeps of command inputs.","abstract_html":"Both optimal control methods and learning-based methods have been widely used for the control of legged locomotion. While optimal control formulations allow the designer to guarantee constraints on the solutions found, learning-based methods can leverage data and past experiences to globally search for solutions robust to noise and errors in the model parameters. This work explores how optimal control methods can be guided and structured by using data-driven techniques such as supervised learning and Bayesian optimization. Two case studies are presented. The first presents a model predictive controller for quadrupedal landing that reasons about body states, reaction forces, and contact timings in an online fashion. This highly nonlinear problem is made tractable by collecting thousands of feasible trajectories offline from trajectory optimizations and learning to generate them from the initial falling conditions. By initializing the search for a solution with the approximation from a deep neural network, the MIT Mini Cheetah is shown to be able to recover from significant falls in simulation and hardware in real time. The second studies the effect of learning command-dependent weights for a convex model predictive controller. The weights for the running costs are adjusted dynamically as a function of the command input. Using black-box optimization techniques and a defined higher-level reward, a function mapping the command input to the weights can be determined by sampling the trajectories from sweeps of command inputs.","abstract_has_math":false,"creators":["Jeon, Se Hwan"],"institution":"Massachusetts Institute of Technology","degree_name":"Master","degree_level":null,"degree_discipline":null,"degree_department":"Massachusetts Institute of Technology. Department of Mechanical Engineering","school":null,"contributors":[],"advisors":["Kim, Sangbae"],"committee_chairs":[],"committee_members":[],"year":2022,"date_issued":"2022-05","date_published":"2022-05","updated_at":"2026-07-22T22:21:46Z","subjects":[],"languages":[],"rights":["In Copyright - Educational Use Permitted","Copyright MIT"],"rights_urls":["http://rightsstatements.org/page/InC-EDU/1.0/"],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/1721.1/144623","outbound_label":"Handle","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor.advisor","label":"Advisor","values":["Kim, Sangbae"]},{"key":"dc:contributor.department","label":"Department","values":["Massachusetts Institute of Technology. Department of Mechanical Engineering"]},{"key":"dc:creator","label":"Author","values":["Jeon, Se Hwan"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2022-08-29T16:00:17Z"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2022-08-29T16:00:17Z"]},{"key":"dc:date.issued","label":"Date","values":["2022-05"]},{"key":"dc:publisher","label":"Institution","values":["Massachusetts Institute of Technology"]},{"key":"dc:type","label":"Dc Type","values":["Thesis"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Master","Master of Science in Mechanical Engineering"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:rights","label":"Dc Rights","values":["In Copyright - Educational Use Permitted","Copyright MIT"]},{"key":"dc:rights.uri","label":"Rights URI","values":["http://rightsstatements.org/page/InC-EDU/1.0/"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["https://hdl.handle.net/1721.1/144623"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["Both optimal control methods and learning-based methods have been widely used for the control of legged locomotion. While optimal control formulations allow the designer to guarantee constraints on the solutions found, learning-based methods can leverage data and past experiences to globally search for solutions robust to noise and errors in the model parameters. This work explores how optimal control methods can be guided and structured by using data-driven techniques such as supervised learning and Bayesian optimization. Two case studies are presented. The first presents a model predictive controller for quadrupedal landing that reasons about body states, reaction forces, and contact timings in an online fashion. This highly nonlinear problem is made tractable by collecting thousands of feasible trajectories offline from trajectory optimizations and learning to generate them from the initial falling conditions. By initializing the search for a solution with the approximation from a deep neural network, the MIT Mini Cheetah is shown to be able to recover from significant falls in simulation and hardware in real time. The second studies the effect of learning command-dependent weights for a convex model predictive controller. The weights for the running costs are adjusted dynamically as a function of the command input. Using black-box optimization techniques and a defined higher-level reward, a function mapping the command input to the weights can be determined by sampling the trajectories from sweeps of command inputs."]},{"key":"dc:description.degree","label":"Dc Description Degree","values":["S.M."]},{"key":"dc:title","label":"Title","values":["Structuring Optimal Control of Legged Locomotion with Learning-based Methods"]}]}],"canonical_facts":{"dc:contributor.advisor":["Kim, Sangbae"],"dc:contributor.department":["Massachusetts Institute of Technology. Department of Mechanical Engineering"],"dc:creator":["Jeon, Se Hwan"],"dc:date.accessioned":["2022-08-29T16:00:17Z"],"dc:date.available":["2022-08-29T16:00:17Z"],"dc:date.issued":["2022-05"],"dc:description.abstract":["Both optimal control methods and learning-based methods have been widely used for the control of legged locomotion. While optimal control formulations allow the designer to guarantee constraints on the solutions found, learning-based methods can leverage data and past experiences to globally search for solutions robust to noise and errors in the model parameters. This work explores how optimal control methods can be guided and structured by using data-driven techniques such as supervised learning and Bayesian optimization. Two case studies are presented. The first presents a model predictive controller for quadrupedal landing that reasons about body states, reaction forces, and contact timings in an online fashion. This highly nonlinear problem is made tractable by collecting thousands of feasible trajectories offline from trajectory optimizations and learning to generate them from the initial falling conditions. By initializing the search for a solution with the approximation from a deep neural network, the MIT Mini Cheetah is shown to be able to recover from significant falls in simulation and hardware in real time. The second studies the effect of learning command-dependent weights for a convex model predictive controller. The weights for the running costs are adjusted dynamically as a function of the command input. Using black-box optimization techniques and a defined higher-level reward, a function mapping the command input to the weights can be determined by sampling the trajectories from sweeps of command inputs."],"dc:description.degree":["S.M."],"dc:identifier.uri":["https://hdl.handle.net/1721.1/144623"],"dc:publisher":["Massachusetts Institute of Technology"],"dc:rights":["In Copyright - Educational Use Permitted","Copyright MIT"],"dc:rights.uri":["http://rightsstatements.org/page/InC-EDU/1.0/"],"dc:title":["Structuring Optimal Control of Legged Locomotion with Learning-based Methods"],"dc:type":["Thesis"],"thesis:degree_name":["Master","Master of Science in Mechanical Engineering"]},"updated_at":"2026-07-22T22:21:46Z"}