{"id":{"repo_id":"mit","oai_identifier":"oai:dspace.mit.edu:1721.1/139247"},"canonical_url":"https://search.dev.ndltd.org/etd/mit/oai:dspace.mit.edu:1721.1/139247","repository":{"repo_id":"mit","name":"MIT","base_url":"https://dspace.mit.edu/oai/request"},"display":{"title":"Physically Constrained PCB Placement Using Deep Reinforcement Learning","abstract":"This thesis provides an in depth exploration of Reinforcement Learning (RL) based PCB component placement with emphasis on physically verified placements. Unlike prior methods that rely on heuristic proxies for placement quality, this work focuses entirely on routing based metrics that result in functioning placements without the need for fine tuning. Additionally, this exploration considers true use cases of PCB auto-placement where a human-in-the-loop pre-places a set list of components and the auto-placer places the remaining. This is achieved by first restricting the placement domain to only ring placements; a domain where routing calculations become accessible. Within the ring placement domain, an RL agent is trained to place components on a simulated PCB canvas such that there are no component overlaps or wire crossings upon manufacture. Through the use of an unbounded reward system, the agent is trained progressively with PCB complexity gradually increasing as training steps are run. The resulting placements are robust to varying numbers of components as well as component shape and size. Finally, this thesis concludes with a discussion about further work and challenges facing the future of PCB auto-placement.","abstract_html":"This thesis provides an in depth exploration of Reinforcement Learning (RL) based PCB component placement with emphasis on physically verified placements. Unlike prior methods that rely on heuristic proxies for placement quality, this work focuses entirely on routing based metrics that result in functioning placements without the need for fine tuning. Additionally, this exploration considers true use cases of PCB auto-placement where a human-in-the-loop pre-places a set list of components and the auto-placer places the remaining. This is achieved by first restricting the placement domain to only ring placements; a domain where routing calculations become accessible. Within the ring placement domain, an RL agent is trained to place components on a simulated PCB canvas such that there are no component overlaps or wire crossings upon manufacture. Through the use of an unbounded reward system, the agent is trained progressively with PCB complexity gradually increasing as training steps are run. The resulting placements are robust to varying numbers of components as well as component shape and size. Finally, this thesis concludes with a discussion about further work and challenges facing the future of PCB auto-placement.","abstract_has_math":false,"creators":["Crocker, Peter"],"institution":"Massachusetts Institute of Technology","degree_name":"Master","degree_level":null,"degree_discipline":null,"degree_department":"Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science","school":null,"contributors":[],"advisors":["Chan, Vincent W.S."],"committee_chairs":[],"committee_members":[],"year":2021,"date_issued":"2021-06","date_published":"2021-06","updated_at":"2026-07-22T22:21:55Z","subjects":[],"languages":[],"rights":["In Copyright - Educational Use Permitted","Copyright MIT"],"rights_urls":["http://rightsstatements.org/page/InC-EDU/1.0/"],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/1721.1/139247","outbound_label":"Handle","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor.advisor","label":"Advisor","values":["Chan, Vincent W.S."]},{"key":"dc:contributor.department","label":"Department","values":["Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science"]},{"key":"dc:creator","label":"Author","values":["Crocker, Peter"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2022-01-14T14:59:15Z"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2022-01-14T14:59:15Z"]},{"key":"dc:date.issued","label":"Date","values":["2021-06"]},{"key":"dc:publisher","label":"Institution","values":["Massachusetts Institute of Technology"]},{"key":"dc:type","label":"Dc Type","values":["Thesis"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Master","Master of Engineering in Electrical Engineering and Computer Science"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:rights","label":"Dc Rights","values":["In Copyright - Educational Use Permitted","Copyright MIT"]},{"key":"dc:rights.uri","label":"Rights URI","values":["http://rightsstatements.org/page/InC-EDU/1.0/"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["https://hdl.handle.net/1721.1/139247"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["This thesis provides an in depth exploration of Reinforcement Learning (RL) based PCB component placement with emphasis on physically verified placements. Unlike prior methods that rely on heuristic proxies for placement quality, this work focuses entirely on routing based metrics that result in functioning placements without the need for fine tuning. Additionally, this exploration considers true use cases of PCB auto-placement where a human-in-the-loop pre-places a set list of components and the auto-placer places the remaining. This is achieved by first restricting the placement domain to only ring placements; a domain where routing calculations become accessible. Within the ring placement domain, an RL agent is trained to place components on a simulated PCB canvas such that there are no component overlaps or wire crossings upon manufacture. Through the use of an unbounded reward system, the agent is trained progressively with PCB complexity gradually increasing as training steps are run. The resulting placements are robust to varying numbers of components as well as component shape and size. Finally, this thesis concludes with a discussion about further work and challenges facing the future of PCB auto-placement."]},{"key":"dc:description.degree","label":"Dc Description Degree","values":["M.Eng."]},{"key":"dc:title","label":"Title","values":["Physically Constrained PCB Placement Using Deep Reinforcement Learning"]}]}],"canonical_facts":{"dc:contributor.advisor":["Chan, Vincent W.S."],"dc:contributor.department":["Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science"],"dc:creator":["Crocker, Peter"],"dc:date.accessioned":["2022-01-14T14:59:15Z"],"dc:date.available":["2022-01-14T14:59:15Z"],"dc:date.issued":["2021-06"],"dc:description.abstract":["This thesis provides an in depth exploration of Reinforcement Learning (RL) based PCB component placement with emphasis on physically verified placements. Unlike prior methods that rely on heuristic proxies for placement quality, this work focuses entirely on routing based metrics that result in functioning placements without the need for fine tuning. Additionally, this exploration considers true use cases of PCB auto-placement where a human-in-the-loop pre-places a set list of components and the auto-placer places the remaining. This is achieved by first restricting the placement domain to only ring placements; a domain where routing calculations become accessible. Within the ring placement domain, an RL agent is trained to place components on a simulated PCB canvas such that there are no component overlaps or wire crossings upon manufacture. Through the use of an unbounded reward system, the agent is trained progressively with PCB complexity gradually increasing as training steps are run. The resulting placements are robust to varying numbers of components as well as component shape and size. Finally, this thesis concludes with a discussion about further work and challenges facing the future of PCB auto-placement."],"dc:description.degree":["M.Eng."],"dc:identifier.uri":["https://hdl.handle.net/1721.1/139247"],"dc:publisher":["Massachusetts Institute of Technology"],"dc:rights":["In Copyright - Educational Use Permitted","Copyright MIT"],"dc:rights.uri":["http://rightsstatements.org/page/InC-EDU/1.0/"],"dc:title":["Physically Constrained PCB Placement Using Deep Reinforcement Learning"],"dc:type":["Thesis"],"thesis:degree_name":["Master","Master of Engineering in Electrical Engineering and Computer Science"]},"updated_at":"2026-07-22T22:21:55Z"}