{"id":{"repo_id":"mit","oai_identifier":"oai:dspace.mit.edu:1721.1/144903"},"canonical_url":"https://search.dev.ndltd.org/etd/mit/oai:dspace.mit.edu:1721.1/144903","repository":{"repo_id":"mit","name":"MIT","base_url":"https://dspace.mit.edu/oai/request"},"display":{"title":"Monkey: A Distributed Orchestrator for a Virtual Pseudo-Homogenous Computational Cluster Consisting of Heterogeneous Sources","abstract":"As machine learning research becomes increasingly ubiquitous, novel algorithms and state-of-the-art models are progressing to an advanced state with considerably more complex and involved procedures. That is, to achieve groundbreaking results in such a climate, a researcher increasingly depends upon immense computational requisites to develop, train, and evaluate such algorithms. As a result, research labs are faced with the challenge of providing ample computational resources, and researchers are detracted from their core research in order to design, code, and configure experiments for the disparate computational resources provided. The framework proposed herein, therefore, strives to bridge the gaps between research labs, researchers, and computational resources by abstracting and automating the standard process of designing, training, and evaluating an algorithm. This framework, built upon the preexisting Monkey framework, will provide a fault-tolerant, decentralized system that is capable of scheduling and reproducing research training jobs. The framework maintains a virtual pseudo-homogenous cluster built on top of existing heterogeneous computational clusters. Moreover, the framework, designed to be flexible and cost-effective, also prioritizes user accessibility by providing access to an integrated machine learning toolkit with hyperparameter optimizers and a visualization dashboard.","abstract_html":"As machine learning research becomes increasingly ubiquitous, novel algorithms and state-of-the-art models are progressing to an advanced state with considerably more complex and involved procedures. That is, to achieve groundbreaking results in such a climate, a researcher increasingly depends upon immense computational requisites to develop, train, and evaluate such algorithms. As a result, research labs are faced with the challenge of providing ample computational resources, and researchers are detracted from their core research in order to design, code, and configure experiments for the disparate computational resources provided. The framework proposed herein, therefore, strives to bridge the gaps between research labs, researchers, and computational resources by abstracting and automating the standard process of designing, training, and evaluating an algorithm. This framework, built upon the preexisting Monkey framework, will provide a fault-tolerant, decentralized system that is capable of scheduling and reproducing research training jobs. The framework maintains a virtual pseudo-homogenous cluster built on top of existing heterogeneous computational clusters. Moreover, the framework, designed to be flexible and cost-effective, also prioritizes user accessibility by providing access to an integrated machine learning toolkit with hyperparameter optimizers and a visualization dashboard.","abstract_has_math":false,"creators":["Stallone, Matthew J."],"institution":"Massachusetts Institute of Technology","degree_name":"Master","degree_level":null,"degree_discipline":null,"degree_department":"Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science","school":null,"contributors":[],"advisors":["Agrawal, Pulkit"],"committee_chairs":[],"committee_members":[],"year":2022,"date_issued":"2022-05","date_published":"2022-05","updated_at":"2026-07-22T22:21:15Z","subjects":[],"languages":[],"rights":["In Copyright - Educational Use Permitted","Copyright MIT"],"rights_urls":["http://rightsstatements.org/page/InC-EDU/1.0/"],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/1721.1/144903","outbound_label":"Handle","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor.advisor","label":"Advisor","values":["Agrawal, Pulkit"]},{"key":"dc:contributor.department","label":"Department","values":["Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science"]},{"key":"dc:creator","label":"Author","values":["Stallone, Matthew J."]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2022-08-29T16:19:53Z"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2022-08-29T16:19:53Z"]},{"key":"dc:date.issued","label":"Date","values":["2022-05"]},{"key":"dc:publisher","label":"Institution","values":["Massachusetts Institute of Technology"]},{"key":"dc:type","label":"Dc Type","values":["Thesis"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Master","Master of Engineering in Electrical Engineering and Computer Science"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:rights","label":"Dc Rights","values":["In Copyright - Educational Use Permitted","Copyright MIT"]},{"key":"dc:rights.uri","label":"Rights URI","values":["http://rightsstatements.org/page/InC-EDU/1.0/"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["https://hdl.handle.net/1721.1/144903"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["As machine learning research becomes increasingly ubiquitous, novel algorithms and state-of-the-art models are progressing to an advanced state with considerably more complex and involved procedures. That is, to achieve groundbreaking results in such a climate, a researcher increasingly depends upon immense computational requisites to develop, train, and evaluate such algorithms. As a result, research labs are faced with the challenge of providing ample computational resources, and researchers are detracted from their core research in order to design, code, and configure experiments for the disparate computational resources provided. The framework proposed herein, therefore, strives to bridge the gaps between research labs, researchers, and computational resources by abstracting and automating the standard process of designing, training, and evaluating an algorithm. This framework, built upon the preexisting Monkey framework, will provide a fault-tolerant, decentralized system that is capable of scheduling and reproducing research training jobs. The framework maintains a virtual pseudo-homogenous cluster built on top of existing heterogeneous computational clusters. Moreover, the framework, designed to be flexible and cost-effective, also prioritizes user accessibility by providing access to an integrated machine learning toolkit with hyperparameter optimizers and a visualization dashboard."]},{"key":"dc:description.degree","label":"Dc Description Degree","values":["M.Eng."]},{"key":"dc:title","label":"Title","values":["Monkey: A Distributed Orchestrator for a Virtual Pseudo-Homogenous Computational Cluster Consisting of Heterogeneous Sources"]}]}],"canonical_facts":{"dc:contributor.advisor":["Agrawal, Pulkit"],"dc:contributor.department":["Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science"],"dc:creator":["Stallone, Matthew J."],"dc:date.accessioned":["2022-08-29T16:19:53Z"],"dc:date.available":["2022-08-29T16:19:53Z"],"dc:date.issued":["2022-05"],"dc:description.abstract":["As machine learning research becomes increasingly ubiquitous, novel algorithms and state-of-the-art models are progressing to an advanced state with considerably more complex and involved procedures. That is, to achieve groundbreaking results in such a climate, a researcher increasingly depends upon immense computational requisites to develop, train, and evaluate such algorithms. As a result, research labs are faced with the challenge of providing ample computational resources, and researchers are detracted from their core research in order to design, code, and configure experiments for the disparate computational resources provided. The framework proposed herein, therefore, strives to bridge the gaps between research labs, researchers, and computational resources by abstracting and automating the standard process of designing, training, and evaluating an algorithm. This framework, built upon the preexisting Monkey framework, will provide a fault-tolerant, decentralized system that is capable of scheduling and reproducing research training jobs. The framework maintains a virtual pseudo-homogenous cluster built on top of existing heterogeneous computational clusters. Moreover, the framework, designed to be flexible and cost-effective, also prioritizes user accessibility by providing access to an integrated machine learning toolkit with hyperparameter optimizers and a visualization dashboard."],"dc:description.degree":["M.Eng."],"dc:identifier.uri":["https://hdl.handle.net/1721.1/144903"],"dc:publisher":["Massachusetts Institute of Technology"],"dc:rights":["In Copyright - Educational Use Permitted","Copyright MIT"],"dc:rights.uri":["http://rightsstatements.org/page/InC-EDU/1.0/"],"dc:title":["Monkey: A Distributed Orchestrator for a Virtual Pseudo-Homogenous Computational Cluster Consisting of Heterogeneous Sources"],"dc:type":["Thesis"],"thesis:degree_name":["Master","Master of Engineering in Electrical Engineering and Computer Science"]},"updated_at":"2026-07-22T22:21:15Z"}