{"id":{"repo_id":"mit","oai_identifier":"oai:dspace.mit.edu:1721.1/9630"},"canonical_url":"https://search.dev.ndltd.org/etd/mit/oai:dspace.mit.edu:1721.1/9630","repository":{"repo_id":"mit","name":"MIT","base_url":"https://dspace.mit.edu/oai/request"},"display":{"title":"Pose imagery and automated three-dimensional modeling of urban environments","abstract":"Three-dimensional (3-D) modeling of urban environments has numerous applications, including virtual environments, urban planning, and physical simulation. Construct­ing 3-D models from photographs (images) is thus an important area of research in computer vision, and increasingly, computer graphics. However, despite many years of research, a system that automatically recovers realistic 3-D models remains elusive; most practical systems require significant human input. Unlike automatic algorithms, human-assisted systems are not scalable, both in terms of the number of images processed and the complexity of the generated 3-D model. This thesis describes novel techniques to automatically extract textured 3-D mod­els of urban environments from pose imagery, i.e., images annotated with camera position and orientation in a single global coordinate system. Physical instruments (e.g., surveying, Global Positioning System (GPS), inertial sensors, etc.) are used to provide accurate initial pose estimates to the proposed algorithms. As these es­timates are not perfect, I first describe two optimization techniques that refine pose estimates using information present in the images: spherical mosaicing recovers rel­ative rotations between images taken from a single position, and mosaic registration accurately locates mosaics in a global coordinate system. Next, I describe an algo­rithm that extracts vertical facades from mosaics annotated with accurate pose. The algorithm employs horizontal line segments to detect likely facade orientations and locates these facades using a space-sweep technique. Textures are robustly computed for the facades by combining information from several mosaics using median statistics. I present results for a large pose image dataset ( consisting of about four thousand images taken from eighty-one positions) of an urban office complex. These techniques were successful in recovering all significant vertical facades in the complex, as well as several neighboring facades.","abstract_html":"Three-dimensional (3-D) modeling of urban environments has numerous applications, including virtual environments, urban planning, and physical simulation. Construct­ing 3-D models from photographs (images) is thus an important area of research in computer vision, and increasingly, computer graphics. However, despite many years of research, a system that automatically recovers realistic 3-D models remains elusive; most practical systems require significant human input. Unlike automatic algorithms, human-assisted systems are not scalable, both in terms of the number of images processed and the complexity of the generated 3-D model. This thesis describes novel techniques to automatically extract textured 3-D mod­els of urban environments from pose imagery, i.e., images annotated with camera position and orientation in a single global coordinate system. Physical instruments (e.g., surveying, Global Positioning System (GPS), inertial sensors, etc.) are used to provide accurate initial pose estimates to the proposed algorithms. As these es­timates are not perfect, I first describe two optimization techniques that refine pose estimates using information present in the images: spherical mosaicing recovers rel­ative rotations between images taken from a single position, and mosaic registration accurately locates mosaics in a global coordinate system. Next, I describe an algo­rithm that extracts vertical facades from mosaics annotated with accurate pose. The algorithm employs horizontal line segments to detect likely facade orientations and locates these facades using a space-sweep technique. Textures are robustly computed for the facades by combining information from several mosaics using median statistics. I present results for a large pose image dataset ( consisting of about four thousand images taken from eighty-one positions) of an urban office complex. These techniques were successful in recovering all significant vertical facades in the complex, as well as several neighboring facades.","abstract_has_math":false,"creators":["Coorg, Satyan R"],"institution":"Massachusetts Institute of Technology","degree_name":null,"degree_level":null,"degree_discipline":null,"degree_department":"Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science","school":null,"contributors":[],"advisors":["Seth Teller."],"committee_chairs":[],"committee_members":[],"year":1998,"date_issued":"1998","date_published":"1998","updated_at":"2026-07-22T22:21:22Z","subjects":["Electrical Engineering and Computer Science"],"languages":["eng"],"rights":["M.I.T. theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. See provided URL for inquiries about permission."],"rights_urls":["http://dspace.mit.edu/handle/1721.1/7582"],"identifier_entries":[]},"links":{"outbound_url":"http://hdl.handle.net/1721.1/9630","outbound_label":"Handle","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor.advisor","label":"Advisor","values":["Seth Teller."]},{"key":"dc:contributor.department","label":"Department","values":["Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science"]},{"key":"dc:creator","label":"Author","values":["Coorg, Satyan R"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2005-08-19T19:03:23Z"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2005-08-19T19:03:23Z"]},{"key":"dc:date.issued","label":"Date","values":["1998"]},{"key":"dc:publisher","label":"Institution","values":["Massachusetts Institute of Technology"]},{"key":"dc:type","label":"Dc Type","values":["Thesis"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Electrical Engineering and Computer Science"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language.iso","label":"Language (ISO)","values":["eng"]},{"key":"dc:rights","label":"Dc Rights","values":["M.I.T. theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. See provided URL for inquiries about permission."]},{"key":"dc:rights.uri","label":"Rights URI","values":["http://dspace.mit.edu/handle/1721.1/7582"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["http://hdl.handle.net/1721.1/9630"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Thesis (Ph. D.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 1998.","Includes bibliographical references (p. 113-121)."]},{"key":"dc:description.abstract","label":"Abstract","values":["Three-dimensional (3-D) modeling of urban environments has numerous applications, including virtual environments, urban planning, and physical simulation. Construct­ing 3-D models from photographs (images) is thus an important area of research in computer vision, and increasingly, computer graphics. However, despite many years of research, a system that automatically recovers realistic 3-D models remains elusive; most practical systems require significant human input. Unlike automatic algorithms, human-assisted systems are not scalable, both in terms of the number of images processed and the complexity of the generated 3-D model. This thesis describes novel techniques to automatically extract textured 3-D mod­els of urban environments from pose imagery, i.e., images annotated with camera position and orientation in a single global coordinate system. Physical instruments (e.g., surveying, Global Positioning System (GPS), inertial sensors, etc.) are used to provide accurate initial pose estimates to the proposed algorithms. As these es­timates are not perfect, I first describe two optimization techniques that refine pose estimates using information present in the images: spherical mosaicing recovers rel­ative rotations between images taken from a single position, and mosaic registration accurately locates mosaics in a global coordinate system. Next, I describe an algo­rithm that extracts vertical facades from mosaics annotated with accurate pose. The algorithm employs horizontal line segments to detect likely facade orientations and locates these facades using a space-sweep technique. Textures are robustly computed for the facades by combining information from several mosaics using median statistics. I present results for a large pose image dataset ( consisting of about four thousand images taken from eighty-one positions) of an urban office complex. These techniques were successful in recovering all significant vertical facades in the complex, as well as several neighboring facades."]},{"key":"dc:description.degree","label":"Dc Description Degree","values":["Ph.D."]},{"key":"dc:format.mimetype","label":"Dc Format Mimetype","values":["application/pdf"]},{"key":"dc:title","label":"Title","values":["Pose imagery and automated three-dimensional modeling of urban environments"]}]}],"canonical_facts":{"dc:contributor.advisor":["Seth Teller."],"dc:contributor.department":["Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science"],"dc:creator":["Coorg, Satyan R"],"dc:date.accessioned":["2005-08-19T19:03:23Z"],"dc:date.available":["2005-08-19T19:03:23Z"],"dc:date.issued":["1998"],"dc:description":["Thesis (Ph. D.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 1998.","Includes bibliographical references (p. 113-121)."],"dc:description.abstract":["Three-dimensional (3-D) modeling of urban environments has numerous applications, including virtual environments, urban planning, and physical simulation. Construct­ing 3-D models from photographs (images) is thus an important area of research in computer vision, and increasingly, computer graphics. However, despite many years of research, a system that automatically recovers realistic 3-D models remains elusive; most practical systems require significant human input. Unlike automatic algorithms, human-assisted systems are not scalable, both in terms of the number of images processed and the complexity of the generated 3-D model. This thesis describes novel techniques to automatically extract textured 3-D mod­els of urban environments from pose imagery, i.e., images annotated with camera position and orientation in a single global coordinate system. Physical instruments (e.g., surveying, Global Positioning System (GPS), inertial sensors, etc.) are used to provide accurate initial pose estimates to the proposed algorithms. As these es­timates are not perfect, I first describe two optimization techniques that refine pose estimates using information present in the images: spherical mosaicing recovers rel­ative rotations between images taken from a single position, and mosaic registration accurately locates mosaics in a global coordinate system. Next, I describe an algo­rithm that extracts vertical facades from mosaics annotated with accurate pose. The algorithm employs horizontal line segments to detect likely facade orientations and locates these facades using a space-sweep technique. Textures are robustly computed for the facades by combining information from several mosaics using median statistics. I present results for a large pose image dataset ( consisting of about four thousand images taken from eighty-one positions) of an urban office complex. These techniques were successful in recovering all significant vertical facades in the complex, as well as several neighboring facades."],"dc:description.degree":["Ph.D."],"dc:format.mimetype":["application/pdf"],"dc:identifier.uri":["http://hdl.handle.net/1721.1/9630"],"dc:language.iso":["eng"],"dc:publisher":["Massachusetts Institute of Technology"],"dc:rights":["M.I.T. theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. See provided URL for inquiries about permission."],"dc:rights.uri":["http://dspace.mit.edu/handle/1721.1/7582"],"dc:subject":["Electrical Engineering and Computer Science"],"dc:title":["Pose imagery and automated three-dimensional modeling of urban environments"],"dc:type":["Thesis"]},"updated_at":"2026-07-22T22:21:22Z"}