{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/81809"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/81809","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"High-Fidelity Image -Based Modeling","abstract":"Image-based modeling is the process of automatically acquiring geometric object and scene models from photographs or video clips. This dissertation addresses three core problems in image-based modeling: static scene reconstruction, high-fidelity camera calibration, and dynamic scene reconstruction. For static scene reconstruction, we propose two novel multi-view stereo algorithms. In the first algorithm, after building a visual hull model purely from geometric constraints associated with image silhouettes, we identify rims where the surface grazes the visual hull model, which is then carved by maximizing a photometric consistency score defined over the surface while fixing the identified rims. A local iterative deformation step is finally used to recover fine surface details. In the second algorithm, we propose a simple method that outputs a set of planar oriented rectangular patches, which are then converted into a polygonal surface. The method does not require any initialization and is capable of detecting and discarding outliers and obstacles visible in the images. It does not perform any smoothing across nearby features, yet is one of the best algorithm available today according to the recent quantitative evaluations of multi-view stereo algorithms. For high-fidelity camera calibration, given a set of camera parameters possibly containing errors, we use multi-view stereo to construct a rough geometric model of a scene, which is then used to establish feature correspondences, Standard bundle adjustment software is used with the established feature correspondences to tighten up camera parameters. The proposed method has been tested on various real data sets including objects without salient textures, where feature correspondences cannot be established without our method. Lastly, for dynamic scene reconstruction, we propose a dense 3D tracking algorithm that uses multi-view stereo in the first frame to reconstruct an initial surface mesh, then tracks its vertices over time by using a local rigid and a global non-rigid motion models. An expansion strategy, which has proven extremely effective for multi-view stereo, is employed for fast and complex motions that existing approaches cannot handle. Qualitative and quantitative experiments are performed for seven real data sets, demonstrating the effectiveness of our approach.","abstract_html":"Image-based modeling is the process of automatically acquiring geometric object and scene models from photographs or video clips. This dissertation addresses three core problems in image-based modeling: static scene reconstruction, high-fidelity camera calibration, and dynamic scene reconstruction. For static scene reconstruction, we propose two novel multi-view stereo algorithms. In the first algorithm, after building a visual hull model purely from geometric constraints associated with image silhouettes, we identify rims where the surface grazes the visual hull model, which is then carved by maximizing a photometric consistency score defined over the surface while fixing the identified rims. A local iterative deformation step is finally used to recover fine surface details. In the second algorithm, we propose a simple method that outputs a set of planar oriented rectangular patches, which are then converted into a polygonal surface. The method does not require any initialization and is capable of detecting and discarding outliers and obstacles visible in the images. It does not perform any smoothing across nearby features, yet is one of the best algorithm available today according to the recent quantitative evaluations of multi-view stereo algorithms. For high-fidelity camera calibration, given a set of camera parameters possibly containing errors, we use multi-view stereo to construct a rough geometric model of a scene, which is then used to establish feature correspondences, Standard bundle adjustment software is used with the established feature correspondences to tighten up camera parameters. The proposed method has been tested on various real data sets including objects without salient textures, where feature correspondences cannot be established without our method. Lastly, for dynamic scene reconstruction, we propose a dense 3D tracking algorithm that uses multi-view stereo in the first frame to reconstruct an initial surface mesh, then tracks its vertices over time by using a local rigid and a global non-rigid motion models. An expansion strategy, which has proven extremely effective for multi-view stereo, is employed for fast and complex motions that existing approaches cannot handle. Qualitative and quantitative experiments are performed for seven real data sets, demonstrating the effectiveness of our approach.","abstract_has_math":false,"creators":["Furukawa, Yasutaka"],"institution":"University of Illinois at Urbana-Champaign","degree_name":"Ph.D.","degree_level":"Dissertation","degree_discipline":"Computer Science","degree_department":null,"school":null,"contributors":["Ponce, Jean"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2015,"date_issued":"2015-09-25T20:20:32Z","date_published":"2015-09-25T20:20:32Z","updated_at":"2026-07-22T22:26:17Z","subjects":["Artificial Intelligence"],"languages":["eng"],"rights":[],"rights_urls":[],"identifier_entries":[{"key":"dc:identifier","label":"Identifier","values":["(MiAaPQ)AAI3314770"],"render_values":[{"text":"(MiAaPQ)AAI3314770","href":null,"code":true}]}]},"links":{"outbound_url":"http://hdl.handle.net/2142/81809","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Ponce, Jean"]},{"key":"dc:creator","label":"Author","values":["Furukawa, Yasutaka"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2015-09-25T20:20:32Z","10000-01-01","2008"]},{"key":"dc:type","label":"Dc Type","values":["text"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Computer Science"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Dissertation"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Ph.D."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois at Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Artificial Intelligence"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["eng"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["http://hdl.handle.net/2142/81809","(MiAaPQ)AAI3314770"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Image-based modeling is the process of automatically acquiring geometric object and scene models from photographs or video clips. This dissertation addresses three core problems in image-based modeling: static scene reconstruction, high-fidelity camera calibration, and dynamic scene reconstruction. For static scene reconstruction, we propose two novel multi-view stereo algorithms. In the first algorithm, after building a visual hull model purely from geometric constraints associated with image silhouettes, we identify rims where the surface grazes the visual hull model, which is then carved by maximizing a photometric consistency score defined over the surface while fixing the identified rims. A local iterative deformation step is finally used to recover fine surface details. In the second algorithm, we propose a simple method that outputs a set of planar oriented rectangular patches, which are then converted into a polygonal surface. The method does not require any initialization and is capable of detecting and discarding outliers and obstacles visible in the images. It does not perform any smoothing across nearby features, yet is one of the best algorithm available today according to the recent quantitative evaluations of multi-view stereo algorithms. For high-fidelity camera calibration, given a set of camera parameters possibly containing errors, we use multi-view stereo to construct a rough geometric model of a scene, which is then used to establish feature correspondences, Standard bundle adjustment software is used with the established feature correspondences to tighten up camera parameters. The proposed method has been tested on various real data sets including objects without salient textures, where feature correspondences cannot be established without our method. Lastly, for dynamic scene reconstruction, we propose a dense 3D tracking algorithm that uses multi-view stereo in the first frame to reconstruct an initial surface mesh, then tracks its vertices over time by using a local rigid and a global non-rigid motion models. An expansion strategy, which has proven extremely effective for multi-view stereo, is employed for fast and complex motions that existing approaches cannot handle. Qualitative and quantitative experiments are performed for seven real data sets, demonstrating the effectiveness of our approach.","Made available in DSpace on 2015-09-25T20:20:32Z (GMT). No. of bitstreams: 2 license.txt: 4848 bytes, checksum: 96035ab3f5e1c23cc7138a224ce498bd (MD5) 3314770.pdf: 3366426 bytes, checksum: 50b58e2f54fb34fdd60ff92928d6a9a0 (MD5) Previous issue date: 2008","Embargo set by: Seth Robbins for item 83090 Lift date: Forever Reason: Restricted to the U of I community idenfinitely during batch ingest of legacy ETDs","Restricted to the U of I community idenfinitely during batch ingest of legacy ETDs","U of I Only","121 p.","Thesis (Ph.D.)--University of Illinois at Urbana-Champaign, 2008."]},{"key":"dc:title","label":"Title","values":["High-Fidelity Image -Based Modeling"]}]}],"canonical_facts":{"dc:contributor":["Ponce, Jean"],"dc:creator":["Furukawa, Yasutaka"],"dc:date":["2015-09-25T20:20:32Z","10000-01-01","2008"],"dc:description":["Image-based modeling is the process of automatically acquiring geometric object and scene models from photographs or video clips. This dissertation addresses three core problems in image-based modeling: static scene reconstruction, high-fidelity camera calibration, and dynamic scene reconstruction. For static scene reconstruction, we propose two novel multi-view stereo algorithms. In the first algorithm, after building a visual hull model purely from geometric constraints associated with image silhouettes, we identify rims where the surface grazes the visual hull model, which is then carved by maximizing a photometric consistency score defined over the surface while fixing the identified rims. A local iterative deformation step is finally used to recover fine surface details. In the second algorithm, we propose a simple method that outputs a set of planar oriented rectangular patches, which are then converted into a polygonal surface. The method does not require any initialization and is capable of detecting and discarding outliers and obstacles visible in the images. It does not perform any smoothing across nearby features, yet is one of the best algorithm available today according to the recent quantitative evaluations of multi-view stereo algorithms. For high-fidelity camera calibration, given a set of camera parameters possibly containing errors, we use multi-view stereo to construct a rough geometric model of a scene, which is then used to establish feature correspondences, Standard bundle adjustment software is used with the established feature correspondences to tighten up camera parameters. The proposed method has been tested on various real data sets including objects without salient textures, where feature correspondences cannot be established without our method. Lastly, for dynamic scene reconstruction, we propose a dense 3D tracking algorithm that uses multi-view stereo in the first frame to reconstruct an initial surface mesh, then tracks its vertices over time by using a local rigid and a global non-rigid motion models. An expansion strategy, which has proven extremely effective for multi-view stereo, is employed for fast and complex motions that existing approaches cannot handle. Qualitative and quantitative experiments are performed for seven real data sets, demonstrating the effectiveness of our approach.","Made available in DSpace on 2015-09-25T20:20:32Z (GMT). No. of bitstreams: 2 license.txt: 4848 bytes, checksum: 96035ab3f5e1c23cc7138a224ce498bd (MD5) 3314770.pdf: 3366426 bytes, checksum: 50b58e2f54fb34fdd60ff92928d6a9a0 (MD5) Previous issue date: 2008","Embargo set by: Seth Robbins for item 83090 Lift date: Forever Reason: Restricted to the U of I community idenfinitely during batch ingest of legacy ETDs","Restricted to the U of I community idenfinitely during batch ingest of legacy ETDs","U of I Only","121 p.","Thesis (Ph.D.)--University of Illinois at Urbana-Champaign, 2008."],"dc:identifier":["http://hdl.handle.net/2142/81809","(MiAaPQ)AAI3314770"],"dc:language":["eng"],"dc:subject":["Artificial Intelligence"],"dc:title":["High-Fidelity Image -Based Modeling"],"dc:type":["text"],"thesis:degree_discipline":["Computer Science"],"thesis:degree_level":["Dissertation"],"thesis:degree_name":["Ph.D."],"thesis:institution_name":["University of Illinois at Urbana-Champaign"]},"updated_at":"2026-07-22T22:26:17Z"}