{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/81619"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/81619","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"Visual Objects and Environments: Capture, Extraction, and Representation","abstract":"The availability of affordable computational power and graphics rendering capabilities is enabling the creation of realistic imagery that are widely used today in special effects and animation. New forms of visual media with a higher degree of interactivity, such as video games and virtual reality (VR), have also emerged as a result of these technological advances. While there are many factors that contribute to the presentation effectiveness of a VR simulation or video game, one of the most important is visual realism, and the use of images as textures is often the key. In this dissertation we introduce a number of tools and techniques that are aimed at making it easier to create photorealistic virtual objects and environments. In the first part we describe an algorithm for extracting object boundaries from images, formulated as a probabilistic alpha channel estimation problem. The closed-form solution proposed enables detailed and possibly diffused object boundaries to be found, so that visual objects embedded in digital images can be extracted. We also describe an interactive tool that allows this object extraction operation to be performed with loosely drawn freehand sketches. The second part of the thesis introduces an image-based representation for visual objects called facted appearance models that can be used for object recognition, and pose estimation. The model was used successfully in object tracking in video streams, as well as to estimate eye gaze by treating it as a pose estimation problem. The gaze estimation algorithm achieved 0.36 degree accuracy, which is to existing, more complicated eye gaze estimation techniques. The final part of the thesis describes a panoramic camera based on mirror pyramids that is capable of capturing visual environments from multiple view points simultaneously at video rates.","abstract_html":"The availability of affordable computational power and graphics rendering capabilities is enabling the creation of realistic imagery that are widely used today in special effects and animation. New forms of visual media with a higher degree of interactivity, such as video games and virtual reality (VR), have also emerged as a result of these technological advances. While there are many factors that contribute to the presentation effectiveness of a VR simulation or video game, one of the most important is visual realism, and the use of images as textures is often the key. In this dissertation we introduce a number of tools and techniques that are aimed at making it easier to create photorealistic virtual objects and environments. In the first part we describe an algorithm for extracting object boundaries from images, formulated as a probabilistic alpha channel estimation problem. The closed-form solution proposed enables detailed and possibly diffused object boundaries to be found, so that visual objects embedded in digital images can be extracted. We also describe an interactive tool that allows this object extraction operation to be performed with loosely drawn freehand sketches. The second part of the thesis introduces an image-based representation for visual objects called facted appearance models that can be used for object recognition, and pose estimation. The model was used successfully in object tracking in video streams, as well as to estimate eye gaze by treating it as a pose estimation problem. The gaze estimation algorithm achieved 0.36 degree accuracy, which is to existing, more complicated eye gaze estimation techniques. The final part of the thesis describes a panoramic camera based on mirror pyramids that is capable of capturing visual environments from multiple view points simultaneously at video rates.","abstract_has_math":false,"creators":["Tan, Kar-Han"],"institution":"University of Illinois at Urbana-Champaign","degree_name":"Ph.D.","degree_level":"Dissertation","degree_discipline":"Computer Science","degree_department":null,"school":null,"contributors":["Ahuja, Narendra"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2015,"date_issued":"2015-09-25T20:19:33Z","date_published":"2015-09-25T20:19:33Z","updated_at":"2026-07-22T22:26:16Z","subjects":["Computer Science"],"languages":["eng"],"rights":[],"rights_urls":[],"identifier_entries":[{"key":"dc:identifier","label":"Identifier","values":["(MiAaPQ)AAI3086196"],"render_values":[{"text":"(MiAaPQ)AAI3086196","href":null,"code":true}]}]},"links":{"outbound_url":"http://hdl.handle.net/2142/81619","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Ahuja, Narendra"]},{"key":"dc:creator","label":"Author","values":["Tan, Kar-Han"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2015-09-25T20:19:33Z","10000-01-01","2003"]},{"key":"dc:type","label":"Dc Type","values":["text"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Computer Science"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Dissertation"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Ph.D."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois at Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Computer Science"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["eng"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["http://hdl.handle.net/2142/81619","(MiAaPQ)AAI3086196"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["The availability of affordable computational power and graphics rendering capabilities is enabling the creation of realistic imagery that are widely used today in special effects and animation. New forms of visual media with a higher degree of interactivity, such as video games and virtual reality (VR), have also emerged as a result of these technological advances. While there are many factors that contribute to the presentation effectiveness of a VR simulation or video game, one of the most important is visual realism, and the use of images as textures is often the key. In this dissertation we introduce a number of tools and techniques that are aimed at making it easier to create photorealistic virtual objects and environments. In the first part we describe an algorithm for extracting object boundaries from images, formulated as a probabilistic alpha channel estimation problem. The closed-form solution proposed enables detailed and possibly diffused object boundaries to be found, so that visual objects embedded in digital images can be extracted. We also describe an interactive tool that allows this object extraction operation to be performed with loosely drawn freehand sketches. The second part of the thesis introduces an image-based representation for visual objects called facted appearance models that can be used for object recognition, and pose estimation. The model was used successfully in object tracking in video streams, as well as to estimate eye gaze by treating it as a pose estimation problem. The gaze estimation algorithm achieved 0.36 degree accuracy, which is to existing, more complicated eye gaze estimation techniques. The final part of the thesis describes a panoramic camera based on mirror pyramids that is capable of capturing visual environments from multiple view points simultaneously at video rates.","Made available in DSpace on 2015-09-25T20:19:33Z (GMT). No. of bitstreams: 2 license.txt: 4848 bytes, checksum: 96035ab3f5e1c23cc7138a224ce498bd (MD5) 3086196.pdf: 4427786 bytes, checksum: aad75ad59f38519c052ea932ab2a0940 (MD5) Previous issue date: 2003","Embargo set by: Seth Robbins for item 82900 Lift date: Forever Reason: Restricted to the U of I community idenfinitely during batch ingest of legacy ETDs","Restricted to the U of I community idenfinitely during batch ingest of legacy ETDs","U of I Only","110 p.","Thesis (Ph.D.)--University of Illinois at Urbana-Champaign, 2003."]},{"key":"dc:title","label":"Title","values":["Visual Objects and Environments: Capture, Extraction, and Representation"]}]}],"canonical_facts":{"dc:contributor":["Ahuja, Narendra"],"dc:creator":["Tan, Kar-Han"],"dc:date":["2015-09-25T20:19:33Z","10000-01-01","2003"],"dc:description":["The availability of affordable computational power and graphics rendering capabilities is enabling the creation of realistic imagery that are widely used today in special effects and animation. New forms of visual media with a higher degree of interactivity, such as video games and virtual reality (VR), have also emerged as a result of these technological advances. While there are many factors that contribute to the presentation effectiveness of a VR simulation or video game, one of the most important is visual realism, and the use of images as textures is often the key. In this dissertation we introduce a number of tools and techniques that are aimed at making it easier to create photorealistic virtual objects and environments. In the first part we describe an algorithm for extracting object boundaries from images, formulated as a probabilistic alpha channel estimation problem. The closed-form solution proposed enables detailed and possibly diffused object boundaries to be found, so that visual objects embedded in digital images can be extracted. We also describe an interactive tool that allows this object extraction operation to be performed with loosely drawn freehand sketches. The second part of the thesis introduces an image-based representation for visual objects called facted appearance models that can be used for object recognition, and pose estimation. The model was used successfully in object tracking in video streams, as well as to estimate eye gaze by treating it as a pose estimation problem. The gaze estimation algorithm achieved 0.36 degree accuracy, which is to existing, more complicated eye gaze estimation techniques. The final part of the thesis describes a panoramic camera based on mirror pyramids that is capable of capturing visual environments from multiple view points simultaneously at video rates.","Made available in DSpace on 2015-09-25T20:19:33Z (GMT). No. of bitstreams: 2 license.txt: 4848 bytes, checksum: 96035ab3f5e1c23cc7138a224ce498bd (MD5) 3086196.pdf: 4427786 bytes, checksum: aad75ad59f38519c052ea932ab2a0940 (MD5) Previous issue date: 2003","Embargo set by: Seth Robbins for item 82900 Lift date: Forever Reason: Restricted to the U of I community idenfinitely during batch ingest of legacy ETDs","Restricted to the U of I community idenfinitely during batch ingest of legacy ETDs","U of I Only","110 p.","Thesis (Ph.D.)--University of Illinois at Urbana-Champaign, 2003."],"dc:identifier":["http://hdl.handle.net/2142/81619","(MiAaPQ)AAI3086196"],"dc:language":["eng"],"dc:subject":["Computer Science"],"dc:title":["Visual Objects and Environments: Capture, Extraction, and Representation"],"dc:type":["text"],"thesis:degree_discipline":["Computer Science"],"thesis:degree_level":["Dissertation"],"thesis:degree_name":["Ph.D."],"thesis:institution_name":["University of Illinois at Urbana-Champaign"]},"updated_at":"2026-07-22T22:26:16Z"}