{"id":{"repo_id":"houston","oai_identifier":"oai:uh-ir.tdl.org:10657/4417"},"canonical_url":"https://search.dev.ndltd.org/etd/houston/oai:uh-ir.tdl.org:10657/4417","repository":{"repo_id":"houston","name":"University of Houston","base_url":"https://uh-ir.tdl.org/server/oai/request"},"display":{"title":"Face Recognition in Unconstrained Conditions: Improving Face Alignment and Constructing a Pose-Invariant Compact Biometric Template","abstract":"Face recognition has been significantly advanced in the past decade; however, challenges remain under unconstrained conditions regarding variations in pose, illumination, and occlusion. Existing solutions tackle the unconstrained face recognition problem in two ways: (i) controlling the variations of the input to the recognition system, and (ii) improving the robustness of the recognition system to these variations. In the first method, the face frontalization module in 3D-aided face recognition significantly reduces the pose variation by mapping a facial image to a frontalized texture space with the help of a 3D facial model. However, because face frontalization relies heavily on the projection matrix generated by face alignment, its performance has been largely constrained by the robustness of face alignment under unconstrained conditions. In the second method, using an ensemble deep neural network model for recognition has been demonstrated to be robust to pose variations. However, the biometric template generated by the ensemble model is much larger than the template generated by an individual model. This dissertation presents solutions to both problems. To improve the robustness of face alignment under unconstrained conditions and significantly reduce the biometric template size, the first contribution is a Globally Optimized Dual-Pathway (GoDP) landmark detector algorithm that is robust to head pose variations up to 90\\degrees. The second contribution is a pose estimation algorithm namely Annotated Face Model-based Alignment (AFMA) that estimates a head pose without landmarks. The third contribution is a pose estimation algorithm with the name Sensible-Points based reinforced Hypothesis Refinement (SHR) which is robust to facial occlusion. The fourth contribution is a pose estimation algorithm with the name Convolutional Point-set Representation-based Face Alignment (CPRFA), it is robust to facial occlusion and large head pose variations. The fifth contribution is a neural network architecture that reduces the template size of an ensemble deep model by more than an order-of-magnitude based on self-occlusion masks, we name it Mask-Guided Compact Template Learning (MGCTL). When plugging GoDP and MGCTL into a 3D-aided face recognition pipeline, state-of-the-art performance is achieved on multiple databases in terms of both face recognition accuracy and template matching speed.","abstract_html":"Face recognition has been significantly advanced in the past decade; however, challenges remain under unconstrained conditions regarding variations in pose, illumination, and occlusion. Existing solutions tackle the unconstrained face recognition problem in two ways: (i) controlling the variations of the input to the recognition system, and (ii) improving the robustness of the recognition system to these variations. In the first method, the face frontalization module in 3D-aided face recognition significantly reduces the pose variation by mapping a facial image to a frontalized texture space with the help of a 3D facial model. However, because face frontalization relies heavily on the projection matrix generated by face alignment, its performance has been largely constrained by the robustness of face alignment under unconstrained conditions. In the second method, using an ensemble deep neural network model for recognition has been demonstrated to be robust to pose variations. However, the biometric template generated by the ensemble model is much larger than the template generated by an individual model. This dissertation presents solutions to both problems. To improve the robustness of face alignment under unconstrained conditions and significantly reduce the biometric template size, the first contribution is a Globally Optimized Dual-Pathway (GoDP) landmark detector algorithm that is robust to head pose variations up to 90\\degrees. The second contribution is a pose estimation algorithm namely Annotated Face Model-based Alignment (AFMA) that estimates a head pose without landmarks. The third contribution is a pose estimation algorithm with the name Sensible-Points based reinforced Hypothesis Refinement (SHR) which is robust to facial occlusion. The fourth contribution is a pose estimation algorithm with the name Convolutional Point-set Representation-based Face Alignment (CPRFA), it is robust to facial occlusion and large head pose variations. The fifth contribution is a neural network architecture that reduces the template size of an ensemble deep model by more than an order-of-magnitude based on self-occlusion masks, we name it Mask-Guided Compact Template Learning (MGCTL). When plugging GoDP and MGCTL into a 3D-aided face recognition pipeline, state-of-the-art performance is achieved on multiple databases in terms of both face recognition accuracy and template matching speed.","abstract_has_math":false,"creators":["Wu, Yuhang 1991-"],"institution":"University of Houston","degree_name":"Doctor of Philosophy","degree_level":"Doctoral","degree_discipline":"Computer Science","degree_department":null,"school":null,"contributors":[],"advisors":["Kakadiaris, Ioannis A."],"committee_chairs":[],"committee_members":["Shah, Shishir Kirit","Vilalta, Ricardo","Prasad, Saurabh"],"year":2018,"date_issued":"2018-12","date_published":"2018-12","updated_at":"2026-07-24T02:32:27Z","subjects":["Face recognition","Face alignment","Compact template learning","Deep learning"],"languages":["eng"],"rights":["The author of this work is the copyright owner. UH Libraries and the Texas Digital Library have their permission to store and provide access to this work. Further transmission, reproduction, or presentation of this work is prohibited except with permission of the author(s)."],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/10657/4417","outbound_label":"Handle","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor.advisor","label":"Advisor","values":["Kakadiaris, Ioannis A."]},{"key":"dc:contributor.committeemember","label":"Committee Member","values":["Shah, Shishir Kirit","Vilalta, Ricardo","Prasad, Saurabh"]},{"key":"dc:creator","label":"Author","values":["Wu, Yuhang 1991-"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2019-09-10T15:18:11Z"]},{"key":"dc:date.issued","label":"Date","values":["2018-12"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Computer Science"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Doctoral"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Doctor of Philosophy"]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Houston"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Face recognition","Face alignment","Compact template learning","Deep learning"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language.iso","label":"Language (ISO)","values":["eng"]},{"key":"dc:rights","label":"Dc Rights","values":["The author of this work is the copyright owner. UH Libraries and the Texas Digital Library have their permission to store and provide access to this work. Further transmission, reproduction, or presentation of this work is prohibited except with permission of the author(s)."]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["https://hdl.handle.net/10657/4417"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["Face recognition has been significantly advanced in the past decade; however, challenges remain under unconstrained conditions regarding variations in pose, illumination, and occlusion. Existing solutions tackle the unconstrained face recognition problem in two ways: (i) controlling the variations of the input to the recognition system, and (ii) improving the robustness of the recognition system to these variations. In the first method, the face frontalization module in 3D-aided face recognition significantly reduces the pose variation by mapping a facial image to a frontalized texture space with the help of a 3D facial model. However, because face frontalization relies heavily on the projection matrix generated by face alignment, its performance has been largely constrained by the robustness of face alignment under unconstrained conditions. In the second method, using an ensemble deep neural network model for recognition has been demonstrated to be robust to pose variations. However, the biometric template generated by the ensemble model is much larger than the template generated by an individual model. This dissertation presents solutions to both problems. To improve the robustness of face alignment under unconstrained conditions and significantly reduce the biometric template size, the first contribution is a Globally Optimized Dual-Pathway (GoDP) landmark detector algorithm that is robust to head pose variations up to 90\\degrees. The second contribution is a pose estimation algorithm namely Annotated Face Model-based Alignment (AFMA) that estimates a head pose without landmarks. The third contribution is a pose estimation algorithm with the name Sensible-Points based reinforced Hypothesis Refinement (SHR) which is robust to facial occlusion. The fourth contribution is a pose estimation algorithm with the name Convolutional Point-set Representation-based Face Alignment (CPRFA), it is robust to facial occlusion and large head pose variations. The fifth contribution is a neural network architecture that reduces the template size of an ensemble deep model by more than an order-of-magnitude based on self-occlusion masks, we name it Mask-Guided Compact Template Learning (MGCTL). When plugging GoDP and MGCTL into a 3D-aided face recognition pipeline, state-of-the-art performance is achieved on multiple databases in terms of both face recognition accuracy and template matching speed."]},{"key":"dc:format.mimetype","label":"Dc Format Mimetype","values":["application/pdf"]},{"key":"dc:title","label":"Title","values":["Face Recognition in Unconstrained Conditions: Improving Face Alignment and Constructing a Pose-Invariant Compact Biometric Template"]}]}],"canonical_facts":{"dc:contributor.advisor":["Kakadiaris, Ioannis A."],"dc:contributor.committeemember":["Shah, Shishir Kirit","Vilalta, Ricardo","Prasad, Saurabh"],"dc:creator":["Wu, Yuhang 1991-"],"dc:date.accessioned":["2019-09-10T15:18:11Z"],"dc:date.issued":["2018-12"],"dc:description.abstract":["Face recognition has been significantly advanced in the past decade; however, challenges remain under unconstrained conditions regarding variations in pose, illumination, and occlusion. Existing solutions tackle the unconstrained face recognition problem in two ways: (i) controlling the variations of the input to the recognition system, and (ii) improving the robustness of the recognition system to these variations. In the first method, the face frontalization module in 3D-aided face recognition significantly reduces the pose variation by mapping a facial image to a frontalized texture space with the help of a 3D facial model. However, because face frontalization relies heavily on the projection matrix generated by face alignment, its performance has been largely constrained by the robustness of face alignment under unconstrained conditions. In the second method, using an ensemble deep neural network model for recognition has been demonstrated to be robust to pose variations. However, the biometric template generated by the ensemble model is much larger than the template generated by an individual model. This dissertation presents solutions to both problems. To improve the robustness of face alignment under unconstrained conditions and significantly reduce the biometric template size, the first contribution is a Globally Optimized Dual-Pathway (GoDP) landmark detector algorithm that is robust to head pose variations up to 90\\degrees. The second contribution is a pose estimation algorithm namely Annotated Face Model-based Alignment (AFMA) that estimates a head pose without landmarks. The third contribution is a pose estimation algorithm with the name Sensible-Points based reinforced Hypothesis Refinement (SHR) which is robust to facial occlusion. The fourth contribution is a pose estimation algorithm with the name Convolutional Point-set Representation-based Face Alignment (CPRFA), it is robust to facial occlusion and large head pose variations. The fifth contribution is a neural network architecture that reduces the template size of an ensemble deep model by more than an order-of-magnitude based on self-occlusion masks, we name it Mask-Guided Compact Template Learning (MGCTL). When plugging GoDP and MGCTL into a 3D-aided face recognition pipeline, state-of-the-art performance is achieved on multiple databases in terms of both face recognition accuracy and template matching speed."],"dc:format.mimetype":["application/pdf"],"dc:identifier.uri":["https://hdl.handle.net/10657/4417"],"dc:language.iso":["eng"],"dc:rights":["The author of this work is the copyright owner. UH Libraries and the Texas Digital Library have their permission to store and provide access to this work. Further transmission, reproduction, or presentation of this work is prohibited except with permission of the author(s)."],"dc:subject":["Face recognition","Face alignment","Compact template learning","Deep learning"],"dc:title":["Face Recognition in Unconstrained Conditions: Improving Face Alignment and Constructing a Pose-Invariant Compact Biometric Template"],"thesis:degree_discipline":["Computer Science"],"thesis:degree_level":["Doctoral"],"thesis:degree_name":["Doctor of Philosophy"],"thesis:institution_name":["University of Houston"]},"updated_at":"2026-07-24T02:32:27Z"}