{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/129245"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/129245","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"Towards efficient and powerful machine learning vision systems for mobile: unifying interactive segmentation and matting","abstract":"Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2025-10-19 without embargo terms","abstract_html":"Submission original under an indefinite embargo labeled &#x27;Open Access&#x27;. The submission was exported from vireo on 2025-10-19 without embargo terms","abstract_has_math":false,"creators":["Wu, Mingyuan"],"institution":"University of Illinois Urbana-Champaign","degree_name":"M.S.","degree_level":"Thesis","degree_discipline":"Computer Science","degree_department":null,"school":null,"contributors":["Nahrstedt, Klara"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2025,"date_issued":"2025-04-28","date_published":"2025-04-28","updated_at":"2026-07-22T22:25:04Z","subjects":["Machine Learning","Segmentation","Matting"],"languages":["en","eng"],"rights":["Copyright 2025 Mingyuan Wu"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/2142/129245","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Nahrstedt, Klara"]},{"key":"dc:creator","label":"Author","values":["Wu, Mingyuan"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2025-04-28","2025-05"]},{"key":"dc:type","label":"Dc Type","values":["text","Thesis"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Computer Science"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Thesis"]},{"key":"thesis:degree_name","label":"Degree Name","values":["M.S."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Machine Learning","Segmentation","Matting"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["en","eng"]},{"key":"dc:rights","label":"Dc Rights","values":["Copyright 2025 Mingyuan Wu"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["https://hdl.handle.net/2142/129245"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2025-10-19 without embargo terms","The student, Mingyuan Wu, accepted the attached license on 2025-04-24 at 13:08.","The student, Mingyuan Wu, submitted this Thesis for approval on 2025-04-24 at 13:23.","This Thesis was approved for publication on 2025-04-28 at 12:49.","DSpace SAF Submission Ingestion Package generated from Vireo submission #21955 on 2025-10-19 at 18:10:53","Recent advancements in hardware have significantly enhanced the capabilities of smartphones, tablets and head-mounted devices, transforming these mobile devices from mere communication tools into powerful multimedia and creative platforms. This transformation has created unprecedented demand for advanced image editing functionalities running behind mobile applications. Today, users no longer like waiting until they return home to their desktop computers and launch software like Photoshop for sophisticated image manipulation. Instead, they expect equally powerful editing features that allow immediate modifications at the moment they capture or discover content worth sharing on their mobile devices. These features empower users across various contexts, from social media creators to professionals, profoundly influencing the way visual content is produced, consumed, and shared, and more importantly, the way how people interact, connect, and share their interests and cherished moments with those they love. These image editing functionalities are largely backed with recent innovations in efficient deep learning in computer vision areas. However, deploying heavy machine learning algorithms efficiently on mobile platforms remains challenging due to inherent constraints such as limited computation power, battery capacity, and the need for real-time respond. These resource limitations directly affect user experience, making it essential to develop highly optimized image editing algorithms specifically for the mobile environment. These algorithms must deliver high accuracy, high computational efficiency, low latency, real-time user interactivity, and robust generalization across diverse visual contents. In this thesis, we focus on segmentation and matting, which serve as foundational editing techniques for background replacement, portrait, and precise foreground object extraction, etc. We develop efficient neural network based algorithms designed for the resource-constrained mobile environment. Moreover, we tailor our editing functionalities to practical user scenarios, adding interactivity in segmentation tasks and reducing input requirements for matting processes, for better user experience and compatibility across wider range of mobile devices. Specifically, TraceNet ackles efficient and interactive instance segmentation by explicitly locating the user-selected instance via receptive field tracing, and thus significantly reduce computations in neural networks. Meanwhile, I-Matting addresses trimap-free image matting via a hierarchical adversarial training mechanism and a patch-rank component, and improves the matting accuracy while reducing computations. Collectively, these methods constitute a meaningful step towards efficient mobile-based image editing, demonstrating their effectiveness through accuracy and efficiency improvements in extensive datasets."]},{"key":"dc:format","label":"Dc Format","values":["application/pdf"]},{"key":"dc:title","label":"Title","values":["Towards efficient and powerful machine learning vision systems for mobile: unifying interactive segmentation and matting"]}]}],"canonical_facts":{"dc:contributor":["Nahrstedt, Klara"],"dc:creator":["Wu, Mingyuan"],"dc:date":["2025-04-28","2025-05"],"dc:description":["Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2025-10-19 without embargo terms","The student, Mingyuan Wu, accepted the attached license on 2025-04-24 at 13:08.","The student, Mingyuan Wu, submitted this Thesis for approval on 2025-04-24 at 13:23.","This Thesis was approved for publication on 2025-04-28 at 12:49.","DSpace SAF Submission Ingestion Package generated from Vireo submission #21955 on 2025-10-19 at 18:10:53","Recent advancements in hardware have significantly enhanced the capabilities of smartphones, tablets and head-mounted devices, transforming these mobile devices from mere communication tools into powerful multimedia and creative platforms. This transformation has created unprecedented demand for advanced image editing functionalities running behind mobile applications. Today, users no longer like waiting until they return home to their desktop computers and launch software like Photoshop for sophisticated image manipulation. Instead, they expect equally powerful editing features that allow immediate modifications at the moment they capture or discover content worth sharing on their mobile devices. These features empower users across various contexts, from social media creators to professionals, profoundly influencing the way visual content is produced, consumed, and shared, and more importantly, the way how people interact, connect, and share their interests and cherished moments with those they love. These image editing functionalities are largely backed with recent innovations in efficient deep learning in computer vision areas. However, deploying heavy machine learning algorithms efficiently on mobile platforms remains challenging due to inherent constraints such as limited computation power, battery capacity, and the need for real-time respond. These resource limitations directly affect user experience, making it essential to develop highly optimized image editing algorithms specifically for the mobile environment. These algorithms must deliver high accuracy, high computational efficiency, low latency, real-time user interactivity, and robust generalization across diverse visual contents. In this thesis, we focus on segmentation and matting, which serve as foundational editing techniques for background replacement, portrait, and precise foreground object extraction, etc. We develop efficient neural network based algorithms designed for the resource-constrained mobile environment. Moreover, we tailor our editing functionalities to practical user scenarios, adding interactivity in segmentation tasks and reducing input requirements for matting processes, for better user experience and compatibility across wider range of mobile devices. Specifically, TraceNet ackles efficient and interactive instance segmentation by explicitly locating the user-selected instance via receptive field tracing, and thus significantly reduce computations in neural networks. Meanwhile, I-Matting addresses trimap-free image matting via a hierarchical adversarial training mechanism and a patch-rank component, and improves the matting accuracy while reducing computations. Collectively, these methods constitute a meaningful step towards efficient mobile-based image editing, demonstrating their effectiveness through accuracy and efficiency improvements in extensive datasets."],"dc:format":["application/pdf"],"dc:identifier":["https://hdl.handle.net/2142/129245"],"dc:language":["en","eng"],"dc:rights":["Copyright 2025 Mingyuan Wu"],"dc:subject":["Machine Learning","Segmentation","Matting"],"dc:title":["Towards efficient and powerful machine learning vision systems for mobile: unifying interactive segmentation and matting"],"dc:type":["text","Thesis"],"thesis:degree_discipline":["Computer Science"],"thesis:degree_level":["Thesis"],"thesis:degree_name":["M.S."],"thesis:institution_name":["University of Illinois Urbana-Champaign"]},"updated_at":"2026-07-22T22:25:04Z"}