{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/125676"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/125676","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"Engaging people in ethical AI development: Design, dataset creation, and decision-making","abstract":"Submission published under a 24 month embargo labeled 'U of I Access', the embargo will last until 2026-08-01","abstract_html":"Submission published under a 24 month embargo labeled &#x27;U of I Access&#x27;, the embargo will last until 2026-08-01","abstract_has_math":false,"creators":["Sharma, Tanusree"],"institution":"University of Illinois at Urbana-Champaign","degree_name":"Ph.D.","degree_level":"Dissertation","degree_discipline":"Informatics","degree_department":null,"school":null,"contributors":["Wang, Yang","Huang, Yun","Miller, Andrew","Das, Sauvik"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2024,"date_issued":"2024-07-01","date_published":"2024-07-01","updated_at":"2026-07-22T22:25:02Z","subjects":["Ai Governance","Security/privacy","Ethical Ai, Democratic Ai","Dataset","Design","Decentralized Governance"],"languages":["en","eng"],"rights":["Copyright 2024 Tanusree Sharma"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"https://hdl.handle.net/2142/125676","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Wang, Yang","Huang, Yun","Miller, Andrew","Das, Sauvik"]},{"key":"dc:creator","label":"Author","values":["Sharma, Tanusree"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2024-07-01","2024-08"]},{"key":"dc:type","label":"Dc Type","values":["text","Thesis"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Informatics"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Dissertation"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Ph.D."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois at Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Ai Governance","Security/privacy","Ethical Ai, Democratic Ai","Dataset","Design","Decentralized Governance"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["en","eng"]},{"key":"dc:rights","label":"Dc Rights","values":["Copyright 2024 Tanusree Sharma"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["https://hdl.handle.net/2142/125676"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Submission published under a 24 month embargo labeled 'U of I Access', the embargo will last until 2026-08-01","The student, Tanusree Sharma, accepted the attached license on 2024-06-25 at 16:21.","The student, Tanusree Sharma, submitted this Dissertation for approval on 2024-06-25 at 16:44.","This Dissertation was approved for publication on 2024-07-01 at 10:53.","DSpace SAF Submission Ingestion Package generated from Vireo submission #20873 on 2025-02-04 at 21:16:08","Major limitations of past AI development include but are not limited to, the absence of thorough documentation and traceability in design, training with a small number of specific datasets, and lack of transparent and democratic decision-making processes for deploying AI models. These limitations could lead to adverse outcomes, such as discrimination, lack of privacy and inclusivity, and breaches of legal regulations. Underserved populations such as people with disabilities and minority groups, in particular, are disproportionately affected by these design decisions. In this dissertation, I employ qualitative, quantitative, and design methods to uncover users’ expectations for privacy and security in AI systems. I then develop methods to engage users in the dataset creation and decision-making phases of AI development. More specifically, this thesis covers different facets of AI development. I start with empirical studies on how end users think about privacy/security in application domains that are powered by AI such as targeted ads and search engines. I then dive into ethical dataset creation for AI model development. Lastly, I present how to engage a diverse set of populations toward democratic decision-making regarding how AI should behave and align with people’s values. One such privacy measure in AI systems is “Data Minimization” which means data should be adequate, relevant, and limited to what is necessary for the purposes for which it is processed. By examining user reactions to data minimization in AI systems, I find that users’ assessments of privacy and security risks can be limited by bounded rationality. My analysis of the users-centric data minimization in the AI system design phase reveals how users reason about the necessity of data based on service quality, specific types of data, or volume and recency of data influenced by their mental capacity and available information. I then investigate how the data creation phase for AI development contributes to privacy risk, misrepresentation, and bias. To address ethical concerns, I developed a method for ethically creating a dataset with an underserved population (blind users) and built the first disability-first dataset, BivPriv to support novel algorithm development for blind users’ visual privacy management (e.g., automatically identifying private or sensitive content in their pictures and videos before they share them with others). The findings highlight users’ willingness to actively engage in the AI development lifecycle and their expectations to be informed about how researchers or companies use the data they share and how their contribution benefits AI systems. Leveraging these insights, I then designed Inclusive.AI, a platform equipped with decentralized governance mechanisms (e.g., quadratic voting), aiming to engage a wide range of people (e.g., underserved groups) in democratic decision-making processes about AI development and governance. This involves introducing technical interventions, particularly Decentralized Autonomous Organizations (DAOs), that incorporate various deliberation, and preference aggregation through voting mechanisms to allow active user participation in influencing how AI should behave, for instance, determining the level of personalization users would experience in AI systems. Through large-scale experiments on different use cases (e.g., text-to-image models, multimodal language models), Inclusive.AI demonstrates promising improvements in scalability, expressiveness, and governance quality in incorporating users’ preferences. This thesis highlights the importance of critically reflecting on when and how to engage end users in the AI development lifecycle. While engaging users is crucial, it is important to build a method for fair deliberation that considers the resilience of the decision-making process, the power structure among stakeholders, and the aggregation method used for gathering user preferences with practical constraints. Key contributions of this thesis include- (i) privacy and security expectations in AI systems (e.g., search engines, targeted ads) across South Asia and EU/UK, accounting for cultural, socioeconomical, and regulatory differences to conceptualize low-fidelity designs based on contextual and situational privacy concerns; (ii) a novel method to ethically create public disability-first dataset, “BivPriv,” to support AI models development in identifying private visual content; (iii) systematically assess the level of decentralization of various DAOs to identify design metrics for democratic decision-making platform; and (iv) design and evaluation of a democratic tool “InclusiveAI,” for decision-making on controversial AI topic (e.g. stereotype, politics), guiding future research in “Democratic AI.”"]},{"key":"dc:format","label":"Dc Format","values":["application/pdf"]},{"key":"dc:title","label":"Title","values":["Engaging people in ethical AI development: Design, dataset creation, and decision-making"]}]}],"canonical_facts":{"dc:contributor":["Wang, Yang","Huang, Yun","Miller, Andrew","Das, Sauvik"],"dc:creator":["Sharma, Tanusree"],"dc:date":["2024-07-01","2024-08"],"dc:description":["Submission published under a 24 month embargo labeled 'U of I Access', the embargo will last until 2026-08-01","The student, Tanusree Sharma, accepted the attached license on 2024-06-25 at 16:21.","The student, Tanusree Sharma, submitted this Dissertation for approval on 2024-06-25 at 16:44.","This Dissertation was approved for publication on 2024-07-01 at 10:53.","DSpace SAF Submission Ingestion Package generated from Vireo submission #20873 on 2025-02-04 at 21:16:08","Major limitations of past AI development include but are not limited to, the absence of thorough documentation and traceability in design, training with a small number of specific datasets, and lack of transparent and democratic decision-making processes for deploying AI models. These limitations could lead to adverse outcomes, such as discrimination, lack of privacy and inclusivity, and breaches of legal regulations. Underserved populations such as people with disabilities and minority groups, in particular, are disproportionately affected by these design decisions. In this dissertation, I employ qualitative, quantitative, and design methods to uncover users’ expectations for privacy and security in AI systems. I then develop methods to engage users in the dataset creation and decision-making phases of AI development. More specifically, this thesis covers different facets of AI development. I start with empirical studies on how end users think about privacy/security in application domains that are powered by AI such as targeted ads and search engines. I then dive into ethical dataset creation for AI model development. Lastly, I present how to engage a diverse set of populations toward democratic decision-making regarding how AI should behave and align with people’s values. One such privacy measure in AI systems is “Data Minimization” which means data should be adequate, relevant, and limited to what is necessary for the purposes for which it is processed. By examining user reactions to data minimization in AI systems, I find that users’ assessments of privacy and security risks can be limited by bounded rationality. My analysis of the users-centric data minimization in the AI system design phase reveals how users reason about the necessity of data based on service quality, specific types of data, or volume and recency of data influenced by their mental capacity and available information. I then investigate how the data creation phase for AI development contributes to privacy risk, misrepresentation, and bias. To address ethical concerns, I developed a method for ethically creating a dataset with an underserved population (blind users) and built the first disability-first dataset, BivPriv to support novel algorithm development for blind users’ visual privacy management (e.g., automatically identifying private or sensitive content in their pictures and videos before they share them with others). The findings highlight users’ willingness to actively engage in the AI development lifecycle and their expectations to be informed about how researchers or companies use the data they share and how their contribution benefits AI systems. Leveraging these insights, I then designed Inclusive.AI, a platform equipped with decentralized governance mechanisms (e.g., quadratic voting), aiming to engage a wide range of people (e.g., underserved groups) in democratic decision-making processes about AI development and governance. This involves introducing technical interventions, particularly Decentralized Autonomous Organizations (DAOs), that incorporate various deliberation, and preference aggregation through voting mechanisms to allow active user participation in influencing how AI should behave, for instance, determining the level of personalization users would experience in AI systems. Through large-scale experiments on different use cases (e.g., text-to-image models, multimodal language models), Inclusive.AI demonstrates promising improvements in scalability, expressiveness, and governance quality in incorporating users’ preferences. This thesis highlights the importance of critically reflecting on when and how to engage end users in the AI development lifecycle. While engaging users is crucial, it is important to build a method for fair deliberation that considers the resilience of the decision-making process, the power structure among stakeholders, and the aggregation method used for gathering user preferences with practical constraints. Key contributions of this thesis include- (i) privacy and security expectations in AI systems (e.g., search engines, targeted ads) across South Asia and EU/UK, accounting for cultural, socioeconomical, and regulatory differences to conceptualize low-fidelity designs based on contextual and situational privacy concerns; (ii) a novel method to ethically create public disability-first dataset, “BivPriv,” to support AI models development in identifying private visual content; (iii) systematically assess the level of decentralization of various DAOs to identify design metrics for democratic decision-making platform; and (iv) design and evaluation of a democratic tool “InclusiveAI,” for decision-making on controversial AI topic (e.g. stereotype, politics), guiding future research in “Democratic AI.”"],"dc:format":["application/pdf"],"dc:identifier":["https://hdl.handle.net/2142/125676"],"dc:language":["en","eng"],"dc:rights":["Copyright 2024 Tanusree Sharma"],"dc:subject":["Ai Governance","Security/privacy","Ethical Ai, Democratic Ai","Dataset","Design","Decentralized Governance"],"dc:title":["Engaging people in ethical AI development: Design, dataset creation, and decision-making"],"dc:type":["text","Thesis"],"thesis:degree_discipline":["Informatics"],"thesis:degree_level":["Dissertation"],"thesis:degree_name":["Ph.D."],"thesis:institution_name":["University of Illinois at Urbana-Champaign"]},"updated_at":"2026-07-22T22:25:02Z"}