{"id":{"repo_id":"nus","oai_identifier":"oai:scholarbank.nus.edu.sg:10635/244765"},"canonical_url":"https://search.dev.ndltd.org/etd/nus/oai:scholarbank.nus.edu.sg:10635/244765","repository":{"repo_id":"nus","name":"National University of Singapore","base_url":"https://scholarbank.nus.edu.sg/oai/request"},"display":{"title":"TOWARDS HUMAN-CENTRIC AI: INVERSE REINFORCEMENT LEARNING MEETS ALGORITHMIC FAIRNESS","abstract":"This thesis focuses on two aspects of human-centric AI - value alignment and fairness. Our first work explores Inverse reinforcement learning (IRL), a potential solution to value alignment. We introduce BO-IRL, an IRL algorithm that uses a novel kernel to explore the reward function space efficiently. The second work introduces SCALES, a framework that translates various fairness principles into fair decisions by translating them to a combination of utility, non-causal and causal components which are, in turn, mapped to elements of a Constrained Markov Decision Process. Our final work unifies the concepts of value alignment and fairness by extending the IRL problem to scenarios where the expert agent is fairness abiding. We propose FAIR-BOIRL, a new BO-based IRL algorithm that searches for solutions across both the reward function space and fairness principles. FAIR-BOIRL uses IM-GPTS, a novel acquisition function that uses an implicit multi-arm bandit strategy, to perform an efficient search.","abstract_html":"This thesis focuses on two aspects of human-centric AI - value alignment and fairness. Our first work explores Inverse reinforcement learning (IRL), a potential solution to value alignment. We introduce BO-IRL, an IRL algorithm that uses a novel kernel to explore the reward function space efficiently. The second work introduces SCALES, a framework that translates various fairness principles into fair decisions by translating them to a combination of utility, non-causal and causal components which are, in turn, mapped to elements of a Constrained Markov Decision Process. Our final work unifies the concepts of value alignment and fairness by extending the IRL problem to scenarios where the expert agent is fairness abiding. We propose FAIR-BOIRL, a new BO-based IRL algorithm that searches for solutions across both the reward function space and fairness principles. FAIR-BOIRL uses IM-GPTS, a novel acquisition function that uses an implicit multi-arm bandit strategy, to perform an efficient search.","abstract_has_math":false,"creators":["SREEJITH BALAKRISHNAN"],"institution":null,"degree_name":null,"degree_level":null,"degree_discipline":null,"degree_department":null,"school":null,"contributors":[],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2023,"date_issued":"2023-02-22","date_published":"2023-02-22","updated_at":"2026-07-24T03:31:13Z","subjects":["Human-centric AI, Value Alignment, Algorithmic Fairness, Inverse Reinforcement Learning, Bayesian optimization, Reinforcement Learning"],"languages":[],"rights":[],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":null,"outbound_label":null,"outbound_source":null},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:creator","label":"Author","values":["SREEJITH BALAKRISHNAN"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.issued","label":"Date","values":["2023-02-22"]},{"key":"dc:relation.isreferencedby","label":"Dc Relation Isreferencedby","values":["https://scholarbank.nus.edu.sg/handle/10635/244765"]},{"key":"dc:type","label":"Dc Type","values":["Thesis"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Human-centric AI, Value Alignment, Algorithmic Fairness, Inverse Reinforcement Learning, Bayesian optimization, Reinforcement Learning"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["https://scholarbank.nus.edu.sg/bitstreams/22425ef1-0070-4b89-9940-53b27ad307a8/download"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["This thesis focuses on two aspects of human-centric AI - value alignment and fairness. Our first work explores Inverse reinforcement learning (IRL), a potential solution to value alignment. We introduce BO-IRL, an IRL algorithm that uses a novel kernel to explore the reward function space efficiently. The second work introduces SCALES, a framework that translates various fairness principles into fair decisions by translating them to a combination of utility, non-causal and causal components which are, in turn, mapped to elements of a Constrained Markov Decision Process. Our final work unifies the concepts of value alignment and fairness by extending the IRL problem to scenarios where the expert agent is fairness abiding. We propose FAIR-BOIRL, a new BO-based IRL algorithm that searches for solutions across both the reward function space and fairness principles. FAIR-BOIRL uses IM-GPTS, a novel acquisition function that uses an implicit multi-arm bandit strategy, to perform an efficient search."]},{"key":"dc:format.checksum.md5","label":"Dc Format Checksum Md5","values":["eac19808ab55054e814f7caf574e44b4","d33bda206263b8c8f9bd2cd7efd95f06"]},{"key":"dc:title","label":"Title","values":["TOWARDS HUMAN-CENTRIC AI: INVERSE REINFORCEMENT LEARNING MEETS ALGORITHMIC FAIRNESS"]}]}],"canonical_facts":{"dc:creator":["SREEJITH BALAKRISHNAN"],"dc:date.issued":["2023-02-22"],"dc:description.abstract":["This thesis focuses on two aspects of human-centric AI - value alignment and fairness. Our first work explores Inverse reinforcement learning (IRL), a potential solution to value alignment. We introduce BO-IRL, an IRL algorithm that uses a novel kernel to explore the reward function space efficiently. The second work introduces SCALES, a framework that translates various fairness principles into fair decisions by translating them to a combination of utility, non-causal and causal components which are, in turn, mapped to elements of a Constrained Markov Decision Process. Our final work unifies the concepts of value alignment and fairness by extending the IRL problem to scenarios where the expert agent is fairness abiding. We propose FAIR-BOIRL, a new BO-based IRL algorithm that searches for solutions across both the reward function space and fairness principles. FAIR-BOIRL uses IM-GPTS, a novel acquisition function that uses an implicit multi-arm bandit strategy, to perform an efficient search."],"dc:format.checksum.md5":["eac19808ab55054e814f7caf574e44b4","d33bda206263b8c8f9bd2cd7efd95f06"],"dc:identifier.uri":["https://scholarbank.nus.edu.sg/bitstreams/22425ef1-0070-4b89-9940-53b27ad307a8/download"],"dc:relation.isreferencedby":["https://scholarbank.nus.edu.sg/handle/10635/244765"],"dc:subject":["Human-centric AI, Value Alignment, Algorithmic Fairness, Inverse Reinforcement Learning, Bayesian optimization, Reinforcement Learning"],"dc:title":["TOWARDS HUMAN-CENTRIC AI: INVERSE REINFORCEMENT LEARNING MEETS ALGORITHMIC FAIRNESS"],"dc:type":["Thesis"]},"updated_at":"2026-07-24T03:31:13Z"}