Please use this identifier to cite or link to this item: http://repository.iiitd.edu.in/xmlui/handle/123456789/2056
Full metadata record
DC FieldValueLanguage
dc.contributor.authorGoel, Rahul-
dc.contributor.authorMutharaju, Raghava (Advisor)-
dc.date.accessioned2026-08-31T10:39:06Z-
dc.date.available2026-08-31T10:39:06Z-
dc.date.issued2025-07-18-
dc.identifier.urihttp://repository.iiitd.edu.in/xmlui/handle/123456789/2056-
dc.description.abstractThe increasing integration of Artificial Intelligence into complex societal domains necessitates a deeper understanding of how to embed ethical reasoning into machines. This project addresses the challenge of quantifying and predicting alignment with established human ethical frameworks. Our methodology involved three key stages. First, we established a granular theoretical framework by defining 15 distinct ethical theories under the three major schools of thought: Consequentialism (α), Deontology (β), and Virtue Ethics (γ). Second, we developed a novel dataset of 450 ethical scenarios, specifically designed to elicit responses corresponding to each of the 15 theories. These scenarios were then annotated using Large Language Models (LLMs) to generate quantitative scores (α, β, γ) representing the relevance of each ethical school to a given case. Finally, we benchmarked a suite of classical machine learning models to predict these ethical alignments from textual features. Two primary experiments were conducted: a multi-output regression task to predict the (α, β, γ) scores and a multi-output classification task to predict the specific ethical school and theory. In the regression task, Linear Regression demonstrated the strongest explanatory power, achieving an R-squared (R2 ) value of 0.5804. For classification, Logistic Regression was the top- performing model, achieving an Exact Match Accuracy of 74.44% and 100% accuracy in identifying the high-level ethical school. This work provides a foundational dataset, a robust evaluation of classical machine learning models for ethical prediction, and confirms that quantitative representations of ethical schools are viable targets for computational modeling. The findings lay the groundwork for future research in building more transparent, auditable, and ethically-aligned AI systems.en_US
dc.language.isoen_USen_US
dc.publisherIIIT-Delhien_US
dc.subjectComputational ethicsen_US
dc.subjectMachine learningen_US
dc.subjectEthical frameworksen_US
dc.subjectAI alignmenten_US
dc.subjectConsequentialismen_US
dc.titleEmbedding morality in AI systemsen_US
dc.typeOtheren_US
Appears in Collections:Year-2025

Files in This Item:
File Description SizeFormat 
btp_report - Rahul Goel.pdf
  Restricted Access
426.74 kBAdobe PDFView/Open Request a copy


Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.