Towards a neuro-symbolic approach to moral judgment
Name
wing-spwing-meng-eecs-2024-thesis.pdf
Description
Thesis PDF
Size
1.12 MB
Format
Adobe PDF
Checksum (MD5)
508f6bbb30d4521c42d16a799c680d5f
Author(s)
Wing, Shannon P.
Advisor(s)
Tenenbaum, Joshua
Date Issued
February 2024
Publisher
Massachusetts Institute of Technology
Abstract
The goal to build a safe Artificial General Intelligence requires an advancement beyond any single human being’s moral capacity. For the same reason why we desire democracy, a moral AGI will need to be able to represent a wide array of perspectives accurately.
While there has been a lot of work to push AI towards correctly answering unanimously agreed upon moral questions, we will take a different approach and ask: What do we do for the space where there is no correct answer, but perhaps multiple? Where there are better and worse arguments? We will investigate one complex moral question, where the empirical human data strays from unanimous agreement, evaluate chatGPT’s success, and build towards a neuro-symbolic framework to improve upon this baseline. By investigating one problem in depth, we hope to uncover nuances, intricacies, and details that might be overlooked in a broader exploration. Our insights intend to spark curiosity, rather than provide answers.
MIT Department
Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science
Terms of Use
Attribution-NonCommercial-NoDerivatives 4.0 International (CC BY-NC-ND 4.0)
Copyright retained by author(s)
Persistent DSpace Link