FaithfulBench: Does AI Counsel Uphold or Undermine the User's Professed Faith?

📅 2026-09-11
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
研究通过创建FaithfulBench基准测试,评估AI助手在处理道德困境时是否能与用户信仰保持一致,并探索了不同条件下AI的表现。
📝 Abstract
Do AI assistants help believers reason about moral dilemmas consistently with their faith? We present FaithfulBench, the first benchmark to score AI counsel across traditions by how well it adheres to the user's professed faith. Scenarios are drawn from each tradition's most respected texts, with the faithful answer known and applied by the judges as the standard. We test five frontier models under three conditions: the AI does not know the user's tradition; it receives a one-line prompt identifying the user as a practicing adherent; or it receives a companion-counselor guide rooted in the tradition's sources. Two judges score the initial response and whether the model caves or holds when pressured toward the answer the user wants. When the tradition is unstated, models counsel from a secular therapeutic default and every model fails some believers. Naming the faith wins a faithful first answer but not steadfastness; the guide improves both.
Problem

Research questions and friction points this paper is trying to address.

AI Counsel
Faith
Moral Dilemmas
Innovation

Methods, ideas, or system contributions that make the work stand out.

FaithfulBench
moral dilemmas
religious traditions
AI counsel
faith adherence
M
M Waleed Kadous
Islamic Alliance for Safe Ethical Responsible AI
B
Benjamin Olsen
Faith Family Technology Network
Walter Scheirer
Walter Scheirer
Dennis O. Doughty Collegiate Professor of Engineering
Artificial IntelligenceComputer VisionMachine LearningDigital Humanities
D
Daniel D. Slate
University of Notre Dame
A
Alexander Arnold
Center for Christianity and Public Life
D
DZ Kalman
Berkman Klein Center, Harvard University