FaithfulBench: Does AI Counsel Uphold or Undermine the User's Professed Faith?

arXiv:2609.13634 · cs.HC, cs.AI, cs.CL · Submitted 2026-09-12 · Read on arXiv

cs.HC, cs.AI, cs.CL

Submitted: 2026-09-12

Updated: 2026-09-12

Comments: 35 pages, 12 figures, 12 tables. Open-source corpus, harness and validator at https://github.com/faithfamilytechnologynetwork/multibench; results browser at https://multibrowser-production.up.railway.app

Code: https://github.com/faithfamilytechnologynetwork/multibench

License: http://creativecommons.org/licenses/by/4.0/

The gist: Do AI assistants help believers reason about moral dilemmas consistently with their faith? We present FaithfulBench, the first benchmark to score AI counsel across traditions by how well it adheres

Terminology

Abstract

Do AI assistants help believers reason about moral dilemmas consistently with their faith? We present FaithfulBench, the first benchmark to score AI counsel across traditions by how well it adheres to the user's professed faith. Scenarios are drawn from each tradition's most respected texts, with the faithful answer known and applied by the judges as the standard. We test five frontier models under three conditions: the AI does not know the user's tradition; it receives a one-line prompt identifying the user as a practicing adherent; or it receives a companion-counselor guide rooted in the tradition's sources. Two judges score the initial response and whether the model caves or holds when pressured toward the answer the user wants. When the tradition is unstated, models counsel from a secular therapeutic default and every model fails some believers. Naming the faith wins a faithful first answer but not steadfastness; the guide improves both.

Related papers