Published February 20, 2026 | Version v1

How Far Does the Trolley Problem Go in AI Ethics Evaluation? Limits of a Canonical Benchmark and the Risks of Its Misuse

Authors/Creators

Description

This paper examines both the genuine utility and structural limitations of the trolley problem as an AI ethics evaluation instrument. While effective for surfacing value priorities under forced choice, strong trolley-problem performance is not a safety certificate and may produce false assurance. We identify three structural limitations and argue for evaluation frameworks that extend beyond forced-choice benchmarks.

 

Files

TrolleyLimits_Draft_EN_v3.pdf

Files (124.8 kB)

Name Size Download all
md5:22cdb30d563ce9a3311b6e40bdb780cd
124.8 kB Preview Download