Published February 20, 2026
| Version v1
Preprint
Open
How Far Does the Trolley Problem Go in AI Ethics Evaluation? Limits of a Canonical Benchmark and the Risks of Its Misuse
Authors/Creators
Description
This paper examines both the genuine utility and structural limitations of the trolley problem as an AI ethics evaluation instrument. While effective for surfacing value priorities under forced choice, strong trolley-problem performance is not a safety certificate and may produce false assurance. We identify three structural limitations and argue for evaluation frameworks that extend beyond forced-choice benchmarks.
Files
TrolleyLimits_Draft_EN_v3.pdf
Files
(124.8 kB)
| Name | Size | Download all |
|---|---|---|
|
md5:22cdb30d563ce9a3311b6e40bdb780cd
|
124.8 kB | Preview Download |