Published October 24, 2025 | Version v1

Hallucination as Semantic Misalignment: A Categorical Approach via Lax Institutions

  • 1. ZYX Corp

Description

Large Language Models (LLMs) frequently produce hallucinations—outputs that are syntactically fluent yet semantically unfaithful to facts, sources, or context. While existing research has identified statistical mechanisms, detection methods, and mitigation strategies, a unified semantic foundation remains elusive. This paper reframes hallucination as semantic misalignment, formalized through the lens of Institution theory (Goguen & Burstall, 1992) extended to 2-categories. We model hallucination as a breakdown of the satisfaction condition: the naturality between sentences and models fails under vocabulary morphisms, quantified by lax natural transformations. This formalism captures the dual nature of hallucination—a structural error in fact-sensitive domains (medicine, law, science) yet a creative resource in artistic contexts (fiction, poetry, metaphor). We propose a domain-layered naturality control framework with strict naturality for factual claims, semi-strict for opinions, and lax for creative expression. Combined with external semantic verification (NLI consistency, semantic entropy, retrieval divergence) and evaluation redesign that does not over-penalize uncertainty, our approach provides a principled path toward tuning rather than eliminating hallucinations. We critically examine the pitfalls of guardrail-centric approaches and advocate for transparent, semantics-first alignment.

Files

Kano2025_Hallucination_Semantic_Misalignment.pdf

Files (2.0 MB)

Name Size Download all
md5:82e43942146af3c0890a5b33c66a0201
890.9 kB Preview Download
md5:2a5716bb45ae1ba1824ce6f37ffcbd61
580.9 kB Preview Download
md5:b0bde3589b760d609997eefd62ec667c
110.8 kB Preview Download
md5:fb5e3f8bfba19b4e30a45ee06e5fdbab
395.4 kB Preview Download

Additional details

Software

Development Status
Active