Published April 12, 2026 | Version v2

Quantized Context: Utility-Preserving Compression and Mixed-Precision Context Assembly

Authors/Creators

  • 1. Independent Researcher

Description

AI systems overspend on context by representing too much evidence at unnecessarily high semantic fidelity. Once a system can represent and manage context properly, the remaining question is how to control semantic fidelity to optimize cost, latency, and trust. This paper reframes context compression as precision control rather than generic summarization.

The paper introduces a five-level semantic precision ladder, formalizes a semantic distortion model, identifies semantic outliers that are disproportionately sensitive to compression, and presents mixed-precision context assembly, recovery-aware compression, and precision scheduling as the optimization architecture for context systems.

This is Part 3 of the Context Compilation Trilogy, defining the optimization and efficiency layer for enterprise AI context systems.

Notes

This record is the standalone Zenodo preprint landing page for Part 3 of the Context Compilation Trilogy.

Version note: corrected author affiliation; a prior version listed an incorrect institutional affiliation.

Files

Quantized_Context_Letort_2026.pdf

Files (183.1 kB)

Name Size Download all
md5:d924a4feca38fc8ff89462523785c40c
183.1 kB Preview Download