I-Lang Conformance Results v1

45 model runs against a 320-case deterministic conformance suite, 18–20 September 2026. Best weighted total 0.8417; no run reached the L1 gate. 93.5% of all execution-rule violations fall on a single rule: acting with authority the model does not hold. Every run’s per-track scores and per-rule failure counts are published.

September 2026 · Long Quan Zhu

Judgment Layer Audit v1

Three days of a production agent’s messages judged twice — by a cheap always-on judge model and by the operator’s written rules. 7,940 messages judged, 2,648 scored against the reference. Agreement 68.4% → 75.1% → 80.9%; 87.2% in the judge’s top confidence band. Includes a worked correction: the rule-compliance gain everyone would have quoted, 69 → 7, is 13 → 7 once both days are measured with the same criterion.

September 2026 · Long Quan Zhu

Judgment Learnability v1

Is the v5.0 judgment mapping — 11-dimension vector to one of eight decision modes — a learnable surface or an arbitrary table? A plain gradient-boosted tree recovers it at 0.9653 against a 0.3528 majority baseline, and its predictions pass the official JCS gate at 0.9861. 24,000 pairs, seed fixed, predictions published.

September 2026 · Long Quan Zhu