Skip to content

docs: GLM-5.3-Flash attribution correction — wrong haraqat, not convention - #116

Merged
ronaldtse merged 1 commit into
mainfrom
docs/glm53-attribution
Sep 1, 2026
Merged

docs: GLM-5.3-Flash attribution correction — wrong haraqat, not convention#116
ronaldtse merged 1 commit into
mainfrom
docs/glm53-attribution

Conversation

@ronaldtse

Copy link
Copy Markdown
Contributor

Corrects the paper's protocol note and PUBLICATION-NOTES to match the measured attribution (rababa #69): wrong haraqat 10.05 percent of positions vs GLM-5.2's 2.64 (matching our r7 teacher's 2.62 to 0.02pp); missing 1.01, extra 0.20; the entire dagger-alif U+0670 convention effect is 0.125pp under position-derived rules with r7 and GLM-5.2 as <=0.009pp controls. Stronger dedicated-model finding: the newest generalist lost classical-Arabic mark knowledge its predecessor had. Method (decomposition + convention normalization with controls) noted as standard for any future LLM row.

…nvention 0.125pp

The first reading (orthography-shaped) was wrong; the per-position
decomposition with convention normalization and controls localizes
the regression to wrong haraqat at ~4x GLM-5.2's rate — which matches
our r7 teacher's to within 0.02pp. Missing 1.01%, extra 0.20%, the
whole U+0670 effect 0.125pp (310 marks, zero in GT; controls
<=0.009pp). Method note: apply the decomposition to any future LLM
row before interpreting its DER.
@ronaldtse
ronaldtse merged commit 9f546ad into main Sep 1, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant