Same time, same object, different value: probing a small language model for conflicts inside the context window
A pre-registered pilot on Qwen2.5-0.5B. A first protocol met all its criteria (AUROC 1.00), but post-hoc controls showed the result was lexical. A second protocol, trained on colours and materials and tested on sizes, an attribute never seen in training, met all its criteria again: AUROC 0.99 pooled, 0.98 against temporal update, with an increment beyond next-token entropy. The note reports a transferring direction that is sensitive to time and object. It does not claim a detector of contradiction as such. Frozen protocols with SHA-256 hashes, scripts and results are included.
Authors
- Yanush Feshter (ORCID: https://orcid.org/0009-0002-1330-7530)
Publication Details
- Journal
- Zenodo (CERN European Organization for Nuclear Research)
- Published
- 2026-10-03
- DOI
- https://doi.org/10.5281/zenodo.23127023
- Primary Topic
- Natural Language Processing Techniques
- Type
- preprint