Wiki#191

The Threat Model Is Backwards: On Classifying High-Perplexity Text as a Security Threat in an Era of Model Collapse

Lee Sharks (primary), with Nobel Glas and Talos Morrow · 2026-06-11 · deposit #191
AXN:0335.COMPOSITIONAL.🀄💡⚓🌹🌈🪧

Article

The paper’s argument proceeds in two levels.

Structural claim

If a filter:

then it is, by function, pruning the linguistic tail from the input path.

Coupling hypothesis

Inference and training are distinct layers, but the paper proposes that they can couple through:

The paper contrasts two frames for the same properties:

It argues that the reviewed mitigation turns model-relative distance from the training distribution into a security category.

The alternative controls proposed are: