papersTODAY 04:00 UTC
Taxonomy of Indirect Linguistic Encoding for LLM-Based Coded Language Detection
A revised arXiv preprint proposes a mechanism-oriented taxonomy of indirect linguistic expressions, the disguised phrasing such as algospeak and euphemisms that users adopt to hide sensitive meaning from platforms. The work organizes these encoding strategies by how they work rather than how they look, aiming to give LLM-based detection systems a more general basis for spotting obfuscated content. It targets the gap between surface-form moderation filters and adversarial evasion in social media text.