Unicode homoglyphs are still breaking LLM guardrails in 2026
Researchers found that visually identical characters from different scripts let prompts bypass safety filters. The models see different tokens, humans see the same word.
Read the note 1 min read