Aktualności1 sierpnia 2026 A fundamental LLM flaw: forged chain-of-thought bypasses safety
Researchers showed at ICML that large language models identify text roles by style, not by tags. A forged chain-of-thought bypasses safety in OpenAI, Anthropic, Alibaba and DeepSeek models — and the problem may be unsolvable at the model level.