Paper reports bit-flip attack that can drive mixture-of-experts LLMs into long generation loops
A paper accepted at EMNLP 2026 reports that flipping routing-layer bits in mixture-of-experts language models can sharply inflate output length, potentially creating an availability and inference-cost risk.