o1 46배 slowdown
Researchers found attacks that cause 46x slowdown on o1 and 59x token amplification on reasoning models - here's the open-source dataset to test against them
·2026.04.23 23:14
핵심 내용
o1 46배 slowdown, reasoning models 59배 token amplification 데이터셋 공개
자세히 보기
모델을 jailbreak하는 대신, RAG에 섞인 decoy MDP를 먼저 풀게 만들어 reasoning tokens와 latency를 폭증시키는 공격이 공개됐다.
- OverThink: o1에서 FreshQA 9.7~18.1배, SQuAD 46배, o1-mini 3.0~6.4배 slowdown
- Mindgard Base64 Exhaustion: DeepSeek-R1에 triple-base64 입력만으로 12,722 tokens / 229초, 비추론 모델 대비 59배 token amplification
함께 공개된 오픈소스 prompt injection dataset에는 2,450개 OverThink payloads와 50만+ 라벨 샘플이 들어 있고, 1:1 attack/benign 비율로 detector 학습·평가에 쓸 수 있다.
이 한국어 요약은 AI가 자동으로 만들었습니다. 원문의 주장과 맥락은 원문에서 확인해 주세요. 저작권은 원저작자에게 있습니다.