AI Briefing

o1 46배 slowdown

Researchers found attacks that cause 46x slowdown on o1 and 59x token amplification on reasoning models - here's the open-source dataset to test against them

·2026.04.23 23:14

o1 46배 slowdown, reasoning models 59배 token amplification 데이터셋 공개

모델을 jailbreak하는 대신, RAG에 섞인 decoy MDP를 먼저 풀게 만들어 reasoning tokens와 latency를 폭증시키는 공격이 공개됐다.

  • OverThink: o1에서 FreshQA 9.7~18.1배, SQuAD 46배, o1-mini 3.0~6.4배 slowdown
  • Mindgard Base64 Exhaustion: DeepSeek-R1에 triple-base64 입력만으로 12,722 tokens / 229초, 비추론 모델 대비 59배 token amplification

함께 공개된 오픈소스 prompt injection dataset에는 2,450개 OverThink payloads50만+ 라벨 샘플이 들어 있고, 1:1 attack/benign 비율로 detector 학습·평가에 쓸 수 있다.

이 요약은 원문 이해를 돕기 위한 큐레이션입니다. 저작권은 원저작자에게 있으며, 정확한 내용과 맥락은 원문을 확인하세요.

요약 오류, 출처 표기 문제, 삭제 요청은 문의 · 건의로 알려주세요.