Anthropic, 안전장치에 막힌 요청에도 다시 요금을 매긴다 — 생물학·증류 공격·프런티어 모델 개발 세 분야
Anthropic이 오늘부터 Claude가 답하기 전에 안전장치가 막은 요청에도 다시 요금을 매긴다고 밝혔다. 모든 요청이 아니라 오탐률이 낮은 세 분야에만 해당한다. 생물학, 증류 공격(Claude의 답을 대량으로 뽑아 다른 모델을 학습시키려는 시도), 프런티어 LLM 개발이다. 최근 몇 주 사이 자사 시스템을 겨냥한 조직적인 공격을 봤다는 게 이유다.
댓글 반응 2개
- @ncsuian♥2
How does charging in any way serve as a defense against attacks? Just makes people pay @AnthropicAI when it suspects they might be asking a sensitive question it decides not to answer. In other words - payment for services not rendered. - @DataDiscovered
exactly. blocked requests still burn gpu cycles, so free blocks just mean unlimited free attack probes. the real tell: 99.7% of users never hit one. if your billable blocks are nonzero you already know why