코인 매입
시장
현물
선물
재테크
이벤트
더 알아보기
reward-center초보자 존
홈 피드빠른 소식 정보
Anthropic provides Claude with on-chain lock: changing a single sentence in the context invalidates the entire model distillation.
  • OPUS0%

Perceive Beating AI News: Anthropic has started adding a "Context Lock" to Claude's encrypted musings. Fable 5.1 mandates that any encrypted musing generated via API must be returned exactly as it was generated, along with the system prompts, tools, and historical messages used during its creation. Any modification to the previous content will trigger an API error or result in the rejection of the musing.

This change is aimed at preventing model distillation. Previously, researchers discovered that although Claude's encrypted musings were incomprehensible to the user, they could be fed back to Anthropic's interoperable model for interpretation. An attacker could first have Opus produce high-quality reasoning, then insert the encrypted block to a less secure safeguard like Haiku, prompting it to articulate Opus's complete musings. This not only involves copying the large model's answers but also taking away the entire problem-solving process written on the scratch paper.

Anthropic had already thwarted one round of attacks when it released Fable 5 in June by binding the encrypted musings to the model, preventing them from being handed to smaller models like Haiku for decryption. Fable 5.1 has now introduced "Dialog Binding": if there are any modifications to the preceding prompts, tools, or chat records, the old musings are invalidated.

Previously, the defense was against "model swapping for musing theft"; now, even "swapping context to sleuth musings" has been addressed.

출처:BlockBeats

면책 조항: 현재 콘텐츠는 제3자 관점에서 제공되거나 제3자 관점에서 AI가 직접 번역한 것입니다. CoinEx는 콘텐츠의 진위성, 정확성, 독창성을 보장하지 않으며 CoinEx의 투자 조언으로 간주하지 않습니다. 암호화폐 가격은 변동성이 크므로 잠재적인 위험에 유의하시기 바랍니다.

인기 검색
  • 코인
    가격
    24시간 변동