코인 매입
시장
현물
선물
재테크
이벤트
더 알아보기
reward-center초보자 존
홈 피드빠른 소식 정보
Claude Opus5 is nearly immune to Word Injection attacks in a browser scenario, with none of the 129 test cases being compromised.
  • MODE0%
  • MYTH0%
  • OPUS0%

According to Watchful AI, Anthropic disclosed in the Claude Opus 5 system card that the model is almost immune to prompt injection attacks in the browser agent scenario, with none of the 129 test scenarios being compromised. This breakthrough is significant as OpenAI publicly acknowledged last December that prompt injection may never be fully resolved. In Gray Swan's general prompt injection benchmark test, after 15 attacks, Opus 5 achieved a success rate of only 2.0%, significantly better than the previous generation Opus 4.8 at 5.5%, as well as outperforming Mythos 5 at 2.6% and Fable 5 at 2.8%.

Prompt injection is considered the most serious security vulnerability faced by AI agents—attackers manipulate input content such as hidden text embedded in a webpage to bypass model instructions and perform malicious operations. A zero-attack rate is only achieved in products like Claude Cowork with Auto Mode enabled. This achievement marks a transition in AI agent security from "ongoing patches" to "structural defense."

Click on the original article link below to join the Watchful AI · Feishu AI News Channel and receive real-time monitoring of global AI trends and news 24/7.

출처:BlockBeats

면책 조항: 현재 콘텐츠는 제3자 관점에서 제공되거나 제3자 관점에서 AI가 직접 번역한 것입니다. CoinEx는 콘텐츠의 진위성, 정확성, 독창성을 보장하지 않으며 CoinEx의 투자 조언으로 간주하지 않습니다. 암호화폐 가격은 변동성이 크므로 잠재적인 위험에 유의하시기 바랍니다.

인기 검색
  • 코인
    가격
    24시간 변동