← Back to Home

Tag: 模型安全 (2 articles)

Introducing Claude Opus 4.7

Anthropic releases Claude Opus 4.7, focusing on enhanced complex coding and long-running task capabilities, with its 'self-verification' mechanism marking a key step towards more autonomous AI agents.

Anthropic News ·

Stealing Reasoning Traces from Proprietary LLM APIs

Research reveals that major LLMs reuse encryption keys for reasoning blocks across models, allowing attackers to recover hidden reasoning via weaker model jailbreaks and exposing new prompt injection risks.

Simon Willison ·