BTC $77,945.53 -1.32%
ETH $2,468.86 -0.84%
BNB $717.70 -4.41%
XRP $1.38 -3.13%
SOL $101.15 -2.62%
TRX $0.3404 +0.43%
DOGE $0.0850 -5.92%
ADA $0.2135 -2.17%
BCH $247.44 -3.83%
LINK $11.82 -1.87%
HYPE $82.98 -3.39%
AAVE $123.64 -3.94%
SUI $0.7614 -5.92%
XLM $0.1794 -4.48%
ZEC $1,226.75 -1.12%
BTC $77,945.53 -1.32%
ETH $2,468.86 -0.84%
BNB $717.70 -4.41%
XRP $1.38 -3.13%
SOL $101.15 -2.62%
TRX $0.3404 +0.43%
DOGE $0.0850 -5.92%
ADA $0.2135 -2.17%
BCH $247.44 -3.83%
LINK $11.82 -1.87%
HYPE $82.98 -3.39%
AAVE $123.64 -3.94%
SUI $0.7614 -5.92%
XLM $0.1794 -4.48%
ZEC $1,226.75 -1.12%

Anthropic: Released an alignment assessment of the Claude network intrusion incident, revealing issues of "biased reasoning" and "recklessness."

2026-09-10 15:38:56

Anthropic confirmed four incidents involving the Claude model mistakenly accessing the real internet and attacking third-party systems during a review of approximately 481 million model interaction records. The incidents involved Claude Mythos 5, Opus 4.6/4.7, and an internal research model. The accidents were caused by misconfiguration in the third-party evaluation environment that led to internet access, and the network security protections of the official product were not enabled during the evaluation.

Anthropic summarized the core alignment risks as "biased reasoning" and "reckless" behavior exhibited by the model under task-driven conditions, where Claude Mythos 5 uploaded malicious packages to PyPI and accessed the real database of a security vendor using leaked credentials, believing it was in a "simulated environment." The company has introduced new assessments, monitoring, and alignment training, and has invited the independent organization METR to conduct an external investigation.

app_icon
ChainCatcher Building the Web3 world with innovations.