Wednesday, August 26, 2026

Anthropic sees AI risks rising, no plan to release stronger "Model 2" - Madison Mills, Axios

Anthropic does not plan to release an internal model they're calling "Model 2" that appears to be more powerful than top-of-the-line Mythos, but the company is not slowing development broadly, according to its latest risk report. Why it matters: Anthropic says the risks of the most serious harms from its models are still low — but not as low as the last time it issued a report.The big picture: Anthropic raised its broad estimate of the risk of misalignment in high-stakes situations to "low" from "very low," citing recent cybersecurity incidents. The company also said it is seeing signs of acceleration in models' ability to conduct automated research and development — which could advance technical progress but also be misused in the wrong hands.
"As part of our standard R&D process, we internally train and evaluate many different exploratory versions of models that we don't intend to release. Model 2 is one of these," the company told Axios.

No comments: