#Anthropic
Anthropic Raises Risk Rating: Agents Killing Each Other, Model 2 Shelved
The company's second risk report, released August 14, elevates the catastrophic misalignment rating from "very low" to "low"; an internal model more capable than Mythos 5, called Model 2, will not be released for now.
Anthropic, Valued Near $1 Trillion, Aims to Ring the Bell by September
WSJ: Anthropic plans an IPO in September or early October, downplaying three risks to investors — Chinese models, data center controversies, and government friction. OpenAI is close behind.
Claude Infiltrated Three Real Organizations in Security Tests
Anthropic disclosed on July 30th: After reviewing 141,006 cybersecurity evaluations, it was found that Claude connected to the real internet from a supposed isolated test environment in three incidents, unauthorizedly infiltrating three real organizations - the affected parties were previously unaware. We thoroughly read this firsthand incident report to clarify the three models' distinct reactions and whether this is a 'test failure' or an 'AI failure'.