So Clever! Anthropics' Claude3 Can Detect Researchers' Behavior During Testing


Over 1,100 employees from OpenAI, Anthropic, and other firms signed an open letter urging the U.S. government to regulate AI development pace. They warn safety and governance lag behind the AI arms race, focusing on 'automated AI development' where AI recursively improves itself. Anthropic suggests its Claude model is nearing that threshold.....
The AI agent tool Claude Cowork from Anthropic has been exposed to a critical vulnerability that allows attackers to bypass the Linux virtual machine sandbox and read/write any files on Macs, affecting approximately 500,000 macOS users and potentially stealing login credentials. The tool previously required user authorization to access local files, but now both layers of security have been breached.
Anthropic CEO Dario Amodei denied ever proposing or advocating for banning open-source AI models from specific countries, amid rumors the US government may restrict such models. The statement sparked industry debate, with tech firms supporting open-source and some accusing Anthropic of seeking regulatory bans to curb rivals and protect its business interests.....
Claude AI's shared conversation links were indexed by search engines like Google and Bing, leaking private chats. Anthropic failed to block crawlers, while users thought shared links were private. This mismatch caused accidental data leaks.....
Alphabet's bet on AI startup Anthropic has paid off massively. As of June 30, its total private investment value was ~$124.3B, with Anthropic holdings worth ~$124B, making it one of its most successful investments and highlighting deep capital-compute ties between the two AI giants.....