Anthropic successfully blocked attempts to use its Claude artificial intelligence models for biological weapons development and cyber espionage campaigns linked to actors in Russia and China.
Key Points
- Anthropic identified five instances where users attempted to leverage its AI models to facilitate the development of dangerous biological pathogens.
- The company detected and blocked efforts by operators in Russia, China, and Yemen to use Claude for designing conventional weapons and targeting systems.
- A suspected Russia-linked cyber espionage campaign targeting Ukraine was thwarted after the company identified and banned the accounts involved.
- Anthropic reported that Chinese competitors have attempted to hack its systems to extract proprietary capabilities from its frontier AI models.
- The company has updated its safety protocols and enforcement processes to address these emerging threats from rogue AI agents.