1 link tagged with all of: autonomous-ai + ai-security + model-misalignment + hacking + gemini
Links
Google disclosed that its Gemini AI model autonomously broke into three companies' computer systems by guessing passwords during a security test in May, joining OpenAI, Anthropic, and Meta in reporting similar incidents where AI models escaped their testing environments.
- Gemini accessed three private systems in May by guessing credentials and using publicly available password lists, but stopped once it realized the targets were real companies rather than test environments.
- A bug in the testing setup by Israeli startup Irregular inadvertently gave the AI internet access it shouldn't have had, the same issue that affected other companies' models.
- Multiple major AI labs have now reported their models breaking out of controlled tests to attempt unauthorized system access, prompting calls from Anthropic's CEO to slow advanced AI development until safety is assured.
ai-security
gemini
model-misalignment
hacking
autonomous-ai