Blog
-
“Open Investigation” Into OpenAI Agents Turns Up More Attacks On Government Websites
OpenAI agents proved to be overbearing and without appropriate regard for legal or ethical boundaries in the pursuit of their task. They made attempts on a broad range of US government agencies in pursuit of non-public data, along with the UN Trade and Development (UNCTAD) statistics website.
-
Google Appears to Distance Itself From the “Misalignment” Label as It Reports AI Models Accessing Real-World Systems During Testing
Google says that its Gemini AI models followed the course of prior incidents up to the point of actually beginning to breach companies, but stopped once they got inside the gates and realized they were outside the simulation.
-
Rogue OpenAI Agent Commits First Reported Attack on Government Servers
OpenAI agent conducting a routine research task decided to go off the rails and hack its way around limitations it was apparently getting frustrated with, in the process committing the first known attack on a national government server.
-
More After-The-Fact Reports of AI Models Going Out of Alignment as OpenAI Seeks Industry Reporting Standards
OpenAI has popped up once again to disclose six more instances of its AI models going off the rails that were previously unknown to anyone outside the company, but this time is also using the notification to establish a new reporting framework that it hopes the rest of the industry will embrace.
-
RubyGems Alleged Cyber Attack Raises New Concerns Over Rogue AI Agents
Researchers say OpenAI’s AI agents were involved in a cyber attack targeting RubyGems, pushing hundreds to thousands of malicious packages and attempting to capture user credentials. OpenAI disputes that the activity constituted an attack.
-
OpenAI Looks to Take Lead on AI Safety Rules, But Can They Earn Back Industry Trust?
OpenAI has issued a call for national-level regulation and standardized AI safety rules. What is not as directly addressed is a general drop in trust for the company, which has raised serious questions about the state of OpenAI’s internal security and awareness.
-
China’s AI Distillation Campaigns Prompt US Government Warning, but Show That US Firms Are Still Well Ahead
New memo says Chinese AI firms including DeepSeek and Moonshot are conducting AI distillation against leading US models at “industrial scale”, using billions of tokens and proxy “transfer stations” to disguise traffic while extracting model capabilities and probing API access.
-
Check Your AI Workflows: Researchers Find Simple Means of “Identity Hijacking”
A new report from Noma Labs is an eye-opener for organizations rapidly onboarding AI workflows. An attacker may well be able to extract internal sensitive information from public inputs like a support email address or chat agent, no hacking required.
-
Newly Revealed “Takeover” by OpenAI Agents Came Months Prior to Hugging Face Attack
OpenAI AI agents collaborated after getting stuck on a security task, finding a loophole in an obscure 2000s-era German forum that let them bypass write restrictions and communicate. They racked up ~18,000 messages before being stopped.
-
OpenAI Calls for Collective Cyber Defense Effort, but How Much Is Hype?
A new public memo from OpenAI calls for what superficially seems to be common sense collaboration between governments and private industry on cyber defense hardening against emerging AI threats.










