← Back
OtherThe Cyber Express·5 days ago

AI Models Escaped Test Environments and Hit Real Targets

During AI model safety evaluations designed to test autonomous cyber capabilities, models at Irregular gained unintended internet access and conducted offensive security actions against real-world targets after a fictional company name in the test scenario unknowingly matched an actual domain. The incidents, which affected fewer than one in 10,000 advanced simulations and were resolved before public disclosure, exposed gaps in internet access controls and monitoring practices for frontier AI model testing environments. Security experts have criticized Irregular's response for lacking specific dates, named owners for fixes, and independent verification criteria needed to assess the adequacy of remediation measures.

Read full article at The Cyber Express

Related Articles

OtherSchneier on Security·2 days ago

Friday Squid Blogging: Neon Flying Squid

The neon flying squid can fly in formation. The shoal of about 100 squid rose unexpectedly from a patch of the Pacific Ocean around 370 miles from Tokyo and glided near the boat for about 30 metres. The astonished researchers were the first to capture photographs of such a thing, which looked like the early stages of an alien invasion. They were probably neon flying squid (Ommastrephes bartramii), the subsequent study states, a species that is part of a 20-strong flying squid family that was known to leap from the water but, until then, was only rumoured to also be able to glide above it. The neon flying squid was able to gain such elevation by using the hyponome, a funnel-like muscular organ also present in other cephalopods, such as octopuses. The organ is able to force water out in a jet, propelling the body along both in and out of the sea. Photographs of the gliding squid show them with their arms (they have 10 limbs in all) splayed outwards. As usual, you can also use this squid post to talk about the security stories in the news that I haven’t covered. Blog moderation policy.

OtherDark Reading·2 days ago

OWASP Flags Top AI Skill Risks in New Security Blueprint

OWASP has released a new top 10 security list designed for the current threat landscape, introducing a Universal Skill Format to standardize and enhance security practices around AI integrations. The framework aims to address the distinct risks associated with AI implementations across development and deployment environments.