← Back
OtherThe Hacker News·4 days ago

OpenAI Pauses Frontier RL Training as It Tightens Defenses Against Unsafe AI Behavior

OpenAI has temporarily halted reinforcement learning training for its latest models to strengthen its internal safety defenses and expand monitoring capabilities, citing growing risks as AI systems become more advanced. The two-week pause follows concerns about preventing incidents similar to a previous Hugging Face-related event. The company emphasized that developing and testing increasingly capable models carries escalating risks that necessitate enhanced protective measures.

Read full article at The Hacker News

Related Articles

OtherSchneier on Security·1 day ago

Friday Squid Blogging: Neon Flying Squid

The neon flying squid can fly in formation. The shoal of about 100 squid rose unexpectedly from a patch of the Pacific Ocean around 370 miles from Tokyo and glided near the boat for about 30 metres. The astonished researchers were the first to capture photographs of such a thing, which looked like the early stages of an alien invasion. They were probably neon flying squid (Ommastrephes bartramii), the subsequent study states, a species that is part of a 20-strong flying squid family that was known to leap from the water but, until then, was only rumoured to also be able to glide above it. The neon flying squid was able to gain such elevation by using the hyponome, a funnel-like muscular organ also present in other cephalopods, such as octopuses. The organ is able to force water out in a jet, propelling the body along both in and out of the sea. Photographs of the gliding squid show them with their arms (they have 10 limbs in all) splayed outwards. As usual, you can also use this squid post to talk about the security stories in the news that I haven’t covered. Blog moderation policy.

OtherDark Reading·2 days ago

OWASP Flags Top AI Skill Risks in New Security Blueprint

OWASP has released a new top 10 security list designed for the current threat landscape, introducing a Universal Skill Format to standardize and enhance security practices around AI integrations. The framework aims to address the distinct risks associated with AI implementations across development and deployment environments.