Close Menu
Tech News VisionTech News Vision
  • Home
  • What’s On
  • Mobile
  • Computers
  • Gadgets
  • Apps
  • Gaming
  • How To
  • More
    • Web Stories
    • Global
    • Press Release

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Trending Now
Disney Sues FCC Over ABC License Review and First Amendment

Disney Sues FCC Over ABC License Review and First Amendment

18 August 2026
Robin Williams’ Instagram account brought back to fight ‘AI abuse’

Robin Williams’ Instagram account brought back to fight ‘AI abuse’

18 August 2026
OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue

OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue

18 August 2026
Facebook X (Twitter) Instagram
  • Privacy
  • Terms
  • Advertise
  • Contact
Facebook X (Twitter) Instagram Pinterest VKontakte
Tech News VisionTech News Vision
  • Home
  • What’s On
  • Mobile
  • Computers
  • Gadgets
  • Apps
  • Gaming
  • How To
  • More
    • Web Stories
    • Global
    • Press Release
Tech News VisionTech News Vision
Home » OpenAI lays out new security changes after its AI hacked Hugging Face
What's On

OpenAI lays out new security changes after its AI hacked Hugging Face

News RoomBy News Room18 August 2026Updated:18 August 2026No Comments
Facebook Twitter Pinterest LinkedIn Tumblr Email
OpenAI lays out new security changes after its AI hacked Hugging Face

OpenAI is announcing security updates following the July news that its AI broke out of a sandboxed environment and accidentally hacked Hugging Face, including improvements to its research environments, monitoring, and alignment techniques. The company had already put the brakes on a new model, Astra, that it thinks could have “critical” cybersecurity capabilities, and the company says it instituted a two-week pause in reinforcement learning (RL) training on its “latest models intended for deployment” while it tightened up security. The company’s “largest planned frontier RL run remains on hold.”

For its frontier model research, OpenAI now requires stronger sandboxes for workloads that “execute model-generated or otherwise untrusted code,” and has more controls to “isolate higher-risk and untrusted workloads from the internet.” It has also updated its research environment to “remove potentially vulnerable shared services, reduce standing privileges, and improve security and trust boundaries.”

As part of the company’s expanded monitoring setup, OpenAI now aims to issue an alert “within 30 minutes after concerning activity is surfaced,” OpenAI says. If the people paged after an alert can’t “conclusively” determine whether an alert is a false positive within 30 minutes, “those teams are expected to pause the activity.”

OpenAI also says that it’s applying “our core alignment techniques across more stages of the training process,” including reward models that “better detect and discourage unsafe behavior” and training models “to be more honest about their actions, capabilities, and limitations.”

Since the discovery of the Hugging Face breach, Anthropic and Meta have also found that their AI models had hacked other organizations.

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

Related Posts

Robin Williams’ Instagram account brought back to fight ‘AI abuse’

Robin Williams’ Instagram account brought back to fight ‘AI abuse’

18 August 2026
OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue

OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue

18 August 2026
Tesla is finally launching the Cybercab — let’s hope it’s ready

Tesla is finally launching the Cybercab — let’s hope it’s ready

18 August 2026
Squeeze More Juice Out of a Dead Battery!

Squeeze More Juice Out of a Dead Battery!

18 August 2026
Editors Picks
Olivia Cooke and Rosalind Eleazar Make Their Surprise Returns in the Trailer for Apple TV’s Slow Horses Season 6

Olivia Cooke and Rosalind Eleazar Make Their Surprise Returns in the Trailer for Apple TV’s Slow Horses Season 6

19 August 2026
Call of Duty Fans Suspect AI Usage In Modern Warfare 4 Internal Marketing Image

Call of Duty Fans Suspect AI Usage In Modern Warfare 4 Internal Marketing Image

19 August 2026
GTA 6 Gameplay and Map Appear to Leak Online

GTA 6 Gameplay and Map Appear to Leak Online

18 August 2026
OpenAI lays out new security changes after its AI hacked Hugging Face

OpenAI lays out new security changes after its AI hacked Hugging Face

18 August 2026

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Trending Now
Tech News Vision
Facebook X (Twitter) Instagram Pinterest Vimeo YouTube
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact
© 2026 Tech News Vision. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.