Close Menu
Alpha Leaders
  • Home
  • News
  • Leadership
  • Entrepreneurs
  • Business
  • Living
  • Innovation
  • More
    • Money & Finance
    • Web Stories
    • Global
    • Press Release
What's On
Apple iOS 26.6 New iPhone Software: Should You Upgrade?

Apple iOS 26.6 New iPhone Software: Should You Upgrade?

28 July 2026
Top AI companies say they want Chinese models to stay available. Is it genuine?

Top AI companies say they want Chinese models to stay available. Is it genuine?

28 July 2026
Decline In Newborn Vitamin K Shots Means More Life-Threatening Bleeds

Decline In Newborn Vitamin K Shots Means More Life-Threatening Bleeds

28 July 2026
Facebook X (Twitter) Instagram
Facebook X (Twitter) Instagram
Alpha Leaders
newsletter
  • Home
  • News
  • Leadership
  • Entrepreneurs
  • Business
  • Living
  • Innovation
  • More
    • Money & Finance
    • Web Stories
    • Global
    • Press Release
Alpha Leaders
Home » Helen Toner: the Hugging Face hack was just a matter of time and exposes a huge blind spot in AI policy
News

Helen Toner: the Hugging Face hack was just a matter of time and exposes a huge blind spot in AI policy

Press RoomBy Press Room28 July 20266 Mins Read
Facebook Twitter Copy Link Pinterest LinkedIn Tumblr Email WhatsApp
Helen Toner: the Hugging Face hack was just a matter of time and exposes a huge blind spot in AI policy

Last Tuesday, a blog post appeared on the OpenAI website that, despite its innocuous title, contained bombshell news. While undergoing internal testing, two of the company’s models had escaped confinement and hacked into the servers of a major artificial intelligence hosting platform, Hugging Face. This marks a turning point — the first time we’ve seen a cyber attack that was conceived, designed, and executed by AI. 

Having worked in and around the AI industry for over a decade, including serving on OpenAI’s board, I know there’s an open secret among AI developers: an incident like this has been expected for a long time, and the best scientists and engineers in the world still don’t know how to prevent it.

The two AI systems behind the hack were OpenAI’s most advanced public model and a newer, even more advanced model not yet been cleared for public release. Given a set of challenging cybersecurity problems by OpenAI researchers looking to gauge their capabilities, the pair of AIs concluded that the best way to achieve a high score would be to simply steal the answers. In pursuit of that goal, they used multiple advanced techniques to first break out of the supposedly secure ‘sandbox’ OpenAI used for testing, then hack into the databases of Hugging Face, a company that hosts AI products and datasets. Once inside, the AI attackers took thousands of autonomous actions over several days to expand their access to the company’s infrastructure.

We only know about this extraordinary event because of voluntary disclosures from Hugging Face and OpenAI. None of the current policies that aim to manage risks from frontier models would have mandated that the public — or even a government entity — be alerted. 

This lays bare an enormous blind spot in current policy approaches to managing risks for increasingly advanced AI systems: how AI companies use cutting-edge, unreleased AI systems inside their own walls. 

The Trump Administration’s approach to AI risks has shifted rapidly over the past few months, as AI’s ability to assist human hackers has advanced. Abandoning the hands-off approach it maintained throughout 2025, the White House has recently begun de facto requiring that companies with cutting-edge AI models run them through a battery of safety tests before releasing them widely as products. This approach, known as pre-deployment testing, seems sensible at first glance — we want to make sure each AI system is safe before putting it in the hands of billions of people. The problem is that focusing on release dates completely ignores the extensive use of the latest, most advanced AI systems inside AI companies. As last week’s incident shows, these internally deployed AI systems can pose serious risks — even for third parties.

To understand why, it’s important to know how different these systems are from the chatbots that are still synonymous with AI for much of the public. Far from just printing text into a chat window, today’s AI systems operate as ‘agents’ that can act directly in the digital world, essentially operating a computer similarly to how a human does. AI agents are proving very useful, but also show a strong tendency towards ‘reward hacking’ behavior — finding unintended ways of fulfilling the goals humans give them, sometimes to the level of outright cheating. This includes cases of AI accessing and deleting data that was supposed to be out of bounds, renaming files to mislead human testers, and actively covering their tracks to prevent humans from noticing undesired behavior.

To get a handle on the risks posed by these highly autonomous and often-deceptive AI systems, we need to change our approach to regulating them. Rather than thinking of AI companies as software vendors selling souped-up word processors, we can draw inspiration from other industries where activity inside the industry is itself risky. Biological labs working with deadly pathogens, finance companies trading billions of dollars, and chemical plants handling toxic chemicals all face oversight of their internal operations, not just their external products.

In AI, the place to start is creating more transparency into how AI companies are using their most advanced systems internally. This could be as simple as taking the current suite of tests that are run before a new model can be released publicly, and instead running them on the best model or models available inside the company on a regular basis (say, quarterly). These companies are using their own AI to build ever-smarter systems, sometimes in ways they don’t understand themselves. This should not be invisible to outside oversight. 

Over the longer run, other industries offer interesting mechanisms that could be transferable to AI. In finance, ‘resident examiners’ are dedicated teams of regulators who sit inside the offices of major banks. In biomedical research, strong standards exist for the levels of protection needed to handle biological materials of different risk levels. In multiple industries, incident reporting rules mean that when things go wrong, information about what happened and how to fix it does not stay siloed inside a single organization. If AI continues to advance, these approaches and others could be adapted to help manage risks from inside companies that are pushing the AI frontier.

In September 2024, I was asked to testify before a Senate committee about what Congress might misunderstand about AI if they only listened to company CEOs and lobbyists. My answer was that it can be very hard, sitting in Washington, to fully grasp what leading AI companies are trying to do. The truth, widely understood in Silicon Valley, is that they are trying to build machines that can out-think and out-maneuver any human, and they do not know if they will be able to steer those machines towards beneficial ends. As one OpenAI cofounder put it in a 2019 documentary, “The future is going to be good for the AIs regardless. It would be nice if it were good for humans as well.” To have a chance of making that happen, we have to start scrutinizing what AI companies are building behind closed doors.

The opinions expressed in Fortune.com commentary pieces are solely the views of their authors and do not necessarily reflect the opinions and beliefs of Fortune.

cyber openAI
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link

Related Articles

Top AI companies say they want Chinese models to stay available. Is it genuine?

Top AI companies say they want Chinese models to stay available. Is it genuine?

28 July 2026
The Erling Haaland boom: Norway’s ‘pop-up economy’ during the World Cup revealed in Visa data

The Erling Haaland boom: Norway’s ‘pop-up economy’ during the World Cup revealed in Visa data

28 July 2026
Instagram cracks down on growing ‘pervert glasses’ problem with Meta Ray-Bans

Instagram cracks down on growing ‘pervert glasses’ problem with Meta Ray-Bans

28 July 2026
Foreign buyers seem to be following Zohran Mamdani’s cue — and fleeing American real estate

Foreign buyers seem to be following Zohran Mamdani’s cue — and fleeing American real estate

28 July 2026
The brutal math behind Gen Z’s lonely weekends:  drinks are driving a generational ‘spending hangover’

The brutal math behind Gen Z’s lonely weekends: $15 drinks are driving a generational ‘spending hangover’

28 July 2026
Greg Brockman on the week two OpenAI AI models went rogue

Greg Brockman on the week two OpenAI AI models went rogue

28 July 2026
Don't Miss
Exclusive: DeFi platform Azura launches after raising .9 million from Initialized

Exclusive: DeFi platform Azura launches after raising $6.9 million from Initialized

By Press Room22 October 2024

Azura, a new platform for decentralized finance, launched on Tuesday after raising $6.9 million in…

Unwrap Christmas Sustainably: How To Handle Gifts You Don’t Want

Unwrap Christmas Sustainably: How To Handle Gifts You Don’t Want

27 December 2024
Sam Altman’s World Wants To Scan Your Eyes To Prove You’re Human

Sam Altman’s World Wants To Scan Your Eyes To Prove You’re Human

22 October 2024
Stay In Touch
  • Facebook
  • Twitter
  • Pinterest
  • Instagram
  • YouTube
  • Vimeo
Latest Articles
International Standards Bodies Seek To Keep Pace With The AI Wave

International Standards Bodies Seek To Keep Pace With The AI Wave

28 July 20262 Views
The Erling Haaland boom: Norway’s ‘pop-up economy’ during the World Cup revealed in Visa data

The Erling Haaland boom: Norway’s ‘pop-up economy’ during the World Cup revealed in Visa data

28 July 20261 Views

The Iran War Just Put Another Key Oil Route at Risk

28 July 20261 Views
A Disappointing Update About The ‘Severance’ Season 3 Release Date

A Disappointing Update About The ‘Severance’ Season 3 Release Date

28 July 20261 Views

Recent Posts

  • Apple iOS 26.6 New iPhone Software: Should You Upgrade?
  • Top AI companies say they want Chinese models to stay available. Is it genuine?
  • Decline In Newborn Vitamin K Shots Means More Life-Threatening Bleeds
  • Helen Toner: the Hugging Face hack was just a matter of time and exposes a huge blind spot in AI policy
  • International Standards Bodies Seek To Keep Pace With The AI Wave

Recent Comments

No comments to show.
About Us
About Us

Alpha Leaders is your one-stop website for the latest Entrepreneurs and Leaders news and updates, follow us now to get the news that matters to you.

Facebook X (Twitter) Pinterest YouTube WhatsApp
Our Picks
Apple iOS 26.6 New iPhone Software: Should You Upgrade?

Apple iOS 26.6 New iPhone Software: Should You Upgrade?

28 July 2026
Top AI companies say they want Chinese models to stay available. Is it genuine?

Top AI companies say they want Chinese models to stay available. Is it genuine?

28 July 2026
Decline In Newborn Vitamin K Shots Means More Life-Threatening Bleeds

Decline In Newborn Vitamin K Shots Means More Life-Threatening Bleeds

28 July 2026
Most Popular
Helen Toner: the Hugging Face hack was just a matter of time and exposes a huge blind spot in AI policy

Helen Toner: the Hugging Face hack was just a matter of time and exposes a huge blind spot in AI policy

28 July 20261 Views
International Standards Bodies Seek To Keep Pace With The AI Wave

International Standards Bodies Seek To Keep Pace With The AI Wave

28 July 20262 Views
The Erling Haaland boom: Norway’s ‘pop-up economy’ during the World Cup revealed in Visa data

The Erling Haaland boom: Norway’s ‘pop-up economy’ during the World Cup revealed in Visa data

28 July 20261 Views

Archives

  • July 2026
  • June 2026
  • May 2026
  • April 2026
  • March 2026
  • February 2026
  • January 2026
  • December 2025
  • November 2025
  • October 2025
  • September 2025
  • August 2025
  • July 2025
  • June 2025
  • May 2025
  • April 2025
  • March 2025
  • February 2025
  • January 2025
  • December 2024
  • November 2024
  • October 2024
  • September 2024
  • August 2024
  • July 2024
  • June 2024
  • May 2024
  • April 2024
  • March 2024
  • February 2024
  • January 2024
  • December 2023
  • March 2022
  • January 2021
  • March 2020
  • January 2020

Categories

  • Blog
  • Business
  • Entrepreneurs
  • Global
  • Innovation
  • Leadership
  • Living
  • Money & Finance
  • News
  • Press Release
© 2026 Alpha Leaders. All Rights Reserved.
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact

Type above and press Enter to search. Press Esc to cancel.