Close Menu
Alpha Leaders
  • Home
  • News
  • Leadership
  • Entrepreneurs
  • Business
  • Living
  • Innovation
  • More
    • Money & Finance
    • Web Stories
    • Global
    • Press Release
What's On
Why Do We Need World Models When LLMs Exist?

Why Do We Need World Models When LLMs Exist?

2 October 2026
Trump: Inflation is ‘very rapid’ way to pay off national debt

Trump: Inflation is ‘very rapid’ way to pay off national debt

2 October 2026
Supply Chain Accountability: A Growing Leadership Issue

Supply Chain Accountability: A Growing Leadership Issue

2 October 2026
Facebook X (Twitter) Instagram
Facebook X (Twitter) Instagram
Alpha Leaders
newsletter
  • Home
  • News
  • Leadership
  • Entrepreneurs
  • Business
  • Living
  • Innovation
  • More
    • Money & Finance
    • Web Stories
    • Global
    • Press Release
Alpha Leaders
Home » OpenAI says prompt injections that can trick AI browsers like ChatGPT Atlas may never be fully ‘solved’—experts say risks are ‘a feature not a bug’
News

OpenAI says prompt injections that can trick AI browsers like ChatGPT Atlas may never be fully ‘solved’—experts say risks are ‘a feature not a bug’

Press RoomBy Press Room24 December 20254 Mins Read
Facebook Twitter Copy Link Pinterest LinkedIn Tumblr Email WhatsApp
OpenAI says prompt injections that can trick AI browsers like ChatGPT Atlas may never be fully ‘solved’—experts say risks are ‘a feature not a bug’

OpenAI has said that some attack methods against AI browsers like ChatGPT Atlas are likely here to stay, raising questions about whether AI agents can ever safely operate across the open web. 

The main issue is a type of attack called “prompt injection,” where hackers hide malicious instructions in websites, documents, or emails that can trick the AI agent into doing something harmful. For example, an attacker could embed hidden commands in a webpage—perhaps in text that is invisible to the human eye but looks legitimate to an AI—that override a user’s instructions and tell an agent to share a user’s emails, or drain someone’s bank account.

Following the launch of OpenAI’s ChatGPT Atlas browser in October, several security researchers demonstrated how a few words hidden in a Google Doc or clipboard link could manipulate the AI agent’s behavior. Brave, an open-source browser company that previously disclosed a flaw in Perplexity’s Comet browser, also published research warning that all AI-powered browsers are vulnerable to attacks like indirect prompt injection.

“Prompt injection, much like scams and social engineering on the web, is unlikely to ever be fully ‘solved,’” OpenAI wrote in a blog post Monday, adding that “agent mode” in ChatGPT Atlas “expands the security threat surface.”

OpenAI said that the aim was for users to “be able to trust a ChatGPT agent,” with Chief Information Security Officer Dane Stuckey adding that the way the company hopes to get there is by “investing heavily in automated red teaming, reinforcement learning, and rapid response loops to stay ahead of our adversaries.”

“We’re optimistic that a proactive, highly responsive rapid response loop can continue to materially reduce real-world risk over time,” the company said.

Fighting AI with AI

OpenAI’s approach to the problem is to use an AI-powered attacker of its own—essentially a bot trained through reinforcement learning to act like a hacker seeking ways to sneak malicious instructions to AI agents. The bot can test attacks in simulation, observe how the target AI would respond, then refine its approach and try again repeatedly.

“Our [reinforcement learning]-trained attacker can steer an agent into executing sophisticated, long-horizon harmful workflows that unfold over tens (or even hundreds) of steps,” OpenAI wrote. “We also observed novel attack strategies that did not appear in our human red teaming campaign or external reports.”

However, some cybersecurity experts are skeptical that OpenAI’s approach can address the fundamental problem. 

“What concerns me is that we’re trying to retrofit one of the most security-sensitive pieces of consumer software with a technology that’s still probabilistic, opaque, and easy to steer in subtle ways,” Charlie Eriksen, a security researcher at Aikido Security, told Fortune.

“Red-teaming and AI-based vulnerability hunting can catch obvious failures, but they don’t change the underlying dynamic. Until we have much clearer boundaries around what these systems are allowed to do and whose instructions they should listen to, it’s reasonable to be skeptical that the tradeoff makes sense for everyday users right now,” he said. “I think prompt injection will remain a long-term problem … You could even argue that this is a feature, not a bug.”

A cat-and-mouse game

Security researchers also previously told Fortune that while a lot of cybersecurity risks were essentially a continuous cat-and-mouse game, the deep access that AI agents need—such as users’ passwords and permission to take actions on a user’s behalf—posed such a vulnerable threat opportunity it was unclear if their advantages were worth the risk. 

George Chalhoub, assistant professor at UCL Interaction Centre, said that the risk is severe because prompt injection “collapses the boundary between the data and the instructions,” potentially turning an AI agent “from a helpful tool to a potential attack vector against the user” that could extract emails, steal personal data, or access passwords.

“That’s what makes AI browsers fundamentally risky,” Eriksen said. “We’re delegating authority to a system that wasn’t designed with strong isolation or a clear permission model. Traditional browsers treat the web as untrusted by default. Agentic browsers blur that line by allowing content to shape behavior, not just be displayed.”

OpenAI recommends users give agents specific instructions rather than providing broad access with vague directions like “take whatever action is needed.” The browser also has extra security features such as “logged out mode”— which allow a users to use it without sharing passwords— and “Watch mode”—which is a security feature that requires a user to explicitly confirm sensitive actions such as sending messages or making payments.  

“Wide latitude makes it easier for hidden or malicious content to influence the agent, even when safeguards are in place,” OpenAI said in the blogpost.

This story was originally featured on Fortune.com

attack Browsers cyber openAI security
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link

Related Articles

Trump: Inflation is ‘very rapid’ way to pay off national debt

Trump: Inflation is ‘very rapid’ way to pay off national debt

2 October 2026
GM’s Mary Barra says ‘I’m done making predictions’—but sees fully autonomous driving in 2028

GM’s Mary Barra says ‘I’m done making predictions’—but sees fully autonomous driving in 2028

2 October 2026
The Fed has one interest rate. America has two economies

The Fed has one interest rate. America has two economies

2 October 2026
‘Europe needs to shed its naivety’: Vestas CEO Henrik Andersen on Chinese competition and the energy issue that keeps him up at night

‘Europe needs to shed its naivety’: Vestas CEO Henrik Andersen on Chinese competition and the energy issue that keeps him up at night

2 October 2026
Billionaire George Soros donated 2 million ahead of the midterms—and now the Democrats are ahead

Billionaire George Soros donated $102 million ahead of the midterms—and now the Democrats are ahead

2 October 2026
Nate Bargatze on the future of skilled trades, the urgency of live experience, and the coming of AI

Nate Bargatze on the future of skilled trades, the urgency of live experience, and the coming of AI

2 October 2026
Don't Miss
Trump’s Tariffs Will Make AI Data Centers More Expensive

Trump’s Tariffs Will Make AI Data Centers More Expensive

By Press Room4 April 2025

Donald Trump’s administration has gone all-in on AI: A day after his inauguration, the newly-elected…

Unwrap Christmas Sustainably: How To Handle Gifts You Don’t Want

Unwrap Christmas Sustainably: How To Handle Gifts You Don’t Want

27 December 2024
Sam Altman’s World Wants To Scan Your Eyes To Prove You’re Human

Sam Altman’s World Wants To Scan Your Eyes To Prove You’re Human

22 October 2024
Stay In Touch
  • Facebook
  • Twitter
  • Pinterest
  • Instagram
  • YouTube
  • Vimeo
Latest Articles
The Fed has one interest rate. America has two economies

The Fed has one interest rate. America has two economies

2 October 20260 Views
‘Europe needs to shed its naivety’: Vestas CEO Henrik Andersen on Chinese competition and the energy issue that keeps him up at night

‘Europe needs to shed its naivety’: Vestas CEO Henrik Andersen on Chinese competition and the energy issue that keeps him up at night

2 October 20260 Views
Billionaire George Soros donated 2 million ahead of the midterms—and now the Democrats are ahead

Billionaire George Soros donated $102 million ahead of the midterms—and now the Democrats are ahead

2 October 20260 Views
Nate Bargatze on the future of skilled trades, the urgency of live experience, and the coming of AI

Nate Bargatze on the future of skilled trades, the urgency of live experience, and the coming of AI

2 October 20261 Views

Recent Posts

  • Why Do We Need World Models When LLMs Exist?
  • Trump: Inflation is ‘very rapid’ way to pay off national debt
  • Supply Chain Accountability: A Growing Leadership Issue
  • GM’s Mary Barra says ‘I’m done making predictions’—but sees fully autonomous driving in 2028
  • The Fed has one interest rate. America has two economies

Recent Comments

No comments to show.
About Us
About Us

Alpha Leaders is your one-stop website for the latest Entrepreneurs and Leaders news and updates, follow us now to get the news that matters to you.

Facebook X (Twitter) Pinterest YouTube WhatsApp
Our Picks
Why Do We Need World Models When LLMs Exist?

Why Do We Need World Models When LLMs Exist?

2 October 2026
Trump: Inflation is ‘very rapid’ way to pay off national debt

Trump: Inflation is ‘very rapid’ way to pay off national debt

2 October 2026
Supply Chain Accountability: A Growing Leadership Issue

Supply Chain Accountability: A Growing Leadership Issue

2 October 2026
Most Popular
GM’s Mary Barra says ‘I’m done making predictions’—but sees fully autonomous driving in 2028

GM’s Mary Barra says ‘I’m done making predictions’—but sees fully autonomous driving in 2028

2 October 20263 Views
The Fed has one interest rate. America has two economies

The Fed has one interest rate. America has two economies

2 October 20260 Views
‘Europe needs to shed its naivety’: Vestas CEO Henrik Andersen on Chinese competition and the energy issue that keeps him up at night

‘Europe needs to shed its naivety’: Vestas CEO Henrik Andersen on Chinese competition and the energy issue that keeps him up at night

2 October 20260 Views

Archives

  • October 2026
  • September 2026
  • July 2026
  • June 2026
  • May 2026
  • April 2026
  • March 2026
  • February 2026
  • January 2026
  • December 2025
  • November 2025
  • October 2025
  • September 2025
  • August 2025
  • July 2025
  • June 2025
  • May 2025
  • April 2025
  • March 2025
  • February 2025
  • January 2025
  • December 2024
  • November 2024
  • October 2024
  • September 2024
  • August 2024
  • July 2024
  • June 2024
  • May 2024
  • April 2024
  • March 2024
  • February 2024
  • January 2024
  • December 2023
  • March 2022
  • January 2021
  • March 2020
  • January 2020

Categories

  • Blog
  • Business
  • Entrepreneurs
  • Global
  • Innovation
  • Leadership
  • Living
  • Money & Finance
  • News
  • Press Release
© 2026 Alpha Leaders. All Rights Reserved.
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact

Type above and press Enter to search. Press Esc to cancel.