Close Menu
    What's Hot
    Technology & Innovation

    Sarai AI Now Handles Multi-Room Hotel Bookings Through Natural Conversation

    By Art RyanJuly 22, 20260

    Hotel bookings are getting less like filling out a form and more like sending a…

    DerbySoft and Travelodge Partner to Make Corporate Hotel Booking Less Messy

    July 22, 2026

    MGX-Led Consortium Completes $40 Billion AI Data Center Deal

    July 22, 2026

    e& UAE and Core42 Want to Make Sovereign AI Compute Easier to Access

    July 22, 2026
    Facebook X (Twitter) Instagram
    Facebook X (Twitter) Instagram
    Breaking AI News
    Wednesday, July 22
    • Home
    • Events
    • Videos
      • Machine Can Think Summit 2026
      • Step Dubai Conference 2026
    • Technology & Innovation

      Sarai AI Now Handles Multi-Room Hotel Bookings Through Natural Conversation

      July 22, 2026

      DerbySoft and Travelodge Partner to Make Corporate Hotel Booking Less Messy

      July 22, 2026

      MGX-Led Consortium Completes $40 Billion AI Data Center Deal

      July 22, 2026

      e& UAE and Core42 Want to Make Sovereign AI Compute Easier to Access

      July 22, 2026

      TikTok Says Better AI Marketing Starts With Cultural Relevance, Not More Content

      July 22, 2026
    • Business & Marketing

      MGX-Led Consortium Completes $40 Billion AI Data Center Deal

      July 22, 2026

      e& UAE and Core42 Want to Make Sovereign AI Compute Easier to Access

      July 22, 2026

      Google Frozen V2 Could Make Gemini AI Cheaper to Run

      July 22, 2026

      Billtrust Brings AI Assistants Into Accounts Receivable Work

      July 22, 2026

      UAE Firms Step Up AI Investments as Customer Expectations Move Faster

      July 21, 2026
    • Industry Applications

      Billtrust Brings AI Assistants Into Accounts Receivable Work

      July 22, 2026

      Naver Cloud and HD Korea Shipbuilding Push AI Deeper Into Shipyards

      July 21, 2026

      Julphar Taps IBM and SAP for Major AI-Ready Digital Overhaul

      July 21, 2026

      DIFC Gets Its First AI-Native Asset Manager as Dubai Pushes Finance Into the AI Era

      July 21, 2026

      Ottawa Backs AI Farming Tools as Climate Pressure Hits Canadian Crops

      July 20, 2026
    • Trends & Insights

      UAE Firms Step Up AI Investments as Customer Expectations Move Faster

      July 21, 2026

      NVIDIA Vera Rubin Pushes AI Compute Toward a New Metric: Intelligence per Dollar

      July 20, 2026

      China’s Moonshot AI Kimi K3 Narrows Gap With US Rivals

      July 19, 2026

      Claude Honeycomb Leak Sparks Fresh Talk Around Anthropic’s Next Big AI Model

      July 18, 2026

      Google Delays Gemini 3.5 Pro as Coding Performance Misses the Mark

      July 18, 2026
    • AI in Travel

      DerbySoft and Travelodge Partner to Make Corporate Hotel Booking Less Messy

      July 22, 2026

      Dubai AI World Series Reveals 26 Finalists as Enterprise AI Moves Past the Pilot Stage

      July 21, 2026

      Flight Centre Turns to AI Agents as Online Travel Booking Becomes a Bigger Fight

      July 20, 2026

      IBS Group Launches Naviq as AI-First Travel Technology Company

      July 14, 2026

      Travel Compositor Brings AskIA to the Front End of AI Trips

      July 14, 2026
    Breaking AI News
    Home » OpenAI Rolls Out Security Update for ChatGPT Atlas to Block Prompt Injection Attacks
    Technology & Innovation

    OpenAI Rolls Out Security Update for ChatGPT Atlas to Block Prompt Injection Attacks

    Art RyanBy Art RyanDecember 24, 2025No Comments4 Mins Read
    Facebook Twitter Pinterest LinkedIn Tumblr Email
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Key Highlights:

    • OpenAI rolled out a major security update for ChatGPT Atlas to prevent prompt injection attacks.
    • The update includes an adversarially trained browser-agent model and system-level security improvements.
    • Atlas can now detect and flag malicious instructions that try to manipulate the AI agent’s behavior.

    It’s not hidden anymore that AI browsers are slowly making their way into the market. Earlier this year, AI companies like Perplexity and OpenAI launched Comet AI browser and ChatGPT Atlas, respectively. The idea behind AI browser is quite fascinating but is it reliable enough to make users transition to these options for everyday web browsing is still a big question.

    Here’s how OpenAI is blocking prompt injection attacks in ChatGPT Atlas

    That’s especially true when the browser market is alone dominated by Google Chrome, which is already making strides by adding AI features. But, for AI-powered browsers that alone isn’t problem, privacy and security risks associated with them are equally a major roadblock. Speaking off which, OpenAI yesterday detailed that it has rolled out a major security update for ChatGPT Atlas, after a report from earlier this month ranked it as the worst browser to exist.

    In the latest security update, OpenAI has increased security against of the most persistent risks related to AI agents. Here I’m talking about prompt injection attacks. If you’ve used ChatGPT Atlas’s agent mode, you must be aware that it is designed to work directly inside a user’s browser. Meaning, it can do everything like a human would do, like opening webpages, clicking links, typing text, and completing workflows.

    While these are great in terms of making your browsing session easy and seamless, OpenAI admits that it makes the browser more appealing for attackers out there. They can manipulate agent behaviour through prompt injection. For those unaware, it’s a deceptive technique that works around embedding hidden or misleading instructions inside content that an AI agent processes, such as emails, documents, or webpages.

    Internal tests with its in-house AI attacker

    OpenAI says it has been working on curbing this threat long before ChatGPT Atlas launched publicly. The company confirmed that the recently released security update includes a newly adversarial trained browser-agent model, alongside robust system-level security. Per the announcement, these measures were taken after OpenAI internally discovered a new class of prompt injection.

    The company used automated red teaming approach. As part of this approach, the company developed an internal AI attacker trained using reinforcement learning rather than human testers. The interesting part is that the said internal AI attacker continuously searches for ways to hack into Atlas by attempting real-world, multistep attacks against the agent.

    According to OpenAI results are promising because its reinforcement-learning attacker can discover long-horizon exploits that unfold over dozens or even hundreds of steps. The internal AI attacker learns from its own successes and failures, and updates its strategies over time, much like a human attacker would do. This gives the company an opportunity to learn about such loopholes internally and develop fixes for them before they even reach to the masses.

    Signals for Reinforcement Learning
    Image credit: OpenAI

    An example of how the attackers exploit AI agents

    One example shared by OpenAI highlights how subtle these attacks can be. In the demonstration, a malicious email planted in a user’s inbox contained hidden instructions telling the agent to send a resignation email. Later, when the user asked Atlas to draft an out-of-office reply, the agent encountered the injected instructions and followed them instead, resigning on the user’s behalf. After the latest update, Atlas now detects and flags this behavior as a prompt injection attempt.

    OpenAI has long admitted that prompt injection remains an open, long-term challenge for everyone out there. The company also advises users to limit logged-in access whenever possible. In addition, users are recommended to review confirmation prompts carefully to stay safe.

    Source: https://www.timesofai.com/
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Art Ryan

    Related Posts

    Sarai AI Now Handles Multi-Room Hotel Bookings Through Natural Conversation

    July 22, 2026

    DerbySoft and Travelodge Partner to Make Corporate Hotel Booking Less Messy

    July 22, 2026

    MGX-Led Consortium Completes $40 Billion AI Data Center Deal

    July 22, 2026

    Comments are closed.

    Latest News

    Sarai AI Now Handles Multi-Room Hotel Bookings Through Natural Conversation

    July 22, 2026

    DerbySoft and Travelodge Partner to Make Corporate Hotel Booking Less Messy

    July 22, 2026

    MGX-Led Consortium Completes $40 Billion AI Data Center Deal

    July 22, 2026

    e& UAE and Core42 Want to Make Sovereign AI Compute Easier to Access

    July 22, 2026
    Facebook X (Twitter) Pinterest Vimeo WhatsApp TikTok Instagram LinkedIn YouTube Spotify Reddit Snapchat Threads

    AI University

    • Global Universities
    • Universities in Africa
    • Universities in Asia
    • Universities in Europe
    • Universities in Latin America
    • Universities in Middle East
    • Universities in North America
    • Universities in Oceania

    AI Tools & Apps Directory

    • AI Productivity Tools
    • AI Coding Tools
    • AI Voice Tools
    • AI Video Tools
    • AI Image Generators
    • AI Writing Tools

    Info

    • Home
    • About Us
    • AI Organizations & Associations
    • Contact Us
    • Cookie Policy
    • Copyright Policy
    • Disclaimer
    • Editorial Policy
    • Terms and Conditions

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    © 2026 Breaking AI News.
    • Privacy Policy

    Type above and press Enter to search. Press Esc to cancel.

    Sign Up

    Want to stay ahead In Artificial Intelligence?

     Sign up now and get exclusive breaking AI news and special updates—FREE!