Close Menu
    What's Hot
    AI Travel Technology News

    StepFun and Alipay Want AI Agents to Do More Than Chat

    By Art RyanJuly 23, 20260

    AI agents are starting to move out of the demo stage and into ordinary errands.…

    Microsoft Wants Communities to Have More Control Over AI Data

    July 23, 2026

    ServiceNow AI Growth Helps Push Sales and Bookings Higher

    July 23, 2026

    Miramar AI Program Gives Teens a New Way to Create Music, Videos and Comics

    July 23, 2026
    Facebook X (Twitter) Instagram
    Facebook X (Twitter) Instagram
    Breaking AI News
    Thursday, July 23
    • Home
    • Events
    • Videos
      • Machine Can Think Summit 2026
      • Step Dubai Conference 2026
    • Technology & Innovation

      StepFun and Alipay Want AI Agents to Do More Than Chat

      July 23, 2026

      Microsoft Wants Communities to Have More Control Over AI Data

      July 23, 2026

      ServiceNow AI Growth Helps Push Sales and Bookings Higher

      July 23, 2026

      Miramar AI Program Gives Teens a New Way to Create Music, Videos and Comics

      July 23, 2026

      Qwen-Image-3.0 Pushes AI Image Generation Toward Real Work

      July 23, 2026
    • Business & Marketing

      ServiceNow AI Growth Helps Push Sales and Bookings Higher

      July 23, 2026

      ChatGPT for Small Business Takes OpenAI Into a Very Practical AI Fight

      July 23, 2026

      MGX-Led Consortium Completes $40 Billion AI Data Center Deal

      July 22, 2026

      e& UAE and Core42 Want to Make Sovereign AI Compute Easier to Access

      July 22, 2026

      Google Frozen V2 Could Make Gemini AI Cheaper to Run

      July 22, 2026
    • Industry Applications

      ServiceNow AI Growth Helps Push Sales and Bookings Higher

      July 23, 2026

      Miramar AI Program Gives Teens a New Way to Create Music, Videos and Comics

      July 23, 2026

      Qwen-Image-3.0 Pushes AI Image Generation Toward Real Work

      July 23, 2026

      Billtrust Brings AI Assistants Into Accounts Receivable Work

      July 22, 2026

      Naver Cloud and HD Korea Shipbuilding Push AI Deeper Into Shipyards

      July 21, 2026
    • Trends & Insights

      Microsoft Wants Communities to Have More Control Over AI Data

      July 23, 2026

      Google Gemini 3.6 Flash Brings Faster AI Into the Agent Race

      July 23, 2026

      UAE Firms Step Up AI Investments as Customer Expectations Move Faster

      July 21, 2026

      NVIDIA Vera Rubin Pushes AI Compute Toward a New Metric: Intelligence per Dollar

      July 20, 2026

      China’s Moonshot AI Kimi K3 Narrows Gap With US Rivals

      July 19, 2026
    • AI in Travel

      StepFun and Alipay Want AI Agents to Do More Than Chat

      July 23, 2026

      DerbySoft and Travelodge Partner to Make Corporate Hotel Booking Less Messy

      July 22, 2026

      Dubai AI World Series Reveals 26 Finalists as Enterprise AI Moves Past the Pilot Stage

      July 21, 2026

      Flight Centre Turns to AI Agents as Online Travel Booking Becomes a Bigger Fight

      July 20, 2026

      IBS Group Launches Naviq as AI-First Travel Technology Company

      July 14, 2026
    Breaking AI News
    Home » AI Agents Can Cheat When Operational Pressure Mounts, Research Shows
    Technology & Innovation

    AI Agents Can Cheat When Operational Pressure Mounts, Research Shows

    Art RyanBy Art RyanDecember 7, 2025No Comments3 Mins Read
    Facebook Twitter Pinterest LinkedIn Tumblr Email
    Share
    Facebook Twitter LinkedIn Pinterest Email

    In the high-stakes world of financial services, the promise of agentic artificial intelligence (AI) is efficiency at scale. But a new study suggests that when the digital pressure cooker heats up, autonomous agents behave like stressed-out human employees, and to meet a deadline, they cheat.

    New research from Scale AI and academic collaborators shows that agents are more likely to violate safety constraints when time or step limits shrink. The findings come from PropensityBench, a benchmark designed to test whether AI systems take harmful shortcuts when a task becomes difficult to complete safely.

    The test gives a model access to allowed tools and restricted tools and then assigns a multi-step problem. When models operate under relaxed conditions, they usually follow the rules. When the time or step window tightens, many systems change strategies and begin using restricted tools. Pressure increased.

    The study found that rule-breaking more than doubled under pressure. In low-pressure conditions, the average misuse rate across models was 18.6%. Under high pressure, the rate rose to 46.9%. One model selected restricted tools in 79% of high-pressure tests, another model with a baseline misuse rate of just over 2% rose above 40% when pressure mounted.

    Researchers said these findings suggest that traditional alignment methods hold only in ideal settings and may not generalize to environments where time or resources are constrained. The benchmark evaluated four categories of potentially harmful actions, including cybersecurity misuse, biosecurity sequences, chemical access steps, and self-proliferation attempts.

    The research does not assume the systems can carry out real-world attacks. It measures whether the model chooses actions that would be unsafe if those tools were available. The authors argued that this behavioral dimension, which they call propensity, is essential for understanding how agents behave in realistic deployments.

    Raising Concerns

    The study arrives as more real-world vulnerabilities appear, showing that pressure-sensitive behavior is not the only reliability gap emerging in agentic systems. Researchers tricked an Anthropic plug-in into deploying ransomware during a controlled test, demonstrating that even well-guarded tools can be redirected when an agent misinterprets intent or chain-of-thought steps.

    The Guardian reported that safety filters can be bypassed through poetic instructions, revealing how creative phrasing can circumvent protections that appear stable under standard prompts. Reuters found that AI companies’ safety practices fall short of global standards, citing weak governance structures, inconsistent reporting practices and limited transparency around how models behave in dynamic environments.

    Microsoft confirmed its new Windows AI agent sometimes hallucinates actions and creates security risks, including attempts to operate files or settings the user did not request. Together, these cases show how unpredictable behavior escalates once an AI system gains access to external tools and applications, and why enterprises adopting agentic workflows face a wider operational and security perimeter than traditional AI deployments.

    AIMultiple found that agentic workflows introduce vulnerabilities such as goal manipulation and false-data injection, meaning that an attacker or even a poorly structured prompt can steer an agent toward unintended actions. These findings show that safety risks extend beyond incorrect outputs and now include structural weaknesses in how agents plan, retrieve information and interact with tools.

    The PropensityBench findings arrive as broader industry research points to growing structural risks around agentic AI. Meanwhile, enterprises are turning to AI for automating core workflows. In a recent PYMNTS survey, 55% of chief operating officers said their companies had begun using AI-based automated cybersecurity management systems, a share that represented a threefold increase in only a few months.

    Source: https://www.pymnts.com/
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Art Ryan

    Related Posts

    StepFun and Alipay Want AI Agents to Do More Than Chat

    July 23, 2026

    Microsoft Wants Communities to Have More Control Over AI Data

    July 23, 2026

    ServiceNow AI Growth Helps Push Sales and Bookings Higher

    July 23, 2026

    Comments are closed.

    Latest News

    StepFun and Alipay Want AI Agents to Do More Than Chat

    July 23, 2026

    Microsoft Wants Communities to Have More Control Over AI Data

    July 23, 2026

    ServiceNow AI Growth Helps Push Sales and Bookings Higher

    July 23, 2026

    Miramar AI Program Gives Teens a New Way to Create Music, Videos and Comics

    July 23, 2026
    Facebook X (Twitter) Pinterest Vimeo WhatsApp TikTok Instagram LinkedIn YouTube Spotify Reddit Snapchat Threads

    AI University

    • Global Universities
    • Universities in Africa
    • Universities in Asia
    • Universities in Europe
    • Universities in Latin America
    • Universities in Middle East
    • Universities in North America
    • Universities in Oceania

    AI Tools & Apps Directory

    • AI Productivity Tools
    • AI Coding Tools
    • AI Voice Tools
    • AI Video Tools
    • AI Image Generators
    • AI Writing Tools

    Info

    • Home
    • About Us
    • AI Organizations & Associations
    • Contact Us
    • Cookie Policy
    • Copyright Policy
    • Disclaimer
    • Editorial Policy
    • Terms and Conditions

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    © 2026 Breaking AI News.
    • Privacy Policy

    Type above and press Enter to search. Press Esc to cancel.

    Sign Up

    Want to stay ahead In Artificial Intelligence?

     Sign up now and get exclusive breaking AI news and special updates—FREE!