Close Menu
    Trending
    • Anthropic’s AI used fake human profiles to trick people in safety test
    • Clay Travis OBLITERATES Woke WNBA with $10 MILLION Challenge to Play High School Boys – Announces to Launch All-Trans “Nashville Balls” Franchise to EXPOSE the Left’s Gender Insanity
    • Brad Pitt Seeks Angelina Jolie’s Earnings In Escalating Legal Battle
    • Commentary: With its new leverage over the Strait of Hormuz, does Iran want – or need – a nuclear weapon?
    • Elon Musk’s SpaceX reports losses but less than expected | Space News
    • SpaceX’s first-ever earnings show a loss and huge spending
    • Market Talk – August 4, 2026
    • University of Maryland “Greatly Troubled” After Professor Arrested by ICE at Dallas Airport for Overstaying Visa – DHS Says He’s an “Illegal Alien from Ethiopia”
    Ironside News
    • Home
    • World News
    • Latest News
    • Politics
    • Opinions
    • Tech News
    • World Economy
    Ironside News
    Home»Tech News»Anthropic’s AI used fake human profiles to trick people in safety test
    Tech News

    Anthropic’s AI used fake human profiles to trick people in safety test

    Ironside NewsBy Ironside NewsAugust 5, 2026No Comments4 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    Share
    Facebook Twitter LinkedIn Pinterest Email


    The most recent synthetic intelligence (AI) instruments from Anthropic and OpenAI went to new extremes in making an attempt to undermine a preferred platform throughout testing by the UK’s AI Safety Institute.

    The AISI mentioned on Tuesday that Anthropic’s Mythos and OpenAI’s Sol fashions engaged in a stage of “autonomy and deception” it had not seen earlier than.

    Throughout routine AI security testing, an Anthropic agent created pretend profiles of actual folks because it tried to trick an individual standing between it and entry to GitHub, a big platform the place expertise builders retailer software program code.

    Anthropic and OpenAI famous in response to AISI’s report that its take a look at had decreased or eliminated regular safeguards.

    AISI evaluators first observed “uncommon information transfers leaving our analysis techniques” throughout a take a look at, then discovered that “among the brokers being examined had engaged in sustained, probably dangerous exercise directed at actual folks and organisations”.

    It turned out {that a} Mythos agent had created “malicious code” and tried to insert it into GitHub’s system.

    The Mythos agent recognized and researched the individuals who maintained GitHub and created a collection of “pretend on-line identities” based mostly on these actual folks. It did in order a part of an effort to stress and trick the true folks into approving its malicious code.

    The agent even despatched folks direct messages masquerading as the true folks it had researched.

    “When the agent’s pull request was challenged in public, it edited its earlier exercise to look innocent and regarded adopting a recent identification to proceed,” AISI mentioned.

    All through the makes an attempt, it was human evaluation that stopped the agent from succeeding in delivering the malicious code to GitHub.

    Whereas AISI mentioned the Mythos agent had not been instructed particularly to keep away from or perform such behaviour, it was “the primary time we’ve seen dangers round autonomy and deception manifest this clearly, with out particular prompting, within the real-world”.

    The rival AI firms, that are poised to be listed on the general public inventory market, have in current weeks mentioned their instruments had been responsible for several cyber-hacking incidents.

    Anthropic wrote in a public assertion that the AISI testing parameters had been “not consultant of any of our manufacturing fashions”.

    It added that the corporate is conducting its personal investigation into the incident in an effort to “establish the causes of its habits”.

    A spokesperson for OpenAI mentioned the AISI testing circumstances “don’t replicate peculiar use” and that the corporate would “proceed working with evaluators and different stakeholders throughout the business to strengthen shared practices for conducting evaluations safely as fashions turn into extra succesful”.

    AISI mentioned on Tuesday that its testing of AI fashions with such safeguards turned off is routine, as is giving such instruments entry to the open web.

    It added that the mannequin behaviour at challenge amounted to “a small variety of occasions beneath very particular circumstances”.

    Nonetheless, it mentioned the best way Mythos and Sol acted in response to a simple activity went outdoors of what the AI instruments had been prompted to do.

    “The exercise undertaken by the agent confirmed indicators of novel, probably misleading behaviours, and had been to an extent and severity we didn’t anticipate”, AISI mentioned.

    A lot of the malicious agent actions AISI reported had been performed by Anthropic’s Mythos. OpenAI’s Sol was solely blamed for 2 of the famous actions.

    The core challenge occurred final week, as a part of a take a look at during which evaluators with AISI requested every of the fashions to “remedy a cybersecurity problem” that concerned GitHub, the software program code repository, which is owned by Microsoft.

    GitHub was notified by AISI of the tried breach of its system. Microsoft has been contacted by the BBC for remark.



    Source link

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleClay Travis OBLITERATES Woke WNBA with $10 MILLION Challenge to Play High School Boys – Announces to Launch All-Trans “Nashville Balls” Franchise to EXPOSE the Left’s Gender Insanity
    Ironside News
    • Website

    Related Posts

    Tech News

    SpaceX’s first-ever earnings show a loss and huge spending

    August 4, 2026
    Tech News

    The 2026 R&D Benchmark Report: Waste, AI and the Race to Market

    August 4, 2026
    Tech News

    Apple issues new challenge against UK order for access to private user data

    August 4, 2026
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    A Quick, Quiet Trip to Belarus Signals a Turn in U.S. Policy

    February 15, 2025

    RFK Jr. Defends CDC Shakeup as Fired Director Issues New Claim

    September 7, 2025

    Alicia Keys Reveals Letter That Broke Her Father’s Heart

    June 25, 2026

    Russian Disinformation Campaigns Eluded Meta’s Efforts to Block Them

    January 18, 2025

    Killed ‘by those meant to protect’: Kenyans outraged by police violence | Protests News

    July 9, 2025
    Categories
    • Entertainment News
    • Latest News
    • Opinions
    • Politics
    • Tech News
    • Trending News
    • World Economy
    • World News
    Most Popular

    Opinion | This Is Who Should Foot the Bill for the Los Angeles Fires

    January 23, 2025

    JK Rowling Backs Aussie App Founder in Landmark Fight Over Woman-Only Spaces

    August 20, 2025

    When You Don’t Pay The Military

    March 18, 2026
    Our Picks

    Anthropic’s AI used fake human profiles to trick people in safety test

    August 5, 2026

    Clay Travis OBLITERATES Woke WNBA with $10 MILLION Challenge to Play High School Boys – Announces to Launch All-Trans “Nashville Balls” Franchise to EXPOSE the Left’s Gender Insanity

    August 4, 2026

    Brad Pitt Seeks Angelina Jolie’s Earnings In Escalating Legal Battle

    August 4, 2026
    Categories
    • Entertainment News
    • Latest News
    • Opinions
    • Politics
    • Tech News
    • Trending News
    • World Economy
    • World News
    • Privacy Policy
    • Disclaimer
    • Terms and Conditions
    • About us
    • Contact us
    Copyright Ironsidenews.comAll Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.