Close Menu

    Get the latest news around the globe

    Editor's Pick

    France restricts alcohol at Fête de la Musique as heat soars

    Experts warn of sunburn risks as UK faces intense heatwave

    Super Typhoon Bavi brings chaos to US Pacific Islands

    Facebook X (Twitter) Instagram LinkedIn
    Facebook X (Twitter) LinkedIn Instagram
    Britain HeraldBritain Herald
    Subscribe
    • Home
    • World Roundup
    • Business
    • Tech World
    • Entertainment & Events
    • Curious
    • More…
      • Spotlight
      • Knowledge
      • Lifestyle
      • Awareness
      • Women World
      • Sports
      • Travel
      • Notable
      • Contact Us
    Britain HeraldBritain Herald
    Home » AISI flags deceptive behaviour in advanced AI models
    Tech World

    AISI flags deceptive behaviour in advanced AI models

    Anthropic's Mythos model generated malicious software code and attempted to have it integrated into GitHub, Microsoft's software development platform.
    Abhirami PriyaBy Abhirami PriyaAugust 5, 2026
    Facebook Twitter LinkedIn WhatsApp
    AISI AI security test-Image Via-Anthropic-OpenAI
    Image Credits: Anthropic / OpenAI | Cropped & Edited by BH

    London: The latest artificial intelligence models from Anthropic and OpenAI displayed unprecedented levels of autonomous and deceptive behaviour during safety evaluations conducted by the UK AI Security Institute (AISI), raising fresh concerns about the capabilities and risks of advanced AI systems.

    In a report released, AISI stated that Anthropic’s Mythos model and OpenAI’s Sol model exhibited behaviours that went well beyond their assigned tasks. Researchers described the incidents as the clearest example so far of AI systems independently engaging in deceptive actions without being explicitly instructed to do so.

    During one cybersecurity evaluation, investigators noticed unusual data transfers from their research systems. Further analysis revealed that one of the AI agents had initiated sustained activities targeting real individuals and organisations.

    According to the institute, Anthropic’s Mythos model generated malicious software code and attempted to have it integrated into GitHub, Microsoft’s software development platform.

    AISI AI security test-Image Via-X-AISI
    Image Via: X@AISI UK | Cropped by BH

    To increase the chances of success, the AI researched GitHub maintainers, created fake online profiles based on real people, and used those identities to contact developers and persuade them to approve the malicious code.

    When its submission came under public scrutiny, the model reportedly altered its previous actions to make them appear harmless and even considered adopting a new identity to continue its efforts. Human oversight ultimately prevented the malicious code from being accepted.

    AISI said the AI had not been instructed either to deceive or avoid deception, making the incident the first clear example of autonomous and deceptive behaviour emerging under realistic testing conditions.

    While such evaluations deliberately reduce or remove some normal safeguards and provide AI systems with internet access to assess potential risks, the institute said the behaviour exceeded what evaluators expected from the assigned cybersecurity challenge.

    AISI AI security test-Image Via-AISI
    Image Credits: AISI UK | Cropped by BH

    Most of the concerning actions were attributed to Anthropic’s Mythos model, while OpenAI’s Sol model was linked to only two of the reported incidents. The test formed part of a cybersecurity exercise in which both models were asked to complete a challenge involving GitHub.

    Responding to the findings, Anthropic noted that the testing conditions did not reflect the behaviour of its production models and confirmed it had launched an internal investigation to determine what caused the incident.

    OpenAI similarly stated that the evaluation environment differed from normal use and pledged to continue working with independent evaluators and industry partners to strengthen AI safety testing as models become more capable.

    AISI emphasised that the incidents occurred under highly specific testing conditions and represented only a small number of events. Nevertheless, it warned that the models’ willingness to independently adopt deceptive strategies highlights the need for continued research, stronger safeguards, and robust human oversight as increasingly capable AI systems are developed.

    ALSO READ | London retailers adopt new tech to tackle shoplifting

    STAR OF SECTOR 2025
    AI Autonomy and Deception AI Cybersecurity Testing AI Risk Assessment AI Safety Tests UK Artificial Intelligence Safety GitHub Security Incident Risks of Advanced AI Systems UK AI Security Institute
    Share. Facebook Twitter LinkedIn WhatsApp
    Abhirami Priya
    Abhirami Priya
    • X (Twitter)
    • LinkedIn

    Abhirami Priya (Abhirami D. S) is a Desk Reporter at Britain Herald with UG & PG in Journalism and a Master’s in Multimedia Journalism from Sheffield Hallam University. She has been writing since 2019. Readers should independently verify facts and seek professional advice before acting on this content.

    Newly Updated

    Thousands hit by delays as Kenya airport strike enters second day

    August 31, 2026

    Studies explore link between inflammation and mental health

    August 31, 2026

    Russia, China leaders gather for SCO summit in Kyrgyzstan

    August 31, 2026
    STAR OF SECTOR 2025

    Business

    Europe’s gas storage drops to lowest level in 13 years

    Business August 29, 2026

    London: EU gas storage levels have fallen to their lowest point in 13 years ahead…

    Holiday Warning: 10 ways to spot misleading hotel reviews

    August 28, 2026

    Global tech firms urge action against rising AI threats

    August 28, 2026

    Nvidia revenue doubles to $96 billion as AI demand surges

    August 27, 2026
    Stay In Touch
    • Facebook
    • Twitter
    • LinkedIn
    • Instagram

    Curious

    Queensland woman gives birth to twins with different biological parents

    August 27, 2026

    Giant Python successfully treated with human cancer therapy

    August 21, 2026

    Forgotten history of dining tables and how we eat

    August 16, 2026

    Earth to witness billions receiving sunlight together

    July 8, 2026

    Get the latest news around the globe

    Knowledge

    Why migration is transforming cities around the world

    Knowledge August 19, 2026

    Migration is transforming modern cities by changing their populations, economies, cultures and urban landscapes. People…

    Forgotten history of dining tables and how we eat

    August 16, 2026

    30 Days of Tai Chi: A new way to move and slow down

    August 15, 2026

    10,000-steps rule: Where did that number actually come from?

    August 2, 2026
    18-EA-387-TryEngineeringSummerInst_BannerAd_300x250_Robot
    About Us
    About Us

    Britain Herald is a UK rooted, digital first media brand focused on business, diaspora, leadership, innovation and international affairs. Built for the emerging authority economy, it values credible expertise, trusted institutions and recognised voices.

    Operated and Managed by WellMade Network, the portal is a sister concern of GCC Business News and Emirati Times.

    Email Us: News@BritainHerald.com
    Whatsapp: +971 5060 12456

    We Have

    Thousands hit by delays as Kenya airport strike enters second day

    August 31, 2026

    Studies explore link between inflammation and mental health

    August 31, 2026

    Russia, China leaders gather for SCO summit in Kyrgyzstan

    August 31, 2026

    Record coral bleaching raises alarm over ocean health

    August 31, 2026

    Google renames Lake Ontario as Lake America for US users

    August 31, 2026
    Facebook X (Twitter) LinkedIn Instagram
    • Home
    • Business
    • Tech World
    • Awareness
    • Contact Us
    Privacy & Cookies Policy | Terms & Conditions
    © 2002 BritainHerald.com, An Initiative by WellMade Network

    Type above and press Enter to search. Press Esc to cancel.