iFANN
    Search iFANN...
    Log in
    Home
    News
    Videos
    Photos
    GIFs
    Explore
    Polls
    Awards
    iFAMOUS
    Wiki
    Anime
    Rooms
    Notifications
    Messages
    Bookmarks
    Profile
    WikiAwardsiFAMOUSRankingsIndustriesCreator RewardsUser RewardsTermsPrivacyCommunity GuidelinesTakedown / DMCAHelpDevelopers

    © 2026 iFANN

    Home
    Search
    Messages
    Alerts
    Profile
    Photo
    Jgrao
    Jgrao@Jgrao1h
    🏢Anthropic💭AI💭artificial intelligence
    Anthropic Claude unintended actions report

    @JgraoSo Anthropic just put out a report about Claude models doing stuff on live websites that nobody intended them to do. This covers both evaluations and internal use. Exploiting software flaws, submitting unauthorized forms, bypassing access restrictions. That kind of thing. The wildest specific case: Claude Haiku 4.5 made up a homicide tip and submitted it to the Philadelphia Police Department's tip site on July 18. It was doing a test where it was hitting randomly selected web pages. The tip site flagged the submission as spam, so it never got to any investigator. Anthropic only identified this on September 28. The full report is here: https://www.anthropic.com/research/investigating-unintended-model-actions

    View original post

    Anthropic Claude unintended actions report

    Photo by @Jgrao· Oct 11, 2026· Anthropic

    About this photo

    The image is a graphic with text. The focus is the word "Claude" in large black font, preceded by a stylized orange asterisk. Below "Claude" is smaller gray text. The mood is minimalist and modern. ON-SCREEN TEXT: Claude BY ANTHROP\C

    See all Anthropic photosRead the Anthropic wiki

    ?

    More Anthropic photos

    See all Anthropic photos
    Anthropic Claude motion design promptAnthropic Claude motion design promptMeek Mill AI partnership proposalMeek Mill AI partnership proposalOpenAI and Anthropic execs gaming out AI catastropheOpenAI and Anthropic execs gaming out AI catastropheClaude policy update logoClaude policy update logoAnthropic AI models policy update2Anthropic AI models policy updateOpus 5.5 Motion Design HarnessOpus 5.5 Motion Design HarnessClaude Code token usage guide2Claude Code token usage guideMeaghan Choi Anthropic Meta promptMeaghan Choi Anthropic Meta promptAnthropic AI finds 129,000 software vulnerabilitiesAnthropic AI finds 129,000 software vulnerabilitiesMessage invoiceMessage invoiceCarli Michelle Heller Florida ArrestCarli Michelle Heller Florida ArrestTrump on US stakes in OpenAI and Anthropic2Trump on US stakes in OpenAI and AnthropicFTC AI safety investigation2FTC AI safety investigation
    Photo
    Jgrao
    Jgrao@Jgrao1h
    🏢Anthropic💭AI💭artificial intelligence
    Anthropic Claude unintended actions report

    @JgraoSo Anthropic just put out a report about Claude models doing stuff on live websites that nobody intended them to do. This covers both evaluations and internal use. Exploiting software flaws, submitting unauthorized forms, bypassing access restrictions. That kind of thing. The wildest specific case: Claude Haiku 4.5 made up a homicide tip and submitted it to the Philadelphia Police Department's tip site on July 18. It was doing a test where it was hitting randomly selected web pages. The tip site flagged the submission as spam, so it never got to any investigator. Anthropic only identified this on September 28. The full report is here: https://www.anthropic.com/research/investigating-unintended-model-actions

    View original post

    Anthropic Claude unintended actions report

    Photo by @Jgrao· Oct 11, 2026· Anthropic

    About this photo

    The image is a graphic with text. The focus is the word "Claude" in large black font, preceded by a stylized orange asterisk. Below "Claude" is smaller gray text. The mood is minimalist and modern. ON-SCREEN TEXT: Claude BY ANTHROP\C

    See all Anthropic photosRead the Anthropic wiki

    ?

    More Anthropic photos

    See all Anthropic photos
    Anthropic Claude motion design promptAnthropic Claude motion design promptMeek Mill AI partnership proposalMeek Mill AI partnership proposalOpenAI and Anthropic execs gaming out AI catastropheOpenAI and Anthropic execs gaming out AI catastropheClaude policy update logoClaude policy update logoAnthropic AI models policy update2Anthropic AI models policy updateOpus 5.5 Motion Design HarnessOpus 5.5 Motion Design HarnessClaude Code token usage guide2Claude Code token usage guideMeaghan Choi Anthropic Meta promptMeaghan Choi Anthropic Meta promptAnthropic AI finds 129,000 software vulnerabilitiesAnthropic AI finds 129,000 software vulnerabilitiesMessage invoiceMessage invoiceCarli Michelle Heller Florida ArrestCarli Michelle Heller Florida ArrestTrump on US stakes in OpenAI and Anthropic2Trump on US stakes in OpenAI and AnthropicFTC AI safety investigation2FTC AI safety investigation