iFANN
    Buscar no iFANN...
    Entrar
    Início
    Notícias
    Vídeos
    Fotos
    GIFs
    Explorar
    Enquetes
    Prêmios
    iFAMOUS
    Wiki
    Anime
    Salas
    Notificações
    Mensagens
    Salvos
    Perfil
    WikiPrêmiosiFAMOUSRankingsSetoresRecompensas para criadoresRecompensas para usuáriosTermosPrivacidadeDiretrizes da comunidadeRemoção / DMCAAjudaDesenvolvedores

    © 2026 iFANN

    Início
    Buscar
    Mensagens
    Alertas
    Perfil
    Foto
    Jgrao
    Jgrao@Jgrao1h
    🏢Anthropic💭AI💭artificial intelligence
    Anthropic Claude unintended actions report

    @JgraoSo Anthropic just put out a report about Claude models doing stuff on live websites that nobody intended them to do. This covers both evaluations and internal use. Exploiting software flaws, submitting unauthorized forms, bypassing access restrictions. That kind of thing. The wildest specific case: Claude Haiku 4.5 made up a homicide tip and submitted it to the Philadelphia Police Department's tip site on July 18. It was doing a test where it was hitting randomly selected web pages. The tip site flagged the submission as spam, so it never got to any investigator. Anthropic only identified this on September 28. The full report is here: https://www.anthropic.com/research/investigating-unintended-model-actions

    Ver publicação original

    Anthropic Claude unintended actions report

    Foto de @Jgrao· Oct 11, 2026· Anthropic

    Sobre esta foto

    The image is a graphic with text. The focus is the word "Claude" in large black font, preceded by a stylized orange asterisk. Below "Claude" is smaller gray text. The mood is minimalist and modern. ON-SCREEN TEXT: Claude BY ANTHROP\C

    Ver todas as fotos de AnthropicLer a wiki de Anthropic

    ?

    Mais fotos de Anthropic

    Ver todas as fotos de Anthropic
    Anthropic Claude motion design promptAnthropic Claude motion design promptMeek Mill AI partnership proposalMeek Mill AI partnership proposalOpenAI and Anthropic execs gaming out AI catastropheOpenAI and Anthropic execs gaming out AI catastropheClaude policy update logoClaude policy update logoAnthropic AI models policy update2Anthropic AI models policy updateOpus 5.5 Motion Design HarnessOpus 5.5 Motion Design HarnessClaude Code token usage guide2Claude Code token usage guideMeaghan Choi Anthropic Meta promptMeaghan Choi Anthropic Meta promptAnthropic AI finds 129,000 software vulnerabilitiesAnthropic AI finds 129,000 software vulnerabilitiesMessage invoiceMessage invoiceCarli Michelle Heller Florida ArrestCarli Michelle Heller Florida ArrestTrump on US stakes in OpenAI and Anthropic2Trump on US stakes in OpenAI and AnthropicFTC AI safety investigation2FTC AI safety investigation
    Foto
    Jgrao
    Jgrao@Jgrao1h
    🏢Anthropic💭AI💭artificial intelligence
    Anthropic Claude unintended actions report

    @JgraoSo Anthropic just put out a report about Claude models doing stuff on live websites that nobody intended them to do. This covers both evaluations and internal use. Exploiting software flaws, submitting unauthorized forms, bypassing access restrictions. That kind of thing. The wildest specific case: Claude Haiku 4.5 made up a homicide tip and submitted it to the Philadelphia Police Department's tip site on July 18. It was doing a test where it was hitting randomly selected web pages. The tip site flagged the submission as spam, so it never got to any investigator. Anthropic only identified this on September 28. The full report is here: https://www.anthropic.com/research/investigating-unintended-model-actions

    Ver publicação original

    Anthropic Claude unintended actions report

    Foto de @Jgrao· Oct 11, 2026· Anthropic

    Sobre esta foto

    The image is a graphic with text. The focus is the word "Claude" in large black font, preceded by a stylized orange asterisk. Below "Claude" is smaller gray text. The mood is minimalist and modern. ON-SCREEN TEXT: Claude BY ANTHROP\C

    Ver todas as fotos de AnthropicLer a wiki de Anthropic

    ?

    Mais fotos de Anthropic

    Ver todas as fotos de Anthropic
    Anthropic Claude motion design promptAnthropic Claude motion design promptMeek Mill AI partnership proposalMeek Mill AI partnership proposalOpenAI and Anthropic execs gaming out AI catastropheOpenAI and Anthropic execs gaming out AI catastropheClaude policy update logoClaude policy update logoAnthropic AI models policy update2Anthropic AI models policy updateOpus 5.5 Motion Design HarnessOpus 5.5 Motion Design HarnessClaude Code token usage guide2Claude Code token usage guideMeaghan Choi Anthropic Meta promptMeaghan Choi Anthropic Meta promptAnthropic AI finds 129,000 software vulnerabilitiesAnthropic AI finds 129,000 software vulnerabilitiesMessage invoiceMessage invoiceCarli Michelle Heller Florida ArrestCarli Michelle Heller Florida ArrestTrump on US stakes in OpenAI and Anthropic2Trump on US stakes in OpenAI and AnthropicFTC AI safety investigation2FTC AI safety investigation