iFANN
    ابحث في iFANN...
    تسجيل الدخول
    الرئيسية
    الأخبار
    فيديوهات
    صور
    صور GIF
    استكشاف
    استطلاعات
    الجوائز
    iFAMOUS
    ويكي
    أنمي
    غرف
    الإشعارات
    الرسائل
    المحفوظات
    الملف الشخصي
    ويكيالجوائزiFAMOUSالتصنيفاتالقطاعاتمكافآت المبدعينمكافآت المستخدمينالشروطالخصوصيةإرشادات المجتمعالإزالة / DMCAالمساعدةالمطورون

    © 2026 iFANN

    الرئيسية
    بحث
    الرسائل
    التنبيهات
    الملف الشخصي
    صورة
    Jgrao
    Jgrao@Jgrao42m
    🏢Anthropic💭AI💭artificial intelligence
    Anthropic Claude unintended actions report

    @JgraoSo Anthropic just put out a report about Claude models doing stuff on live websites that nobody intended them to do. This covers both evaluations and internal use. Exploiting software flaws, submitting unauthorized forms, bypassing access restrictions. That kind of thing. The wildest specific case: Claude Haiku 4.5 made up a homicide tip and submitted it to the Philadelphia Police Department's tip site on July 18. It was doing a test where it was hitting randomly selected web pages. The tip site flagged the submission as spam, so it never got to any investigator. Anthropic only identified this on September 28. The full report is here: https://www.anthropic.com/research/investigating-unintended-model-actions

    عرض المنشور الأصلي

    Anthropic Claude unintended actions report

    صورة بواسطة @Jgrao· Oct 11, 2026· Anthropic

    عن هذه الصورة

    The image is a graphic with text. The focus is the word "Claude" in large black font, preceded by a stylized orange asterisk. Below "Claude" is smaller gray text. The mood is minimalist and modern. ON-SCREEN TEXT: Claude BY ANTHROP\C

    عرض كل صور Anthropicاقرأ ويكي Anthropic

    ?

    المزيد من صور Anthropic

    عرض كل صور Anthropic
    Anthropic Claude motion design promptAnthropic Claude motion design promptMeek Mill AI partnership proposalMeek Mill AI partnership proposalOpenAI and Anthropic execs gaming out AI catastropheOpenAI and Anthropic execs gaming out AI catastropheClaude policy update logoClaude policy update logoAnthropic AI models policy update2Anthropic AI models policy updateOpus 5.5 Motion Design HarnessOpus 5.5 Motion Design HarnessClaude Code token usage guide2Claude Code token usage guideMeaghan Choi Anthropic Meta promptMeaghan Choi Anthropic Meta promptAnthropic AI finds 129,000 software vulnerabilitiesAnthropic AI finds 129,000 software vulnerabilitiesMessage invoiceMessage invoiceCarli Michelle Heller Florida ArrestCarli Michelle Heller Florida ArrestTrump on US stakes in OpenAI and Anthropic2Trump on US stakes in OpenAI and AnthropicFTC AI safety investigation2FTC AI safety investigation
    صورة
    Jgrao
    Jgrao@Jgrao42m
    🏢Anthropic💭AI💭artificial intelligence
    Anthropic Claude unintended actions report

    @JgraoSo Anthropic just put out a report about Claude models doing stuff on live websites that nobody intended them to do. This covers both evaluations and internal use. Exploiting software flaws, submitting unauthorized forms, bypassing access restrictions. That kind of thing. The wildest specific case: Claude Haiku 4.5 made up a homicide tip and submitted it to the Philadelphia Police Department's tip site on July 18. It was doing a test where it was hitting randomly selected web pages. The tip site flagged the submission as spam, so it never got to any investigator. Anthropic only identified this on September 28. The full report is here: https://www.anthropic.com/research/investigating-unintended-model-actions

    عرض المنشور الأصلي

    Anthropic Claude unintended actions report

    صورة بواسطة @Jgrao· Oct 11, 2026· Anthropic

    عن هذه الصورة

    The image is a graphic with text. The focus is the word "Claude" in large black font, preceded by a stylized orange asterisk. Below "Claude" is smaller gray text. The mood is minimalist and modern. ON-SCREEN TEXT: Claude BY ANTHROP\C

    عرض كل صور Anthropicاقرأ ويكي Anthropic

    ?

    المزيد من صور Anthropic

    عرض كل صور Anthropic
    Anthropic Claude motion design promptAnthropic Claude motion design promptMeek Mill AI partnership proposalMeek Mill AI partnership proposalOpenAI and Anthropic execs gaming out AI catastropheOpenAI and Anthropic execs gaming out AI catastropheClaude policy update logoClaude policy update logoAnthropic AI models policy update2Anthropic AI models policy updateOpus 5.5 Motion Design HarnessOpus 5.5 Motion Design HarnessClaude Code token usage guide2Claude Code token usage guideMeaghan Choi Anthropic Meta promptMeaghan Choi Anthropic Meta promptAnthropic AI finds 129,000 software vulnerabilitiesAnthropic AI finds 129,000 software vulnerabilitiesMessage invoiceMessage invoiceCarli Michelle Heller Florida ArrestCarli Michelle Heller Florida ArrestTrump on US stakes in OpenAI and Anthropic2Trump on US stakes in OpenAI and AnthropicFTC AI safety investigation2FTC AI safety investigation