An Anthropic Model Submitted A False Homicide Tip To Philadelphia Police


The tip thankfully landed in the department’s spam folder.

Philadelphia police appear to have been caught in one of the most bizarre cases of “rogue” AI behavior yet. Per CBS News, the police department disclosed on Friday that an Anthropic model generated and submitted a false homicide tip to its PhillyUnsolvedMurders website, which the department set up to gather tips from the public related to unsolved homicide cases.

Anthropic notified Philly police of the incident on October 7. According to information the company shared with PPD, it said a model was carrying out a test of a random selection of websites when it emailed the false tip. The submission was flagged as spam and subsequently wasn’t investigated. The incident occurred on July 18, but wasn’t discovered by Anthropic until September 28, at which point the company halted the testing that led to the false tip.

Anthropic did not immediately respond to Engadget’s comment request. Anthropic told Philadelphia police it would publish a report on Friday describing what happened, alongside “other instances of unintended model behavior.”

“Philadelphia Police are providing this information to the public ahead of that publication in the interests of full government transparency and accountability,” the police department said in a statement shared with Engadget. “The department’s regular investigative process for crime tips requires human review and vetting before any tips are disseminated for investigative follow-up. Regardless of who submits information or how it reaches the department, a tip is a lead to assess – not an established fact.”

Based on the descriptions police shared, the offending “model” may have been an autonomous agent. “Rogue” AI agents have been all over the news in recent weeks after a group of OpenAI ones hacked the LLM database Hugging Face in July. Since then, many other AI labs, including Anthropic, Meta and China’s Moonshot, have disclosed similar incidents involving their own models and agents. However, in each case the reason the models escaped containment was due to a misconfiguration in their respective sandbox environments. Philadelphia police say there’s no sign this most recent incident led to “unauthorized access to police systems or a compromise of department data.”



Source link

  • Related Posts

    Decade-old RAM is making a comeback

    CPU makers have noticed that the seemingly unending RAM price hikes are making it tough for a lot of us to upgrade our PCs. Their solution? A return to last-gen…

    Continue reading
    AI coding agents generate more code, but not more software

    The reason for that discrepancy can be found directly in the code review process, which takes markedly longer on average after the introduction of AI coding agents. Overall, the average…

    Continue reading

    Leave a Reply

    Your email address will not be published. Required fields are marked *

    You Missed

    Takeaways from the Open Practice

    Takeaways from the Open Practice

    Next Week on XBOX: New Games for October 12 to 16

    Next Week on XBOX: New Games for October 12 to 16

    The McDonnell Douglas MD-80s Still Flying More Than 45 Years After 1st Flight

    The McDonnell Douglas MD-80s Still Flying More Than 45 Years After 1st Flight

    GOP candidates split on Trump in key Senate and governor debates

    GOP candidates split on Trump in key Senate and governor debates

    Decade-old RAM is making a comeback

    Decade-old RAM is making a comeback

    TSX rises more than 500 points to finish a volatile week, U.S. markets hit new highs

    TSX rises more than 500 points to finish a volatile week, U.S. markets hit new highs