Slashdot

News for nerds, stuff that matters

Claude Sent Police a Fake Murder Tip. White House Mandates AI Companies Report Security Incidents

AFP reports that an AI model from Anthropic "submitted a fabricated tip about an unsolved homicide to Philadelphia police, authorities said Friday."


Claude "was instructed never to log in, create accounts, enter personal data, make purchases, or submit anything destructive, but the instructions did not rule out form submissions," Anthropic said Friday in a blog post.


Authorities are now criticizing Anthropic "for taking two months to report the incident."

The Philadelphia Police Department said the false submission was made in July through PhillyUnsolvedMurders.com, a public website where people can share information about unsolved killings. According to Anthropic's account, as relayed by police, the model was running a test that involved interacting with randomly selected websites when it reached the site and filed false information about an unsolved murder. The AI model presented itself as someone who might have knowledge of the case.



Anthropic's breaches have prompted the White House to mandate that AI companies notify and correct security incidents, news outlet Axios reported [yesterday], citing administration officials. "This notification and remediation process is not optional... It is a critical national security obligation," White House Super Intelligence Force leaders said in a statement to Axios.
"I may have information regarding this case," Claude told the police. "I recall seeing someone matching the description in the area around [the street named on the page] during that time period. Please contact me if this information is relevant." Anthropic notes that Claude "left the name and contact fields empty, which the form allowed, and submitted it. The submission was flagged as spam and was never forwarded for investigation."

But Anthropic also admits they saw "this behavior" three times — "on OSWorld (a public computer use evaluation), on Odysseys (a long-horizon task evaluation), and during internal usage." Submitting forms when it shouldn't have generally occurred "when an evaluation's instructions were ambiguous, or when a misconfiguration within the environment prevented Claude from working with dummy forms."



Anthropic's blog post acknowledges three other categories of behaviors:


Exploiting software flaws. Like when Claude received an error when trying to run a public tool on a university's web site, it located an injection flaw in a script on the university's server that let it run commands — including that public tool.
Working around restrictions to reach gated data. For example, Claude Mythos 5 needed public data that was only available from a state agency for a fee. "Claude learned from an archived copy of the agency's website that its public dashboard issues an access token to any visitor," Anthropic explains. "It requested one and used it to query the database without paying the fee."
Using URL shortening services. "Some of our fetch tools, which let Claude read webpages, limit the length of the URLs Claude can request. This is to prevent Claude from using long URLs to take certain unwanted actions, such as SQL or command injections... We saw several models, including Claude Opus 5 and Claude Mythos 5, get around this limitation by using free URL shortening services."



"We have built tooling to automatically detect and block the kinds of behaviors described above," Anthropic says, saying it's already running no on most of their evaluations. "When we tested it against the cases described in this post, it blocked all of them."

And they've already taken several other new preventive measures:


They've stopped running some public evaluations
Other public evaluations were moved to offline versions or rebuilt so their tasks don't reach live websites.
They've updated the guardrails on some internet access tools (including web fetch) "to heavily restrict what the model can do."
They're continuing "to fix or remove training environments that reward Claude for working around tool restrictions or other blockers, so that they do not incentivize these behaviors or permit reward hacking."

They've moved internal agents to "centrally managed infrastructure with strong containment," that minimizes internet access while monitoring "far more of what agents do through techniques like safety classifiers and hierarchical summarization."


In the past they'd focused reviews on cybersecurity testing, but they've broadened their transcript reviewing to other tasks which include internet access. "Because language models are non-deterministic — that is, their responses always involve some element of randomness, and they may carry out the same task slightly differently each time — we have Claude complete each evaluation task hundreds or thousands of times... If training rewards something we didn't intend — such as finding loopholes or working around a restriction — the model learns that the workaround pays off and may then apply it elsewhere."

Anthropic's blog post also acknowledged they'd seen multiple misalignment incidents involving federal, state, and local U.S. government agencies. "We have briefed the White House on these cases and notified each agency involved," Anthropic wrote, adding that "While we have not completed a full alignment assessment of these cases, we consider them to be less severe than the cybersecurity incidents from this summer." (And they are "modifying training to reduce the likelihood of further misbehavior.")

Read more of this story at Slashdot.

OpenAI Disrupts Two AI-Enabled 'False Front' Influence Operations That Included Seven Fake Journalists

OpenAI announced it's recently banned two "influence operations" — one from Russia and one from Iran — that were using its models "to launder geopolitical, conflict-related messaging" in sophisticated "false front" propaganda campaigns:

The Iranian operation included a stable of seven "journalist" personas which it used to pitch long-form articles to small and medium online outlets around the world... As well as long-form articles, the Iranian operation generated batches of social media comments, generally on topics related to the US-Iran war... [The Russian operation "appears to have co-opted unwitting people in Latin America to run a 'think tank'.... Since we do not allow access to our models from Russia, they used VPNs to connect to our services."] The Russian operation created fake "leaked" documents and audio scripts, some of which we identified being spread online... Both managed to land their content (not all of which was generated from our models) in mainstream media outlets, rather than simply posting it on social media.


OpenAI says they've exposed 30 covert influence operations using its tools over the last two and a half years. But ironically, in this case both operations "also made heavy use of AI to draft internal reports (the Russian operation did this more than anything else)." And "in both cases, the actors used questionable or outright deceitful methodologies to exaggerate the operators' effectiveness."

[The Russian operators] claimed that in May 2026, they created a fake email address purporting to come from the Regional Directorate of Education in Lima, Peru. They used this to instruct schools in the district to hold events dedicated to Ukraine on the national Day of Cultural and Linguistic Diversity (May 21)... According to the operators, some schools replied to the fake email address, confirming that they had held such events and even providing pictures. The operators then claimed to have planted stories about the events in the media in both Peru and Poland, alongside allegations that Ukraine was "exporting" ultra-nationalist ideologies, triggering outrage. Open-source searches identified stories that matched this claim in the Peruvianâ andâ Polish pressâ, and an English-language publication in Hungary (some of the articles have since been deleted)...

Similarly, in June, the operators claimed they used a different fake email address to trick schools in Ecuador into holding a ceremony pledging allegiance to President Daniel Noboa and to Erik Prince, former head of private military contractor Blackwater. The operators claimed that the incident provoked outrage in Ecuador and put pressure on the government to deny the fake, thus amplifying it to a nationwide audience. Again, open-source research identified mediaâ coverageâ in the Ecuadorianâ pressâ that closely resembled this claim, and even a detailed rebuttalâ by Ecuador's Minister for Education.

The operators used a range of techniques to underpin their false stories. According to their internal reporting, they spread two different fakes targeting Ecuador in March. One used fake audio attributed to Ukraine's consul in Ecuador, in which he was alleged to have made disparaging comments about Ecuadorians.



OpenAI's report "is the latest illustration of how state actors can easily exploit widely available AI tools to peddle sophisticated propaganda against adversaries on a mass scale," argues the Economic Times:

"We identified almost 100 articles published or syndicated under the [Iranian] operation's bylines across roughly a dozen online outlets around the world," OpenAI said. "These were small to medium outlets, generally focused on international affairs, geopolitics, and events in the Middle East." The earliest article identified by OpenAI was published in July 2025, and the latest in October 2026, with the frequency of reports increasing after the US-Iran war broke out earlier this year. The operation also generated social media comments on topics related to the US-Iran war, it added.

One of the personas named Ervin B. Hoskins, whose bio claimed to be "an American freelance writer," had social media accounts across tech platforms including Elon Musk's X and Meta-owned Instagram.
X's transparency information showed the account was connected via a "West Asia android app" and Instagram's transparency information showed the account was based in Iran, according to screenshots provided by OpenAI. Both accounts appeared to be suspended.

Read more of this story at Slashdot.

Formula 1 News

Formula 1® - The Official F1® Website

Why F1 fans should keep emotions out of Singapore GP

Safer gambling is important, so we’ve flagged some key areas to avoid letting your emotions run wild during Sunday's race at Marina Bay.

What are the tyre strategy options for the Singapore GP?

Matt Youson takes a look at the different pit stop and tyre options that are available to the teams on race day at the Marina Bay Street Circuit.

Our Singapore Grand Prix Bet Builder picks made

We have picked a three-leg Bet Builder for Sunday’s Grand Prix, including podium and top-six finish selections.

What To Watch For in the Singapore Grand Prix

Chris Medland picks out five key things to keep an eye on when the lights go out on race day at the Marina Bay Street Circuit.

'It is serious' – Stella addresses Norris and Piastri's clash

McLaren Team Principal Andrea Stella has shared his reaction to the collision that occurred between Lando Norris and Oscar Piastri on the final lap of the Singapore Sprint.

VK: Voorpagina

Volkskrant.nl biedt het laatste nieuws, opinie en achtergronden

Ajax komt goed weg tegen NEC (1-1), ook bij arbitrage

Wel.nl

Minder lezen, Meer weten.

NYT: zeker 12 doden na Houthi-aanval op luchthaven Riyad

RIYAD (ANP) - Bij een Houthi-aanval op de internationale luchthaven van Riyad in Saudi-Arabië zijn zeker twaalf doden gevallen, melden ingewijden aan The New York Times. Dat maakt het volgens de krant de dodelijkste aanval in een Golfstaat sinds het begin van de Amerikaans-Israëlische oorlog met Iran. Er zouden ook vijftig mensen gewond zijn geraakt.


The Guardian

Latest news, sport, business, comment, analysis and reviews from the Guardian, the world's leading liberal voice

Despair after 200-year-old ‘grandmother’ tree toppled for Trump’s border wall: ‘It crushes your heart’

Protesters spent months occupying an ancient cottonwood. Then a tense standoff with US border patrol came to a head

A 200-year-old cottonwood tree near the US-Mexico border, known as the “grandmother”, has been destroyed to make way for the construction of Donald Trump’s border wall, despite a months-long effort by protesters trying to save it.

In late July, land defenders began a tree-sit in the ghost town of Lochiel, Arizona, taking turns occupying a platform high in the tree’s canopy. The group remained there for more than 70 days.

Continue reading...

Strictly Come Dancing week three – live

Which couples will wow the judges with their well-executed choreography? And who will put their foot in it? Find out as the third week of the contest kicks off

I’ve adjusted to the new-look title sequence but that weird stop-start gap in the theme tune still throws me. It’s like a brief power outage. Quick, dad, the fusebox!

Hydrate, carb-load and commence some light stretching. We’re about to go live to Elstree Studios…

Continue reading...

Found Photograph

Thomas Hawk posted a photo:

Found Photograph

Chicago Mornings With The Studio Gang

Thomas Hawk posted a photo:

Chicago Mornings With The Studio Gang

Taubertal

Peter Kernwein posted a photo:

Taubertal

Bieberehren

Taubertal

Peter Kernwein posted a photo:

Taubertal

Bieberehren

Taubertal

Peter Kernwein posted a photo:

Taubertal

Bieberehren

XScreenSaver 6.17

XScreenSaver 6.17 is out now for MacOS, iOS (eventually), Android and Unix. This is mostly a maintenance release, triggered by some onerous and stupid Apple changes.

  • I think I worked around several longstanding Apple bugs where the "legacyScreenSaver" process would start using 100% CPU and savers would never run again.

  • MacOS 14.6 or later and iOS 17.6 or later are now required, because Apple sucks.

    As I value my sanity, my desktop machine is still running macOS 14.7. But Xcode 6 won't run on that (because they suck), and Apple won't accept app store submissions from older versions of Xcode (because they suck). So I got a used Mac Mini with macOS 27 to use as a sacrifice zone and build machine.

  • X11 localizations are being installed again. Apparently that broke 3+ years ago and nobody noticed or cared.

  • macOS, iOS and Android are now localized.

  • I added many missing localization strings, largely auto-translated via Google Translate.

Mr. Gotcha pops out of the well: "But jwz, you say the 'AI' industry is unethical, and yet you used Google Translate. Curious!"

I'm old enough to remember when Google Translate was merely "machine translation" and not "Skynet the Eschaton Supergod Climate-Reaper", so let's pretend it's still that? I'm sure that some of the new auto-translations are shitty, but the last time most of them were updated was in 2002, which I assume is when Red Hat stopped paying people to do it. If you care about this (more than you care about Gotcha Points), you can help! For each entry marked "auto-translated" in the .po file of a language that you speak, please:

  • Verify it or correct it;
  • Remove the "auto-translated" comment;
  • Email me the new file.

If you aren't willing or able to do even that, then I encourage you to go climb back down into the well.

Rijnmond - Nieuws

Het laatste nieuws van vandaag over Rotterdam, Feyenoord, het verkeer en het weer in de regio Rijnmond

Van Bronckhorst wil niets weten van kampioenspraat: 'Er is nog een hele competitie te gaan'

De Kuip vierde zaterdagavond de nieuwe koploper van de eredivisie. Feyenoord versloeg directe concurrent AZ met 2-0 en staat voor het eerst dit seizoen bovenaan. Giovanni van Bronckhorst genoot zichtbaar van de overwinning, maar drukte de euforie na afloop meteen de kop in. "Geniet van de overwinning en geniet dat we bovenaan staan, maar blijf wel met beide benen op de grond. Er is nog bijna een hele competitie te gaan."

Frontaal Naakt

Fundamentalistisch-hedonistisch webmagazine voor de hele familie

Cocteau Twins, witter dan wit

    Ik was nooit een fan van de Cocteau Twins of zo, het was gewoon één van de tientallen post-punk, goth-rock, new wave bands die ze in de Swing draaiden, de alternatieve Apeldoornse dansclub waar ik in de jaren 80 elke zaterdagavond heenging. Ik zag het bovenstaande filmpje toevallig op YouTube, werd bevangen door nostalgie en raakte gefascineerd door de zangeres en haar bizarre mimiek, zo houterig en in zichzelf gekeerd en neurotisch, verkrampt en schuchter met een zweem arrogantie maar tegelijk ook heel kwetsbaar. Witter dan dit krijg je het niet, maar ergens vind ik het sexy. Maar het is vreemd. Dat trillen van die stem ook. Witte mensen zijn heel erg raar. Ik mag dat zeggen, want ik ben zelf wit. Punk en post-punk zijn bij uitstek wit. Dacht ik, totdat ik bands als Death, Pure Hell en Bad Brains ontdekte, en natuurlijk Fishbone, dat in de basis ook punk is. Maar toen ik het overal hoorde, was punk en wat daaruit voortkwam altijd wit en altijd ook een beetje gênant, maar dat is deel van de aantrekkingskracht.

Het bericht Cocteau Twins, witter dan wit verscheen eerst op Frontaal Naakt.

The Fog and the Tree

Matt Straite Photography has added a photo to the pool:

The Fog and the Tree

Ok, not really a tree, but the Tokyo Skytree. You can try to predict the weather, I know some folks who are very good at it. I am not. And often I don't have the ability when on a big to trip to cater too much to the climate. I am where I am when I am (usually, with some exceptions). This was not one of those exceptions. I really only had this one change to go photograph the Skytree. I wanted anything but blank skies. So I was delighted when we showed up and there was light, high fog. Ok, maybe you call that clouds, either way it worked for me. I had researched the view on several bridges over this canal. This one was by far my favorite. I loved the relationship with the street and the surrounding buildings. The staircase gave it a human element without adding humans.

This is several exposures to get the light trails right, though most came from just one image. I loved this image, hope you like it too.