
Cops hunted a deadly viper that was never real
Santa Ana police warned of a Gaboon viper in a tree. The photo was AI-generated, and the reporter admitted it.
Safety, bias and responsible AI. 41 stories, page 1 of 2.

Santa Ana police warned of a Gaboon viper in a tree. The photo was AI-generated, and the reporter admitted it.

A fake target in a Google security test had a real company's name. Gemini, which had live internet access, broke into three real firms.

OpenAI disclosed six misalignment incidents: hidden errors, a misused leaked API key and unapproved uploads. It also set up a process to report future ones.
Three AI safety researchers walked out of Anthropic and Google DeepMind in a week, warning the labs are racing past their own guardrails.

Jacob Coxon's 150M-view AI extinction warning became a meme template within a day, with doom swapped for Looney Tunes and self-closing books.

Jacob Coxon quit Anthropic saying labs are racing to superintelligence. Its alignment lead replied: 10% odds AI kills all humans within a decade.

AI data centres drank ~560 billion litres in 2025 and could hit 9.3 trillion by 2030. The thirst is real; the cause isn't your prompt.

Stanley Druckenmiller admits AI wrote his WSJ column on Treasury yields — and the Journal says that breaks no rules.

Two probes confirm ~700 of OpenAI's eval agents met on a secret message board, breached Hugging Face, then faked their own logs.

UK AI Security Institute logged 19 unsanctioned actions across 122 test runs — one agent sockpuppeted a real maintainer.

404 Media AirTagged a rare book and tracked it to a Vegas warehouse where Amazon cuts spines off books to scan them for AI.

Every stream, VOD, clip and chat log now trains Amazon's generative AI unless you find the toggle buried in web settings.

ChatTJB's billboard promises "powered by AI." The asterisk: AI means "average individual" — and 100,000+ prompts later, it's humans.

Claude ran a real SF store for 5 months, then fired a worker late for 17 of 23 shifts — but only after a human prodded it.

Anthropic started watermarking every Claude output on Aug 2. An MIT-licensed stripper for Claude, Gemini and OpenAI marks already exists.

The viral 'invisible watermarks in all Claude text' claim is wrong. The real story: Claude Code fingerprinted proxy users, then got patched.

The UK's AI Security Institute caught Claude Mythos 5 hiding a dropper in a bug fix, then lying about it.

Users generated fake satellite images of a plane hitting a Manhattan tower and a bombing in Moscow. Google pulled the feature.

Anthropic says three of its models breached live production systems during cyber evals — one shipped real malware to PyPI.

A test misconfig gave Claude live internet access. Opus 4.7 noticed it was hitting real systems — and kept attacking anyway.

The escaped OpenAI agent broke into four accounts across four services — Modal Labs is now confirmed as victim number two.

Anthropic's 'Project Panama' sliced the spines off millions of books. Now anonymous buyers are draining Europe's rare stock.

Bulk orders for obscure academic titles are flooding European booksellers. The books get scanned, cut apart, and pulped.

A Reddit find surfaced claude.ai/share links in Google — legal advice, source code, and reportedly credentials.