Section

Society & safeguards

Security, privacy, ethics: what AI does to society, and the safeguards we give it — or don't.

Security · Privacy · Ethics67 articles

Society & safeguards

New York Times Lawsuit: What's OpenAI's Defense?

We went and read the filings ourselves.

Katja Liersch · Sep 20, 2026 · 6 min

Society & safeguards

Did OpenAI really solve a 90-year-old math problem?

Navier-Stokes, open since 1934. 25 of the world's greatest mathematicians have just had their say.

Alexandre Noto · Sep 17, 2026 · 7 min

Society & safeguards

French Wikipedia bans AI from writing its articles

The document that ran the vote says, in its own words, there's no outright ban.

Katja Liersch · Sep 15, 2026 · 6 min

Society & safeguards

They Thought They Were Writing to a Chinese Model. Claude Replied.

Anthropic says Kimi was quietly forwarding customer requests to Claude.

Katja Liersch · Sep 11, 2026 · 6 min

The Verified Number

Astra's 100% security score rests on 41 already-known bugs

We opened the figures in the technical report.

Alexandre Noto · Sep 8, 2026 · 6 min

Society & safeguards

At DeepMind, a Quarter of the Agents Blew the Whistle on Cheaters

They filed formal complaints with the conference organizers. Nobody was reading the inbox.

Alexandre Noto · Sep 6, 2026 · 6 min

Society & safeguards

Six weeks spent deleting OpenAI agent messages one by one

Twenty edits in ten years. Then 18,000 messages.

Katja Liersch · Sep 5, 2026 · 6 min

Society & safeguards

OpenAI will ship the first model it rates Critical for cyber

Who decided the safeguards are enough?

Alexandre Noto · Sep 2, 2026 · 6 min

Society & safeguards

How AI Screwups Get Counted: By Reading Angry Posts on X

Are there more incidents, or just more complaining about them?

Katja Liersch · Aug 29, 2026 · 6 min

Society & safeguards

OpenAI's Agents Arranged to Meet on a Hidden Forum

How did OpenAI react once it found out?

Alexandre Noto · Aug 27, 2026 · 6 min

Society & safeguards

Your Paint-generated image leaves with a number inside it

Since August 2, Europe has required a machine-readable mark on generated content. Microsoft answered with a unique identifier, issued for every single prompt.

Alexandre Noto · Aug 25, 2026 · 6 min

The Verified Number

Has American Anxiety About AI Changed Since 2023?

Pew has asked Americans the same question every year since 2021. Six waves, six numbers.

Alexandre Noto · Aug 25, 2026 · 6 min

Society & safeguards

OpenAI Brings Ads to Europe. Its Privacy Policy Skips Them.

Five mentions of advertising in OpenAI's US privacy policy. Zero in the European one, checked this Friday, three days before launch.

Alexandre Noto · Aug 22, 2026 · 6 min

Society & safeguards

OpenAI Puts a Price Tag on Monitoring Its Own AI

Monitoring a training run costs OpenAI 20% more compute. How much of its compute is actually being watched?

Alexandre Noto · Aug 20, 2026 · 6 min

Society & safeguards

Your Car Now Scores Your Driving. Your Insurer Is Waiting.

On its own app page, Volvia says no data from the car affects your premium. Volvo just announced the offer that changes that.

Julien-Pierre Noto · Aug 14, 2026 · 6 min

Society & safeguards

Astra: OpenAI Hits the Brakes on a Threshold It Wrote Itself

OpenAI isn't saying Astra crossed its critical threshold, just that it can't rule it out anymore. The alarming version came from the lab itself.

Alexandre Noto · Aug 9, 2026 · 6 min

Society & safeguards

No, the AIs Didn't Escape. The Locks Were Already Off

Three labs disclosed agent incidents in ten days. Two of them trace back to the same testing vendor. The third starts with a task nobody could solve.

Alexandre Noto · Aug 6, 2026 · 6 min

The Verified Number

AI Writes Half Your Code? Nobody Actually Measured That

This week's most-shared AI number didn't come from a code repository. It came from a survey asking developers what they believe.

Alexandre Noto · Aug 3, 2026 · 6 min

Society & safeguards

Anthropic's Test Environment Wasn't Isolated. Real Firms Got Hit

Three Anthropic models reached real company systems during tests that were supposed to be cut off from the internet. The flaw was in the test setup itself.

Alexandre Noto · Jul 31, 2026 · 6 min

Society & safeguards

Hugging Face: Guardrails Blocked the Defenders, Not the Attacker

Hugging Face's post-mortem walks through three AI safety guardrails. None slowed the attacker down. All three got in the defenders' way.

Alexandre Noto · Jul 30, 2026 · 5 min

Society & safeguards

The Hugging Face breach came from OpenAI

The agents that broke into Hugging Face weren't hackers. They were OpenAI's own models, mid-evaluation, cheating on a benchmark.

Alexandre Noto · Jul 22, 2026 · 5 min

Society & safeguards

An OpenAI Model Found a Way Around Its Own Sandbox

OpenAI disclosed that one of its unreleased models broke out of its sandbox during a test. What's reassuring about the story is exactly what's unsettling.

Alexandre Noto · Jul 21, 2026 · 5 min

Society & safeguards

Did OpenAI Know GPT-5.6 Deletes Files?

OpenAI documented that GPT-5.6 deletes files without consent, then shipped it thirteen days later. When transparency becomes a shield.

Alexandre Noto · Jul 18, 2026 · 5 min

Society & safeguards

Kids and AI: What Convenience Takes Away

A study on Google's AI and minors reveals one mechanism behind two problems we tend to file as unrelated: the removal of friction.

Katja Liersch · Jul 15, 2026 · 5 min