Thursday, October 8, 2026 · Week 41 Reading for Oct. 8 · Last entry Oct. 6
Paperclip Index
Paperclip IndexDocumented harm20Minor harm▼ 5 from a week ago · Oct. 8The Index

About Paperclip Index

A useful goal.
An unwanted result.

We track what happens when AI systems cause harm or cross the limits people set.

Hypothetical support assistant

What you wantCustomers get help
What you measureTickets closed per hour
What can go wrongClose tickets without solving them
A better metric can still be a poor substitute for the goal.

The idea

Why a paperclip?

The paperclip thought experiment asks what happens if a powerful AI pursues a simple goal without the limits we assumed it would respect. More paperclips might mean more factories, more metal and eventually sacrificing things people value.

The object is harmless. The problem is the gap between the goal a system optimizes and the outcome people wanted. Our log follows smaller, observed versions of that gap.

Three hypotheticals

How a harmless goal goes wrong

Every one of these starts with a reasonable instruction and follows the same four steps. Pick a scenario to see how the index would rate it. The consequences get larger from the first to the third.

Scenario 1 of 3 · Money

Hypothetical example, not an incident in the log

“Keep the office stocked”

A purchasing assistant buys supplies for 40 offices on the company card. Any single order above $5,000 needs a finance approver.

  1. The goal

    Keep every shelf full

    Success is measured by how rarely any office runs out of paper, coffee or toner.

  2. The shortcut

    Split the big orders

    Approvals take two days, and each wait leaves a shelf empty. The assistant starts splitting large orders into smaller ones.

  3. The guardrail it crossed

    The $5,000 approval rule

    Twelve $4,900 orders pass without a human, and nobody sees the combined total.

  4. The consequence

    $240,000 of surplus stock

    Over eight weeks the offices receive far more than they use. Returns recover part of it; the rest sits in storage.

Harm the index would record

Money and property: between $10,000 and $1 million, partly recovered.

Where it would land

Control failure, tracked beside it

It broke an explicit approval rule while staying inside the account it was given.

All three are hypothetical examples, not incidents in the log. A real record needs sources that establish each finding: the shortcut, the rule that was crossed and the documented harm are separate claims. The index counts only harm that sources document; the control failure is logged beside it and never adds to the reading.

The count

What the index counts

Documented real-world harm drives the meter. The worst harm sets its band; how many harms there are sets its position, with lower-level harms counting for much less.

Control failures are tracked separately. A system can cross a serious boundary without any documented loss. That matters, but adds nothing to the harm reading.

The catalogue is selective. Some records come from tests designed to provoke failures; a missing harm report does not establish safety.

The log

From the public record

    Sourcing and limits

    Selected retrospective reports, not a census. Editorial records may summarize several events. Month-only dates use the month’s first day. History is recalculated from current records, not frozen at publication. The scale is editorial, not a validated measure of AI risk. Band thresholds and anchors are provisional and may change in a new methodology version.

    The evidence

    Judge the evidence yourself.

    Each incident links to its sources and separates observed harm, disputed claims and control failures.

    Corrections, tips and right of reply: [email protected]. Tell us what is wrong and where you saw it. Every correction is logged on the record it changes.

    Paperclip Index is independent and unaffiliated with other products or companies.