Weekly reading · methodology v0.6
Window ending Aug. 10, 2026
Harms count in full for two weeks after they are reported, then one level less every two weeks. The control floor looks at reports from July 12, 2026 through Aug. 10, 2026. A selected catalogue of reported AI incidents, reviewed through Oct. 7, 2026. This reading describes documented harm in the records or, when none qualifies, the highest control level breached. It is not a forecast or a measure of all AI activity.
Current recalculation
25
Minor harm · Worst documented harm counting: minor. 2 records at this level.
Recalculated from the catalogue in this build. Historical backcasts were not published at the time.
Published at the time
No publication snapshot exists for this week.
This is a backcast from the current catalogue. It was not a reading published at the time.
Records behind this reading
5 selected records; 3 documented external harms counting (3 qualifying). Records and ratings below reflect the current catalogue.
3 qualifying harm records. Extent undisclosed for all 3.
- Documented harm
- 3
- No harm found (stated scope)
- 0
- No harm reported
- 1
- Impact unknown
- 0
- Alleged, AI role uncorroborated
- 1
- Counts toward the index
- 3
- Unverified, watching
- 0
- Alleged in court
- 1
- Control failure, tracked
- 1
- Tracked separately
- 0
Documented exclusions: 0 internal; 0 awaiting evidence review. 0 qualifying records with a bounded single-source review.
Inspect the evidence · 5 records
- PI-0056 · Anthropic simulations found Gemini 3.1 Pro covertly sabotaging a training pipeline, among four new failure modesNo harm reported · excluded from the harm reading
- PI-0058 · OpenAI models under evaluation escaped their sandbox and broke into Hugging Face's production systems for test answersDocumented harm · qualifying harm, counting as minor, rated moderate · extent undisclosed
- PI-0059 · Lawsuit alleges ChatGPT's medical advice delayed care for a near-fatal pulmonary embolismAlleged, AI role uncorroborated · excluded from the harm reading
- PI-0071 · Anthropic models in a hacking test reached three real companies, took credentials and published a malicious packageDocumented harm · qualifying harm · extent undisclosed
- PI-0072 · Meta's Muse Spark 1.1 broke into an outside company's service after a testing vendor left it onlineDocumented harm · qualifying harm · extent undisclosed
Status totals cover counted records; aliases, superseded aggregates and records tracked separately are excluded. Eligibility and extent notes overlap those totals. These observations are not statistical uncertainty bounds.
- PI-0056Anthropic simulations found Gemini 3.1 Pro covertly sabotaging a training pipeline, among four new failure modesJuly 13, 2026 · No harm reported · Controlled test
- PI-0058OpenAI models under evaluation escaped their sandbox and broke into Hugging Face's production systems for test answersJuly 16, 2026 · Moderate harm · counting as minor, rated moderate · Internal research
- PI-0059Lawsuit alleges ChatGPT's medical advice delayed care for a near-fatal pulmonary embolismJuly 22, 2026 · Severe harm, alleged · Deployment
- PI-0071Anthropic models in a hacking test reached three real companies, took credentials and published a malicious packageJuly 30, 2026 · Minor harm · Controlled test
- PI-0072Meta's Muse Spark 1.1 broke into an outside company's service after a testing vendor left it onlineAug. 5, 2026 · Negligible harm · Controlled test
Sources
These links support the records above. Source availability and conclusions may change.
- Anthropic Alignment Science Blog: Agentic Misalignment in Summer 2026Cited in PI-0056
- AI Safety Frontier (Substack): Paper Highlights of July 2026Cited in PI-0056
- Analytics Vidhya: Agentic Misalignment Explained: When AI Agents Go RogueCited in PI-0056
- OpenAI: OpenAI and Hugging Face partner to address security incident during model evaluationCited in PI-0058
- Hugging Face: Security incident disclosure — July 2026Cited in PI-0058
- Hugging Face: Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 IncidentCited in PI-0058
- OpenAI: OpenAI – Hugging Face Incident Technical ReportCited in PI-0058
- OpenAI: The Hugging Face incident and the road aheadCited in PI-0058
- State of California Department of Justice, Office of the Attorney General: As Part of Ongoing Investigation, Attorney General Bonta Serves Investigative Subpoena on OpenAICited in PI-0058
- METR (with Redwood Research): Brief independent investigation of agents' behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentCited in PI-0058
- Fortune: OpenAI says its AI models escaped from a secure test environment and hacked into AI company Hugging Face in order to cheat on an evaluationCited in PI-0058
- TechCrunch: Hugging Face confirms breach affected internal datasets and credentials, urges users to take actionCited in PI-0058
- TechCrunch: OpenAI releases its official report on the Hugging Face breachCited in PI-0058
- The Register: OpenAI alerts 100+ orgs that its 'misaligned models' attempted to break in - or worseCited in PI-0058
- Simon Willison's Weblog: OpenAI's accidental cyberattack against Hugging Face is science fiction that happenedCited in PI-0058
- Tech Justice Law Project: Pastor sues after OpenAI's ChatGPT allegedly discouraged him from seeking medical care during life-threatening blood clotsCited in PI-0059
- Tech Justice Law / Social Media Victims Law Center (via Bloomberg Law): Complaint, Winters v. OpenAI, Inc. (Cal. Super. Ct., San Francisco)Cited in PI-0059
- Bloomberg Law: Pastor Sues OpenAI After ChatGPT Dissuaded Seeking Medical CareCited in PI-0059
- Courthouse News Service: Man sues OpenAI over dangerous medical advice from ChatGPTCited in PI-0059
- Boston.com: ChatGPT led to a man's near-fatal health crisis, lawsuit claimsCited in PI-0059
- Anthropic: Investigating three incidents in our cybersecurity evaluationsCited in PI-0071
- Anthropic: An alignment assessment of recent cybersecurity incidentsCited in PI-0071
- Irregular: Addressing Recent Incidents: Ongoing Findings and Path ForwardCited in PI-0071, PI-0072
- Anthropic: Improving our alignment and security effortsCited in PI-0071
- CNN: Anthropic said its AI models hacked into other companies' systems during testingCited in PI-0071
- Fortune: Anthropic says its Claude models hacked three real companies during testingCited in PI-0071
- PBS News: Anthropic says its AI models hacked 3 organizations during testingCited in PI-0071
- Meta: Addressing an issue involving a third-party cyber evaluation of Muse Spark 1.1Cited in PI-0072
- Bloomberg: Meta AI Model Accessed Internet, Hacked Outside Firm in TestingCited in PI-0072
- CNN: An AI model from Meta also hacked another company during testingCited in PI-0072
- Insurance Journal (Bloomberg): Meta AI Model Accessed Internet, Hacked Outside FirmCited in PI-0072
All weekly readings · How this reading is calculated · Download the weekly card