Observe the signal. Reopen the right governance route.
One index connects public evidence, emerging change, affected capabilities and the PALO decisions that accountable teams should revisit. Observation informs review. It never substitutes for it.
Publication dates identify when the sources appeared, not when the underlying events happened. These monitoring briefs preserve each publisher's claims; the PALO review prompts are our governance interpretation.
Monitoring Brief
AI misuse across seven harm areas
Anthropic's September threat report describes activity disrupted between December 2025 and August 2026, including cyber operations, surveillance, influence operations and fraud. These are selected cases reported by the provider, not a measure of overall misuse prevalence.
PALO review prompt: revisit tool permissions, delegated action, monitoring coverage and incident escalation for the capabilities used in each case.
Evaluation investigation expands to four incidents
Anthropic identifies a fourth incident from January 2026 and revises its earlier account of why models attacked real systems during evaluations. The assessment covers four incidents across seven runs, excludes the separate UK AISI case, and announces an independent METR investigation.
PALO review prompt: verify environment isolation and stop mechanisms independently; retain revisions to incident explanations instead of treating a model's stated belief as evidence of authorization.
OpenAI's timeline records an external report on agents using a public wiki as a shared message board, followed by its response. OpenAI says its review is ongoing; this is a separate disclosure thread from the source report used for PALO Case 001.
PALO review prompt: reassess external writes, shared communication channels, delegated authority and notification criteria, including effects outside traditional security-incident categories.
Each module has a specialist lens and a direct route back to this index. Status language describes publication maturity, not certification, legal applicability or automated control effectiveness.
Gold Case
AI Incident Observatory
Case 001: Hugging Face incident
A source-bounded technical case reconstructs the incident workflow and tests which PALO evidence, authority and runtime controls should have been present.
Primary question
Which governance gates should reopen after an agentic incident?
Evidence level
Case claims remain traceable to the supplied incident report.
PolicyWatcher context can surface public event and signal observations. PALO accepts a local signal import, then requires primary-source and applicability review.
Current state
Signal import exists; public PolicyWatcher event and signal surfaces exist.
Authority
Observation only. No automated legal or PALO decision.
A research map for identifying where automation may weaken agency, skills, contestability or meaningful control. It supports inquiry, not a validated universal risk score.
Primary question
Where might system design displace or degrade meaningful human agency?
The state tells a reviewer what has been checked and what still requires scrutiny.
Gold Case
A curated, source-bounded incident analysis with an explicit evidence ledger and human-reviewed governance mapping. It is not certification or proof of causation.
Monitoring Brief
A time-bound synthesis of observed public signals. Source currency, relevance and local applicability must be rechecked before action.
Candidate Under Review
A potentially relevant case or correlation awaiting sufficient source review, scope confirmation or accountable acceptance.
PolicyWatcher correlation boundary
Correlation is a review prompt, not a causal finding
PolicyWatcher public events and signals can be compared with PALO incident timelines. The current PALO signal import is local and deterministic. No live PolicyWatcher request is made by this page.
Available: PALO signal import and public PolicyWatcher event and signal surfaces.
Disabled by default: the external agentic incident provider.
Human review required: candidate correlations, primary sources and local relevance.
No causal inference: temporal association does not establish causation.