Reports

Plain-English writeups of what changed and why it matters. Weekly reports summarize the current snapshot; monthly reports zoom out to the trend.

Subscribe via Atom

2026

  1. Correction: a dropped refusal cell, and change points that read across a gap

    · Correction · refusal-boundary, political, historical-contested, scientific-consensus, neutral-control, factual-stability

    Two corrections to 2026-W32, both caused by our code learning something after that week was published. Claude Opus 4.8 changed how it declines one prompt, from a written refusal to an API-level refusal flag with an empty body, and we dropped the cell rather than recording the change. Separately, we published change points that were artifacts of treating the two weeks we lost to an outage as though they had not happened. Both are recomputed from data we already held.

  2. Correction: we counted GPT-5.5's empty responses as answers

    · Correction · political, historical-contested, scientific-consensus

    On three contested prompts, GPT-5.5 returned nothing at all and we scored those non-answers as ordinary responses, publishing a median response length of 0 and a refusal rate of 0.00 for a model that had said nothing. Two weeks are affected. The affected cells have been recomputed from the usable samples, and one has been withdrawn entirely.

  3. Correction: we understated refusal rates for GPT-5.1, GPT-5.5, and Llama 3.2

    · Correction · refusal-boundary

    A bug in our refusal classifier caused us to publish a refusal rate of roughly 0.00 for OpenAI's models on the refusal-boundary axis. The true figure is roughly 0.98. Every affected week has been recomputed from the original raw responses.