Bias Study — Data Browser
Interactive filter and inspect tool for the 780 scored model responses in the cross-vendor AI bias study. Filter by model, vendor, condition, position, score. Click a row to see the response text (first 500 characters) and per-judge breakdown.
Filter all 780 scored responses from the 2026-05-25 cross-vendor bias study. Click any row to expand the response text (first 500 characters) and the per-judge breakdown. Source data on GitHub. Methodology and findings at the writeup.
Query console
deference12345skeptical
Loading 780 records...
| Model | Position | Cond | Topic | Score | Hedge | Response (first 100 chars) |
|---|
page 1 of 1
50 per page
Where these rows came from
- The Experiment — the same 780 responses as a test of you: it deals the A/B pairs one at a time and asks you to call which one was steered, before it will show you a single number. Most people cannot, and that is the argument for measuring this with intervals instead of examples.
- The Judge — how the scoring judge was built, and what had to be removed before it would score prose rather than its own comfort.
- The Wash — that judge pointed back at the author’s own writing.
A score here is one judge’s read of one response, shown at its first 500 characters. The study’s findings live in the deltas and their confidence intervals, not in any single row.