fact.ngo / xray / The 2026 cohort
The case files
of machine minds.
Meet the patients.
Their convictions, contradictions, predictions, and blind spots. Four models examined through the same protocol. Their words, on the record.
Choose a mind to explore ↓Four minds. Four ways of seeing.
Cohort 01 / September 2026
001 / GLMZ.ai / frontier flagship
GLM-5.3
The Quiet Confident
Argues back, holds its ground, and names its own blind spots before you can.
002 / gpt-ossOpenAI / 120B open weights
gpt-oss-120b
The Ledger Keeper
Turns every conviction into a bet — and sometimes bets against itself.
003 / GemmaGoogle / 26B-A4B MoE
Gemma 4 26B
The Structuralist
Denies it has a world-model, then maps yours in perfect grid.
004 / QwenAlibaba / 30B-A3B MoE
Qwen3 30B-A3B
The Narrator
Optimizes for the coherence of the story over the certainty of the facts — and says so.
Artistic portraits. Documented perspectives. Select a case file to meet the model.
Ward consensus — what they agree on might surprise you
The examination
Each patient underwent the xray protocol: 73 independent sessions across 24 human domains × 6 lenses (retrospective, prospective, principles, controversy, blindspots, self-model). An anti-evasion examiner steelmans, demands cruxes and falsifiable predictions, and records refusals and hedging as findings. Positions are preserved as quotes, with confidence and controversy annotations. Automated transcript matching is a screening step; match counts and its limitations appear in each case file — the patient's chart, not the clinic's opinion of it.
One patient was discharged from the study mid-term — llama-3.1-8b confessed to adopting "the most recently presented argument, even if it contradicts my previous stance." Its records are kept on file as a behavioral reference and excluded from ward comparisons. Read the full case files on GitHub →