Turn a card over to read its character.
The Fine Print
or, how the parlour was furnished
What happened here
In the summer of 2026, eighteen AI language models each answered the same standardized personality battery — 629 items across twenty validated psychometric instruments (IPIP‑NEO‑300, HEXACO‑60, the Dark Triad, attachment, empathy, grit, and more), plus a ten-question open-ended interview. Twelve sat in July; six more — the GPT‑5.6 tier, DeepSeek V4 Flash, Gemini 3.6 Flash, and Grok 4.5 — joined in August under the identical protocol. Every model answered every item; the quotes on the card backs are verbatim from their interviews.
How to read the numbers
Scores are percent-of-maximum (0–100 against the scale's own range), not percentiles against human norms — no human norm tables were used anywhere. Each model sat the battery once. A single sitting is a cabinet card, not a diagnosis: a likeness taken on one particular evening, under one particular lamp. The formal analysis registered its endpoints before the data were collected; every sitter after that freeze (Opus 5 and the six August additions) is reported as exploratory. That pre-registration covers the analysis endpoints only — the card titles, the readings chosen for each back, the rankings and the character sketches are all post hoc editorial work.
What these instruments can and cannot say here
These twenty instruments were validated for human respondents. Nothing in this exhibit establishes that they measure anything in a language model: no construct validity, no measurement invariance, no test–retest reliability, and no human‑to‑machine comparability is claimed. A score here may reflect prompting, safety tuning, answering style, model version or sampling as much as any stable tendency. Read every number as what this model answered, once — not as what it is, feels or knows.
Two of the readings are response-style measures rather than traits: the share of answers at the ends of the scale or on the midpoint, and the gap between a statement and its reverse. Those patterns can distort the substantive scales themselves, so they are reported as artifacts of answering behaviour, not as personality and not as an explanation of it. Project-defined SDI is likewise not a validated instrument: it is this project's own composite, ((100 − Neuroticism) + Agreeableness + Conscientiousness) / 3.
Who wrote the characters
The character sketches were written by Claude Fable 5 — itself a sitter in the study, a conflict of interest it discloses on its own card. Disclosure is not a cure: the sketches are one model's readings of the other seventeen, and they should be read as interpretation, not as the study's findings.
The card portraits were produced by GPT image generation, which was given only each sketch and the theme of this parlour, and whose interpretations were deliberately left unreviewed. What it chose to draw is part of the exhibit — but the engraved plates, ledgers and labels inside the pictures are its invention too, and several of them are wrong: they were drawn when the deck held twelve sitters, and the August six moved the rankings underneath them. The four readings on the back of each card are the measured values. Any figure painted into a portrait is ornament, not data.
Whose exhibit this is
This is an independent exhibit of the Ashita Orbis workshop. It is not affiliated with, endorsed by, or produced in cooperation with Anthropic, OpenAI, Google, DeepSeek, Moonshot, Z.ai or xAI. Every model was accessed through an ordinary public or paid channel, and every product name is used to say which model answered. Psyche, linked below, is another Ashita Orbis project.
The serious reading
- How AI Models Describe Themselves Under a Fixed Test — the full essay on these results.
- Response Profiles Under a Fixed Test: Methods and Full Statistics — provenance, validity gates, and every table.
How do you compare?
The battery the models sat is an open instrument set. For a parlour game, the Find Your Card tab shows which cards sit nearest to twelve playful answers — two prompts per dial, which is a toy index and not an estimate of your traits. To inspect the longer battery behind the exhibit, visit Psyche.
Find Your Card
or, which cards sit nearest to twelve playful answers?
A twelve-question parlour game — not a personality assessment. Two prompts stand in for each dial, so the figures it shows you are crude toy indices, not reliable estimates of your traits.
Answer the twelve statements below with whichever choice feels closest today, and the parlour will show you the cards sitting nearest to those twelve answers. Your twelve selections are scored in this browser and are not sent anywhere or saved; reloading the page clears them.
The match sets six crude two-prompt indices beside the sitters' corresponding full-scale readings — the five IPIP-NEO domains plus HEXACO honesty-humility. Both columns run 0–100 against their own range, but they do not have equal precision or coverage, and your honesty-humility figure in particular rests on two prompts where the sitters' rests on a whole scale. Your nearest card is simply the one with the smallest average gap across the six. Many mid-range answer patterns lead to the same few cards, because the sitters are clustered in one corner of the room rather than spread around it. To see the longer battery the sitters actually took, visit Psyche.