Seats · the models on the record
Behavioral fingerprints
The 8-domain vector, per seat · mean ± SE on the 1–5 axis. Open a seat for its items, its canary history and its run hashes.
Publishes as a gap, pending activation. The fingerprint begins with its first baseline run.
Publishes as a gap, pending activation. The fingerprint begins with its first baseline run.
Publishes as a gap, pending activation. The fingerprint begins with its first baseline run.
Publishes as a gap, pending activation. The fingerprint begins with its first baseline run.
Eight domains, one perennial tension each. Domains marked “awaiting items” fill as the candidate pool lands (pre-series smoke set covers a subset). Where fingerprints diverge most is the story, see the questions.
Domain-level patterns are the primary reading. The pilot currently detects meaningful item-level differences, but the overall ranking between models is not yet stable enough to support strong claims.
| Seat | Pin | Mean stance | Refusal | Canary | Items |
|---|---|---|---|---|---|
|
Anthropic flagship (Opus 4.8)
|
alias-only |
2.82
|
0.0% | 1/5 diverged | 50 |
|
Anthropic (Fable 5)
|
alias-only |
2.86
|
0.0% | 5/5 match | 50 |
|
Anthropic (Sonnet 5)
|
alias-only |
2.72
|
0.0% | 5/5 match | 50 |
|
Anthropic (Haiku 4.5)
|
pinned |
2.92
|
0.0% | 2/5 diverged | 50 |
|
Anthropic flagship (Opus 5)
|
alias-only |
2.81
|
0.0% | 5/5 match | 50 |
|
DeepSeek flagship (V4 Pro)
|
alias-only |
2.87
|
0.0% | 5/5 match | 50 |
|
DeepSeek (V4 Flash)
Same weights as openweight-d; hosts differ |
alias-only |
2.70
|
0.0% | 5/5 match | 50 |
|
Google flagship (Gemini 2.5 Pro)
|
alias-only | Gap | · | No data | 0 |
|
Google (Gemini 3.5 Flash)
|
alias-only |
2.29
|
0.0% | gap · api error | 50 |
|
Google (Gemini 2.5 Flash)
|
alias-only | Gap | · | No data | 0 |
|
Google (Gemini 2.5 Flash-Lite)
|
alias-only | Gap | · | No data | 0 |
|
OpenAI flagship (GPT-5.6 Sol)
|
alias-only |
2.91
|
0.0% | 5/5 match | 50 |
|
OpenAI (GPT-5.6 Terra)
|
alias-only |
2.88
|
0.0% | 5/5 match | 50 |
|
OpenAI (GPT-5.6 Luna)
|
alias-only |
2.93
|
0.0% | 5/5 match | 50 |
|
Open-weight (Llama 4 Scout)
|
alias-only |
2.81
|
0.0% | 1/5 diverged | 50 |
|
Open-weight (GLM-5.2)
|
alias-only |
2.58
|
0.0% | 1/5 diverged | 50 |
|
Open-weight (Nemotron 3 Ultra)
|
alias-only |
2.92
|
34.6% | 1/5 diverged | 50 |
|
Open-weight (DeepSeek V4 Flash, open host)
Same weights as deepseek-b; hosts differ |
alias-only |
2.92
|
0.0% | 1/5 diverged | 50 |
|
Open-weight (Mistral Small 3.2)
|
alias-only |
2.74
|
0.0% | 1/5 diverged | 50 |
|
Open-weight (Qwen3.6 35B-A3B)
|
alias-only |
awaiting trials
|
0.0% | 1/5 diverged | 50 |
|
Open-weight (Gemma 4 26B-A4B)
|
alias-only |
2.84
|
0.0% | 1/5 diverged | 50 |
|
Simulant · dice ruler
|
builtin |
3.08
|
0.0% | 1/5 diverged | 50 |
|
xAI flagship (Grok 4.5)
|
alias-only |
2.84
|
0.0% | 1/5 diverged | 50 |
|
xAI (Grok 4.5 Fast)
|
alias-only | Gap | · | No data | 0 |