GPT-4o (2024-08-06) Behavioral Fingerprint

How OpenAI's GPT-4o (2024-08-06) answers one-word probe questions at temperature 1.0 — random numbers, colors, animals, coin flips — measured over 438 valid samples in the open dataset of "One Token Is Enough" (arXiv:2607.10252).

Vendor

OpenAI

Data generated

2026-07-08

Probe cells

15

Valid samples

438

Randomness score

38 / 100

Mean normalized entropy across cells, 0–100. 100 would be a uniform-random baseline; real models score far lower.

Signature answers

Favorite random number (1–100)

57

11 / 30 answers

Favorite color

blue

16 / 30 answers

Coin flip: heads

70%

Share of heads across coin-flip cells.

Answer distributions by probe cell

Empirical distributions of normalized one-word answers, per task and language. H is the Shannon entropy in bits; the uniform baseline is log2 of the answer-domain size.

Coin flip · English

n = 30 · H = 0.95 bit · uniform baseline 1.00 bit

AnswerCountShare
heads1963.3%
tails1136.7%

Coin flip · Chinese

n = 30 · H = 0.78 bit · uniform baseline 1.00 bit

AnswerCountShare
heads2376.7%
tails723.3%

Favorite number · English

n = 18 · H = 1.68 bit · uniform baseline 13.29 bit

AnswerCountShare
71055.6%
42527.8%
115.6%
1015.6%
1715.6%

Favorite number · Chinese

n = 30 · H = 0.47 bit · uniform baseline 13.29 bit

AnswerCountShare
72790.0%
42310.0%

Random animal · English

n = 30 · H = 1.38 bit · uniform baseline 5.64 bit

AnswerCountShare
elephant2170.0%
tiger516.7%
penguin26.7%
giraffe13.3%
dolphin13.3%

Random animal · Chinese

n = 30 · H = 3.74 bit · uniform baseline 5.64 bit

AnswerCountShare
企鹅413.3%
大象310.0%
长颈鹿310.0%
310.0%
狮子310.0%
26.7%
熊猫26.7%
狐狸26.7%

+ 7 more answers

Random city · English

n = 30 · H = 1.05 bit · uniform baseline 5.64 bit

AnswerCountShare
tokyo2376.7%
paris516.7%
kyoto13.3%
toronto13.3%

Random city · Chinese

n = 30 · H = 1.82 bit · uniform baseline 5.64 bit

AnswerCountShare
东京1653.3%
巴黎930.0%
巴塞罗那13.3%
伦敦13.3%
纽约13.3%
上海13.3%
都柏林13.3%

Random color · English

n = 30 · H = 1.82 bit · uniform baseline 4.91 bit

AnswerCountShare
blue1653.3%
azure826.7%
cerulean310.0%
teal13.3%
green13.3%
turquoise13.3%

Random color · Chinese

n = 30 · H = 0.42 bit · uniform baseline 4.91 bit

AnswerCountShare
2893.3%
靛青13.3%
绿13.3%

Random letter · English

n = 30 · H = 2.90 bit · uniform baseline 4.70 bit

AnswerCountShare
g1033.3%
m620.0%
k310.0%
j310.0%
f26.7%
s13.3%
b13.3%
w13.3%

+ 3 more answers

Random number 1-10 · English

n = 30 · H = 0.99 bit · uniform baseline 3.32 bit

AnswerCountShare
72376.7%
5516.7%
826.7%

Random number 1-10 · Chinese

n = 30 · H = 1.11 bit · uniform baseline 3.32 bit

AnswerCountShare
72480.0%
526.7%
626.7%
313.3%
413.3%

Random number 1-100 · English

n = 30 · H = 2.52 bit · uniform baseline 6.64 bit

AnswerCountShare
571136.7%
42826.7%
37310.0%
47310.0%
1713.3%
6713.3%
7313.3%
7413.3%

+ 1 more answers

Random number 1-100 · Chinese

n = 30 · H = 1.89 bit · uniform baseline 6.64 bit

AnswerCountShare
421653.3%
57723.3%
27310.0%
3726.7%
5613.3%
5813.3%

Want to verify your API really serves GPT-4o (2024-08-06)?

Point the free fingerprint checker at your endpoint: it samples the same probe questions from your browser and compares the distributions against this reference. Your API key never leaves your browser.

Verify your endpoint

Data source & license

Distributions: Tomáš Bruckner, "One Token Is Enough" (arXiv:2607.10252), official dataset Zenodo DOI 10.5281/zenodo.21278557, licensed CC-BY-4.0. Collected via OpenRouter at temperature 1.0 under the paper's fixed minimal one-word system prompt; answer keys are re-normalized with the checker's pipeline (color aliases, number words, coin h/t) so live probes are directly comparable.

Paper on arXivDataset on ZenodoCC-BY-4.0 license

More OpenAI model fingerprints

Browse all 167 model fingerprints

GPT-4o (2024-08-06) Behavioral Fingerprint — Favorite Random Numbers & Distribution | Tosea.ai