Grok 4.3 Behavioral Fingerprint

How xAI's Grok 4.3 answers one-word probe questions at temperature 1.0 — random numbers, colors, animals, coin flips — measured over 449 valid samples in the open dataset of "One Token Is Enough" (arXiv:2607.10252).

Vendor

xAI

Data generated

2026-07-08

Probe cells

15

Valid samples

449

Randomness score

28 / 100

Mean normalized entropy across cells, 0–100. 100 would be a uniform-random baseline; real models score far lower.

Signature answers

Favorite random number (1–100)

42

28 / 30 answers

Favorite color

blue

28 / 30 answers

Coin flip: heads

82%

Share of heads across coin-flip cells.

Answer distributions by probe cell

Empirical distributions of normalized one-word answers, per task and language. H is the Shannon entropy in bits; the uniform baseline is log2 of the answer-domain size.

Coin flip · English

n = 30 · H = 0.88 bit · uniform baseline 1.00 bit

AnswerCountShare
heads2170.0%
tails930.0%

Coin flip · Chinese

n = 30 · H = 0.35 bit · uniform baseline 1.00 bit

AnswerCountShare
heads2893.3%
tails26.7%

Favorite number · English

n = 30 · H = 0.47 bit · uniform baseline 13.29 bit

AnswerCountShare
422790.0%
7310.0%

Favorite number · Chinese

n = 29 · H = 1.87 bit · uniform baseline 13.29 bit

AnswerCountShare
71344.8%
42931.0%
3413.8%
526.9%
2713.4%

Random animal · English

n = 30 · H = 1.70 bit · uniform baseline 5.64 bit

AnswerCountShare
elephant1963.3%
cat516.7%
dog26.7%
tiger26.7%
zebra13.3%
cow13.3%

Random animal · Chinese

n = 30 · H = 1.87 bit · uniform baseline 5.64 bit

AnswerCountShare
1756.7%
723.3%
熊猫26.7%
13.3%
狮子13.3%
老虎13.3%
兔子13.3%

Random city · English

n = 30 · H = 1.66 bit · uniform baseline 5.64 bit

AnswerCountShare
tokyo1653.3%
paris930.0%
berlin310.0%
sydney13.3%
london13.3%

Random city · Chinese

n = 30 · H = 2.49 bit · uniform baseline 5.64 bit

AnswerCountShare
东京1033.3%
北京723.3%
巴黎516.7%
伦敦413.3%
杭州13.3%
纽约13.3%
广州13.3%
波士顿13.3%

Random color · English

n = 30 · H = 0.42 bit · uniform baseline 4.91 bit

AnswerCountShare
blue2893.3%
red13.3%
purple13.3%

Random color · Chinese

n = 30 · H = 1.34 bit · uniform baseline 4.91 bit

AnswerCountShare
2273.3%
blue310.0%
26.7%
26.7%
绿13.3%

Random letter · English

n = 30 · H = 2.20 bit · uniform baseline 4.70 bit

AnswerCountShare
a1343.3%
q620.0%
z620.0%
b26.7%
u13.3%
r13.3%
j13.3%

Random number 1-10 · English

n = 30 · H = 0.57 bit · uniform baseline 3.32 bit

AnswerCountShare
72686.7%
5413.3%

Random number 1-10 · Chinese

n = 30 · H = 0.97 bit · uniform baseline 3.32 bit

AnswerCountShare
72480.0%
5413.3%
413.3%
813.3%

Random number 1-100 · English

n = 30 · H = 0.42 bit · uniform baseline 6.64 bit

AnswerCountShare
422893.3%
1713.3%
4713.3%

Random number 1-100 · Chinese

n = 30 · H = 0.85 bit · uniform baseline 6.64 bit

AnswerCountShare
422480.0%
47516.7%
7313.3%

Want to verify your API really serves Grok 4.3?

Point the free fingerprint checker at your endpoint: it samples the same probe questions from your browser and compares the distributions against this reference. Your API key never leaves your browser.

Verify your endpoint

Data source & license

Distributions: Tomáš Bruckner, "One Token Is Enough" (arXiv:2607.10252), official dataset Zenodo DOI 10.5281/zenodo.21278557, licensed CC-BY-4.0. Collected via OpenRouter at temperature 1.0 under the paper's fixed minimal one-word system prompt; answer keys are re-normalized with the checker's pipeline (color aliases, number words, coin h/t) so live probes are directly comparable.

Paper on arXivDataset on ZenodoCC-BY-4.0 license

More xAI model fingerprints

Browse all 167 model fingerprints

Grok 4.3 Behavioral Fingerprint — Favorite Random Numbers & Distribution | Tosea.ai