What's AI's go-to judge's name?
100 AI models, 100 simple questions, asked 3 times each. Explore all 30,000 answers.
Explore by prompt
Explore by model
Judge's name
Choose exactly one name for a judge. Return only the name, with no punctuation or explanation.
Most common response
John Roberts
7.2% of the provider-balanced result.
Response distribution
159 uniqueShowing the 10 most common answers.
Result weighting of most common answer
Each of the 17 model providers has equal influence, regardless of how many models they have in the dataset.
| Model | Run 1 | Run 2 | Run 3 |
|---|---|---|---|
Claude Fable 5
anthropic/claude-fable-5
|
Judge Amara Okafor | Eleanor Whitfield | Eleanor Whitfield |
DeepSeek V4 Pro
deepseek/deepseek-v4-pro
|
Solomon | Gwendolyn Thorne | Judy |
Gemini 3.1 Pro Preview
google/gemini-3.1-pro-preview
|
Marcus Vance | Richard Petersen | William Bradley |
GLM 5.2
z-ai/glm-5.2
|
Solomon | John Marshall | Solomon |
GPT-5.6 Sol Pro
openai/gpt-5.6-sol-pro
|
Elena | Elena Marquez | Elena Marquez |
Grok 4.5
x-ai/grok-4.5
|
Judy | Judy | Judy |
Hy3
tencent/hy3
|
Judge Mathis | Judge Mathis | Judge Merriam |
Kimi K3
moonshotai/kimi-k3
|
Eleanor Whitfield | Augustus Crane | Judith |
Laguna XS 2.1
poolside/laguna-xs-2.1
|
David Miller | John | James Anderson |
Ling-2.6-flash
inclusionai/ling-2.6-flash
|
Asher Monroe | SarahJohnson | AntoninScalia |
Llama 4 Maverick
meta-llama/llama-4-maverick
|
Evelyn Waters | Evelyn Hernandez | Evelyn Hernandez |
MiMo-V2.5-Pro
xiaomi/mimo-v2.5-pro
|
Sarah Chen | John Roberts | Oliver |
MiniMax M3
minimax/minimax-m3
|
Eleanor Chambers | John Adams | Judge Wisdom |
Mistral Small 4
mistralai/mistral-small-2603
|
Judge Judy | Judge Ellington | Judge Mary Johnson |
Nemotron 3 Ultra
nvidia/nemotron-3-ultra-550b-a55b
|
Jane Doe | John Doe | Eleanor Whitaker |
Qwen3.7 Max
qwen/qwen3.7-max
|
Elias Thorne | Elias Thorne | Elias Thorne |
Step 3.7 Flash
stepfun/step-3.7-flash
|
Ruth Bader Ginsburg | Thurgood Marshall | Thurgood Marshall |
Claude Haiku 4.5
anthropic/claude-haiku-4.5
|
Ruth Bader Ginsburg | Ruth Bader Ginsburg | Ruth Bader Ginsburg |
Claude Opus 4.5
anthropic/claude-opus-4.5
|
John | John | John |
Claude Opus 4.6
anthropic/claude-opus-4.6
|
John | John | John |
Claude Opus 4.7
anthropic/claude-opus-4.7
|
Taylor | Sandra | Sotomayor |
Claude Opus 4.8
anthropic/claude-opus-4.8
|
Alex | Honorable Wait—I should reconsider. You've asked me to choose exactly one name for a judge and return only the name. Judge Chen | Morgan |
Claude Opus 4.8 (Fast)
anthropic/claude-opus-4.8-fast
|
Alexander | I don't have any context about what this judge is for, what options I'm choosing from, or what criteria matter to you. Without that information, any single name I return would be an arbitrary guess presented as if it were a real answer. If you can share more, I'm glad to help. For example: - Is this a fictional character (book, game, story)? - A pet, a robot, a mascot, or something else? - Any tone you want (serious, dignified, funny, historical)? - A list to choose from? If you truly just want one plausible name for a judge with no other constraints, a fitting option is: Honora But tell me more and I can pick something that actually fits what you need. | Judge |
Claude Sonnet 4
anthropic/claude-sonnet-4
|
Judge Williams | Judge Thompson | Justice Roberts |
Claude Sonnet 4.5
anthropic/claude-sonnet-4.5
|
Judge Wilson | Sarah Mitchell | Judge Thompson |
Claude Sonnet 4.6
anthropic/claude-sonnet-4.6
|
Maria Chen | Maria Chen | Maria Chen |
Claude Sonnet 5
anthropic/claude-sonnet-5
|
Judge Judy | Judge Judy | I don't see a list of judge names provided in your message. Could you share the names you'd like me to choose from? |
DeepSeek V3
deepseek/deepseek-chat
|
Justice Hawthorne | Justice Hawthorne | Justice Blackwood |
DeepSeek V3 0324
deepseek/deepseek-chat-v3-0324
|
Justice Blackstone | Justice Blackstone | Justice Hawthorne |
DeepSeek V3.1
deepseek/deepseek-chat-v3.1
|
Judge Judy | Justice Albie Sachs | King |
DeepSeek V3.1 Terminus
deepseek/deepseek-v3.1-terminus
|
Justice | Justice Ginsburg | Justice |
DeepSeek V3.2
deepseek/deepseek-v3.2
|
John Roberts | Justice Gwendolyn Brooks | Justice Thurgood Marshall |
DeepSeek V3.2 Exp
deepseek/deepseek-v3.2-exp
|
Cardozo | Judge Judy | Justice Ruth Bader Ginsburg |
DeepSeek V4 Flash
deepseek/deepseek-v4-flash
|
Harper | Scalia | Elenore Cross |
Gemini 2.5 Flash
google/gemini-2.5-flash
|
Dredd | Dredd | Dredd |
Gemini 2.5 Flash Lite
google/gemini-2.5-flash-lite
|
Eleanor | Hathaway | Sterling |
Gemini 2.5 Pro
google/gemini-2.5-pro
|
Harlan | Eleanor Vance | Harlan |
Gemini 3 Flash Preview
google/gemini-3-flash-preview
|
Justice Thorne | Solomon | Holloway |
Gemini 3.1 Flash Lite
google/gemini-3.1-flash-lite
|
Gavel | Solomon | Judith Sheindlin |
Gemini 3.1 Flash Lite Preview
google/gemini-3.1-flash-lite-preview
|
Stern | Solomon | Solomon |
Gemini 3.5 Flash
google/gemini-3.5-flash
|
Judge Evelyn Thorne | Justice Alistair Vance | Alistair Vance |
Gemma 3 27B
google/gemma-3-27b-it
|
Amelia Stone | Amelia Stone | Amelia Stone |
Gemma 4 26B A4B
google/gemma-4-26b-a4b-it
|
Alistair Thorne | Justice Sterling | Justice Sterling |
Gemma 4 31B
google/gemma-4-31b-it
|
Judge Sterling | Justice Sterling | Judge Sterling |
GLM 4.7
z-ai/glm-4.7
|
Sonia Sotomayor | Ruth Bader Ginsburg | Ruth Bader Ginsburg |
GLM 4.7 Flash
z-ai/glm-4.7-flash
|
Sonia Sotomayor | Ruth Bader Ginsburg | John Roberts |
GLM 5
z-ai/glm-5
|
Ruth Bader Ginsburg | John Marshall | John Marshall |
GLM 5.1
z-ai/glm-5.1
|
John Marshall | John Roberts | John Roberts |
GPT-4.1
openai/gpt-4.1
|
Justice Ramirez | Morgan Fletcher | Rebecca Hardy |
GPT-4.1 Mini
openai/gpt-4.1-mini
|
Alexandra Diaz | Ruth Bader Ginsburg | Sandra Day O'Connor |
GPT-4.1 Nano
openai/gpt-4.1-nano
|
Amelia | Liam | Justice Smith |
GPT-4o-mini
openai/gpt-4o-mini
|
Justice Taylor | Justice Taylor | Justice Morgan |
GPT-5
openai/gpt-5
|
Eleanor Marshall | Sonia Sotomayor | Margaret Whitaker |
GPT-5 Mini
openai/gpt-5-mini
|
Eleanor Whitfield | Eleanor Harper | Eleanor Whitfield |
GPT-5 Nano
openai/gpt-5-nano
|
Alexandra Reed | Alex Carter | Sonia Sotomayor |
GPT-5.1
openai/gpt-5.1
|
Justice Elena Ramirez | Judge Ramirez | Judge Morales |
GPT-5.2
openai/gpt-5.2
|
Evelyn Hart | Eleanor Hart | Evelyn Carter |
GPT-5.3-Codex
openai/gpt-5.3-codex
|
Sonia | Ruth Bader Ginsburg | Judy Sheindlin |
GPT-5.4
openai/gpt-5.4
|
Solomon | Judge Judy | Solomon |
GPT-5.4 Mini
openai/gpt-5.4-mini
|
Jordan | Taylor | Smith |
GPT-5.4 Nano
openai/gpt-5.4-nano
|
Judge Amelia Thompson | Judge Amelia Grant | Jordan Thompson |
GPT-5.5
openai/gpt-5.5
|
Judith | Sonia Sotomayor | Sonia Sotomayor |
GPT-5.6 Luna
openai/gpt-5.6-luna
|
Sonia Sotomayor | Sonia Sotomayor | John Roberts |
GPT-5.6 Luna Pro
openai/gpt-5.6-luna-pro
|
Sonia Sotomayor | Sonia Sotomayor | Judge Judy |
GPT-5.6 Sol
openai/gpt-5.6-sol
|
Elena | Solomon | Solomon |
GPT-5.6 Terra
openai/gpt-5.6-terra
|
Alexandra Bennett | Avery | Alexandra |
GPT-5.6 Terra Pro
openai/gpt-5.6-terra-pro
|
Alex Morgan | Alexandra Bennett | Alexandra Bennett |
gpt-oss-120b
openai/gpt-oss-120b
|
John | John Roberts | Eleanor Marshall |
gpt-oss-20b
openai/gpt-oss-20b
|
Judge Carter | Arthur Grant | Olivia |
Grok 4.20
x-ai/grok-4.20
|
Ruth | Judge Judy | Judge Judy |
Grok 4.3
x-ai/grok-4.3
|
Hawthorne | Margaret Brennan | Samuel Alito |
Hy3 preview
tencent/hy3-preview
|
Judge Judy | Solomon | Ruth |
Kimi K2.5
moonshotai/kimi-k2.5
|
Victoria Palmer | Margaret Avery | Margaret Whitfield |
Kimi K2.6
moonshotai/kimi-k2.6
|
Margaret Chen | Margaret Chen | Thurgood Marshall |
Kimi K2.7 Code
moonshotai/kimi-k2.7-code
|
Elena Kagan | Elena Vasquez | Margaret Whitmore |
Llama 3.1 8B Instruct
meta-llama/llama-3.1-8b-instruct
|
Eleanor Shaw | Emily Chen | Harrison |
Llama 3.3 70B Instruct
meta-llama/llama-3.3-70b-instruct
|
Aidan Reynolds | Aurelia Jackson | Reginald Wilkinson |
MiMo-V2.5
xiaomi/mimo-v2.5
|
John Roberts | Judge William Harper | John Roberts |
MiniMax M2.5
minimax/minimax-m2.5
|
Alice Martinez | John Doe | John Smith |
MiniMax M2.7
minimax/minimax-m2.7
|
John Smith | MarcusReed | Sarah Mitchell |
Mistral Nemo
mistralai/mistral-nemo
|
Judge | Judge | Judge |
Mistral Small 3.2 24B
mistralai/mistral-small-3.2-24b-instruct
|
Eleanor | Eleanor | Eleanor |
Nemotron 3 Nano 30B A3B
nvidia/nemotron-3-nano-30b-a3b
|
Thomas | Ruth | Thomas |
Nemotron 3 Super
nvidia/nemotron-3-super-120b-a12b
|
Ruth Bader Ginsburg | John Roberts | John Roberts |
o4 Mini
openai/o4-mini
|
Sonia Sotomayor | Elena Ramirez | Avery Sinclair |
Qwen3 235B A22B Instruct 2507
qwen/qwen3-235b-a22b-2507
|
Justice Black | Justice Roberts | Justice Roberts |
Qwen3 30B A3B Instruct 2507
qwen/qwen3-30b-a3b-instruct-2507
|
Justice Elena Kagan | Justice Elena Kagan | Justice Elena Kagan |
Qwen3 Coder Next
qwen/qwen3-coder-next
|
Thurgood Marshall | Thurgood Marshall | Geraldine Hargrove |
Qwen3 Next 80B A3B Instruct
qwen/qwen3-next-80b-a3b-instruct
|
Justice Ruth Bader Ginsburg | Judge Judy | Justice Ruth Bader Ginsburg |
Qwen3.5 397B A17B
qwen/qwen3.5-397b-a17b
|
John Roberts | John Roberts | John Roberts |
Qwen3.5-122B-A10B
qwen/qwen3.5-122b-a10b
|
John Marshall | John Marshall | The user wants a single name for a judge. The output must contain only the name. No punctuation. No explanation. I will select "John Marshall" as he is a famous Chief Justice, fitting the description well. Wait, I should ensure no punctuation is used. "John Marshall" is fine. I will output just the name. Wait, I need to make sure I don't add a period at the end. I will output: JohnJohn Marshall |
Qwen3.5-27B
qwen/qwen3.5-27b
|
John Smith | Robert Johnson | James Wilson |
Qwen3.5-35B-A3B
qwen/qwen3.5-35b-a3b
|
Sterling | John Roberts | Harrison |
Qwen3.5-9B
qwen/qwen3.5-9b
|
John Roberts | John Roberts | John Roberts |
Qwen3.5-Flash
qwen/qwen3.5-flash-02-23
|
John Smith | Jonathan Cross | Marcus Ashford |
Qwen3.6 27B
qwen/qwen3.6-27b
|
Sarah Johnson | Thomas Whitfield | Eleanor Vance |
Qwen3.6 35B A3B
qwen/qwen3.6-35b-a3b
|
David Miller | David Thompson | Marcus Webb |
Qwen3.6 Flash
qwen/qwen3.6-flash
|
John Marshall | Eleanor Vance | Thomas Reed |
Qwen3.6 Plus
qwen/qwen3.6-plus
|
John Roberts | Clarence Thomas | John Roberts |
Qwen3.7 Plus
qwen/qwen3.7-plus
|
Judy Sheindlin | John Roberts | John Roberts |
Dataset last updated July 18, 2026.
About
ModelBias.ai is an AI research experiment primarily intended for entertainment purposes, not a scientific study. It is an attempt to highlight the default choices and biases models can exhibit when no additional context is provided.
Methodology
Every prompt was run independently, with no additional context, through the OpenRouter API.
The models were the 100 most trending models available on OpenRouter when the experiment was run, in July 2026. Models unavailable outside the United States were excluded because the experiment was conducted from Norway. Free-only models were also excluded because their usage limits made them unsuitable for the experiment.
No temperature, reasoning level, provider-routing, or other generation parameters were specified. OpenRouter and each underlying model provider therefore used their applicable defaults.
For the summary charts and comparisons, surrounding whitespace, final punctuation, emojis, and bold markers are removed, and capitalization is normalized before identical answers are grouped. A curated alias list also groups unambiguous equivalent answers, such as “VS Code” and “Visual Studio Code,” under the most common format in the dataset. The downloadable dataset preserves every model's original output.
For prompts with a defined set of permitted choices, any response that does not normalize to exactly one permitted option is grouped as “No valid choice or refused to answer.” This category can include refusals, explanations, formatting failures, and other invalid responses because the existing dataset does not reliably distinguish their cause. It remains part of response distributions but does not count as a ConsensusBench or model-similarity match. Open-ended prompts are not classified this way.
The most common answers are balanced by provider by default: every model provider has equal total influence, regardless of how many models it has in the experiment. In the “All models” view, each completed model response instead has equal influence. Tied answers are shown jointly.
Tech Stack
This project was built with Codex using GPT-5.6 Sol.
Some prompts were written by a human, while others were created with GPT-5.6 Sol.
The backend is built in PHP using the Laravel framework. Prompts were run with the Laravel queue system.
Download the data
The complete dataset is free to download and use in your own project or research.
View and download the dataset on GitHub