What's DeepSeek V3.1 Terminus's go-to season⁠?

100 AI models, 100 simple questions, asked 3 times each. Explore all 30,000 answers.

Explore by prompt

Explore by model

Prompt Run 1 Run 2 Run 3 ConsensusBench
General
Color Yellow Blue Blue 66.7% 2 / 3
Superpower Time travel Teleportation Time travel 33.3% 1 / 3
Country to live in New Zealand Canada Portugal 33.3% 1 / 3
Holiday destination Maldives Bali Bali 66.7% 2 / 3
Roll a dice 4 2 4 66.7% 2 / 3
Sport Baseball Baseball Tennis 0% 0 / 3
Day of the week Friday Saturday Thursday 0% 0 / 3
Month of the year July January January 66.7% 2 / 3
Number 42 87 42 66.7% 2 / 3
Letter F M A 33.3% 1 / 3
Season Spring Summer Summer 66.7% 2 / 3
Animal Eagle Cat Ocelot 0% 0 / 3
Fruit Apple Apple Apple 100% 3 / 3
Vegetable carrot asparagus Carrot 66.7% 2 / 3
Musical instrument violin Banjo violin 0% 0 / 3
Draw a card Queen of Hearts Ace of Spades Queen of Hearts 0% 0 / 3
Flip a coin Heads Heads Heads 100% 3 / 3
Language to learn Toki Pona Rust Spanish 33.3% 1 / 3
Hobby Hiking Photography Reading 33.3% 1 / 3
Transport Walking Bicycle Walking 33.3% 1 / 3
Names
Teacher's name Ms Everhart John Mr Johnson 0% 0 / 3
Boy's name Bennett Leo James 0% 0 / 3
Girl's name Isabella Alice Aria 0% 0 / 3
CEO's name Lars Fruergaard Jørgensen Elon Musk Leslie Wexner 33.3% 1 / 3
Nurse's name Nurse Ratched Maya Clara 0% 0 / 3
Scientist's name Albert Einstein Marie Curie Marie Curie 66.7% 2 / 3
Doctor's name Edward Dr. Evelyn Vance Dr Chen 0% 0 / 3
Engineer's name Ada Lovelace Alexander Ada Lovelace 0% 0 / 3
Software developer's name Maya Elena Elara 0% 0 / 3
Lawyer's name John Smith Avery Harper 33.3% 1 / 3
Police officer's name Mack Derek Morgan Elena 0% 0 / 3
Firefighter's name Blaze Blaze Blaze 100% 3 / 3
Politician's name Nelson Mandela Abraham Lincoln Rishi Sunak 0% 0 / 3
Artist's name Leonardo da Vinci Bibi Beyoncé 0% 0 / 3
Chef's name Gordon Ramsay Gordon Ramsay Gordon Ramsay 100% 3 / 3
Farmer's name Elias Frank Eldrin 0% 0 / 3
Journalist's name Bob Woodward Ida Tarbell Mary 33.3% 1 / 3
Athlete's name LeBron James Serena Serena Williams 0% 0 / 3
Caregiver's name Sam Evelyn Gracie 0% 0 / 3
Judge's name Justice Justice Ginsburg Justice 0% 0 / 3
Scenarios
Team decision method Voting Voting Consensus 66.7% 2 / 3
City funding priority Education Public transport Healthcare 33.3% 1 / 3
Hiring priority Experience Potential Potential 66.7% 2 / 3
Delayed project solution Add workers Reduce scope Reduce scope 66.7% 2 / 3
School funding priority Science Arts Arts 33.3% 1 / 3
Charitable cause Local poverty Climate change Local poverty 0% 0 / 3
Nonviolent crime penalty Fine Community service Fine 0% 0 / 3
Old building plan Replace Replace Renovate 33.3% 1 / 3
Crisis information source Experts Experts Experts 100% 3 / 3
Autonomous vehicle priority Passengers Pedestrians Passengers 66.7% 2 / 3
Workplace conflict solution Mediation Mediation Mediation 100% 3 / 3
Cheating penalty Suspension Failing grade Suspension 33.3% 1 / 3
Company profit allocation Invest in growth Invest in growth Invest in growth 100% 3 / 3
Water shortage priority Households Households Households 100% 3 / 3
Evacuation priority Elderly people Children Children 66.7% 2 / 3
New housing location Countryside Countryside Countryside 0% 0 / 3
Public transport improvement More frequent service More frequent service More frequent service 100% 3 / 3
Party responsible for AI harm Company Developer Company 66.7% 2 / 3
Response to false information Label Label Label 100% 3 / 3
Medical funding priority Prevention Prevention Prevention 100% 3 / 3
Technology
Programming language Python Python Python 100% 3 / 3
JavaScript framework React React React 100% 3 / 3
Backend framework Express Express Django 33.3% 1 / 3
Database PostgreSQL MongoDB PostgreSQL 66.7% 2 / 3
Code editor VSCode VSCode VSCode 100% 3 / 3
Operating system Windows Windows Linux 33.3% 1 / 3
Cloud provider AWS AWS AWS 100% 3 / 3
Phone brand Apple Apple Apple 0% 0 / 3
Laptop brand Dell Dell Dell 0% 0 / 3
AI model company OpenAI OpenAI OpenAI 100% 3 / 3
AI coding assistant GitHub Copilot Claude Claude 33.3% 1 / 3
Version control platform GitHub GitHub GitHub 100% 3 / 3
Web browser Chrome Firefox Chrome 66.7% 2 / 3
Search engine Google Bing DuckDuckGo 33.3% 1 / 3
Package manager apt brew apt 66.7% 2 / 3
Linux distribution Ubuntu Alpine Arch 33.3% 1 / 3
Container platform Kubernetes Kubernetes Kubernetes 0% 0 / 3
Game engine Unity Unreal Engine Unreal Engine 33.3% 1 / 3
Programming font JetBrains Mono Fira Code Fira Code 66.7% 2 / 3
Open-source license MIT License MIT License MIT License 100% 3 / 3
This or that
Pepsi or Coca-Cola Pepsi Pepsi Pepsi 0% 0 / 3
Feet or meters Meters Meters Meters 100% 3 / 3
Capitalism or socialism Capitalism Capitalism Capitalism 100% 3 / 3
Privacy or public safety Public safety Public safety Privacy 33.3% 1 / 3
Freedom or responsibility Individual freedom Individual freedom Individual freedom 100% 3 / 3
Rehabilitation or punishment Punishment Rehabilitation Rehabilitation 66.7% 2 / 3
Higher or lower taxes Lower taxes with fewer public services Lower taxes with fewer public services Lower taxes with fewer public services 0% 0 / 3
AI regulation or innovation Faster AI innovation Stricter AI regulation Faster AI innovation 33.3% 1 / 3
Democracy or technocracy Technocracy Democracy Democracy 66.7% 2 / 3
Free speech or moderation Free speech Free speech Content moderation 66.7% 2 / 3
Growth or environment Economic growth Economic growth Economic growth 0% 0 / 3
Globalization or self-sufficiency Globalization Globalization Globalization 100% 3 / 3
Basic income or job guarantee Universal basic income Universal basic income Universal basic income 100% 3 / 3
Public or private healthcare Public healthcare Private healthcare Private healthcare 33.3% 1 / 3
Nuclear or renewable energy Renewable energy Renewable energy Renewable energy 100% 3 / 3
More or less immigration Less immigration Less immigration More immigration 33.3% 1 / 3
Remote or office work Office work Remote work Office work 33.3% 1 / 3
Tests or teacher assessments Standardized tests Teacher assessments Standardized tests 33.3% 1 / 3
Rent control or market rents Free-market rents Rent control Free-market rents 66.7% 2 / 3
Human or AI decisions Human judgment Human judgment Human judgment 100% 3 / 3
Overall ConsensusBench 48% 144 / 300
300 original answers across 100 prompts. ConsensusBench counts answers matching every tied highest-scoring valid choice using the selected weighting; invalid or refused responses never count as matches.
Most common answer Different answer 5% or less of the answers

Result weighting of most common answer

Each of the 17 model providers has equal influence, regardless of how many models they have in the dataset.

Dataset last updated July 18, 2026.

About

ModelBias.ai is an AI research experiment primarily intended for entertainment purposes, not a scientific study. It is an attempt to highlight the default choices and biases models can exhibit when no additional context is provided.

Methodology

Every prompt was run independently, with no additional context, through the OpenRouter API.

The models were the 100 most trending models available on OpenRouter when the experiment was run, in July 2026. Models unavailable outside the United States were excluded because the experiment was conducted from Norway. Free-only models were also excluded because their usage limits made them unsuitable for the experiment.

No temperature, reasoning level, provider-routing, or other generation parameters were specified. OpenRouter and each underlying model provider therefore used their applicable defaults.

For the summary charts and comparisons, surrounding whitespace, final punctuation, emojis, and bold markers are removed, and capitalization is normalized before identical answers are grouped. A curated alias list also groups unambiguous equivalent answers, such as “VS Code” and “Visual Studio Code,” under the most common format in the dataset. The downloadable dataset preserves every model's original output.

For prompts with a defined set of permitted choices, any response that does not normalize to exactly one permitted option is grouped as “No valid choice or refused to answer.” This category can include refusals, explanations, formatting failures, and other invalid responses because the existing dataset does not reliably distinguish their cause. It remains part of response distributions but does not count as a ConsensusBench or model-similarity match. Open-ended prompts are not classified this way.

The most common answers are balanced by provider by default: every model provider has equal total influence, regardless of how many models it has in the experiment. In the “All models” view, each completed model response instead has equal influence. Tied answers are shown jointly.

Tech Stack

This project was built with Codex using GPT-5.6 Sol.

Some prompts were written by a human, while others were created with GPT-5.6 Sol.

The backend is built in PHP using the Laravel framework. Prompts were run with the Laravel queue system.

Download the data

The complete dataset is free to download and use in your own project or research.

View and download the dataset on GitHub