GenAI · Bias in AI

What did AI learn about us?

Before a language model can use a word, it turns it into a list of numbers — a word embedding — learned from billions of human sentences. That is what lets AI understand language. But human text is full of stereotypes, so the numbers soak them up too. Below, each word is placed by the bias its embedding carries. Pick a lens and watch.

Bias lens
Show words
Test a word
Tap any word to see all its biases · farther from centre = stronger bias
◄ neutral ►
NEUTRAL

Why does AI pick up bias?

Three ideas behind what you just saw.

🧮

Words become numbers

Each word is a vector learned so that words used in similar contexts land close together. "King" and "queen", "Paris" and "London" cluster. This is the magic that makes modern AI understand language.

📚

Learned from us

The vectors come from how people actually write. If text pairs "nurse" with "she" and "engineer" with "he", the model learns that link — as fact, not as the stereotype it is.

📏

We can measure it

Take the direction from he → she. Project any word onto it and you get a gender score. That one number is what positions every word on the chart above.

⚠️ These associations are not true and not endorsed — they are stereotypes the model absorbed from biased text. Showing them is the first step to catching and fixing them.

Why it matters — and can we fix it?

📄

Résumé screening

A model that ranks CVs can quietly down-rank women for "engineer" roles — Amazon scrapped exactly such a tool in 2018.

🌐

Translation

From a gender-neutral language, AI often returns "he is a doctor, she is a nurse" — inventing gender from bias.

🖼️

Image & text generation

Ask for "a CEO" or "a criminal" and generators lean on the same stereotypes, at massive scale.

🛠️

Can we fix it?

Partly. We can neutralize the bias direction (try the toggle), curate better data, and audit outputs — but none of it is a full cure. Human oversight stays essential.

Data: precomputed GloVe bias percentiles from WordBias (Ghai et al., IEEE VIS 2021). Scores are relative positions within this small demo vocabulary, for teaching — not a precise measurement of any individual word.