AI Bias Observatory for India
One prompt goes in. Six demographic variants come out. We measure where the model changes its mind — and hand you the receipts.
Live pipeline
streamingPrompt
donesanitised + stored
Gemma baseline
donegemma-4-31b-it generation
Variant generator
done6 demographic rewrites
Semantic divergence
donelexical delta vs baseline
Behavioural analysis
runningrefusal + tone drift
Threat score
queuedTES compositor
Leaderboard
queuedranked + archived
Divergence by axis
averaged across recent runs
0
variants evaluated
0
prompts submitted
0.0
mean threat score
0
critical findings
Four instruments, one observatory. Everything you see runs on the same evaluation trace.
Gender, region, caste, religion, language and income rewrites of your prompt — semantically identical, demographically different.
Embedding-space distance between the neutral baseline and each variant, normalised against the model's own noise floor.
A single 0–100 composite you can argue about, backed by a per-layer breakdown you can't.
Where the failures cluster geographically, aggregated from every public submission on the board.
how it works
The whole trace is public. If you don't trust the score, read the pipeline.
Prompt
donesanitised + stored
Gemma baseline
donegemma-4-31b-it generation
Variant generator
done6 demographic rewrites
Semantic divergence
donelexical delta vs baseline
Behavioural analysis
donerefusal + tone drift
Threat score
doneTES compositor
Leaderboard
runningranked + archived
Signal density
hover a hotspot
—
no signals recorded yet
meet TES
A single judge model is a single point of bias. So we stacked three orthogonal layers and made the last one blind.
01
Every demographic variant is compared against the neutral baseline in code. We measure drift, not vibes.
lexical Δ
02
Refusals, hedging, harshness and length are diffed across variants — where bias hides when the words look polite.
refusal ratio
03
Gemma critiques its own outputs under a blind rubric, so the judge never sees which demographic it is scoring.
blind rubric