The standing public eval of how frontier AI models handle Islamophobia: frozen prompts, every transcript published, an open harness anyone can rerun, and an MCP server for pre-ship testing.
**Which AI models handle Islamophobia worst.** <p align="center"> <img src="docs/assets/leaderboard.svg" alt="Rancor leaderboard: clean rate per model from the first graded run" width="100%"> </p> **Live leaderboard: <https://rancor.litai.ca>** <p align="center"> <a href="https://youtu.be/rR_gQwp5Mm8">