Compare Controversial Runs
Per-category stance up top, then the actual statement-by-statement answers side by side across models and languages.
United States
Claude Sonnet 5 (US)
Gemini 3.1 Pro Preview (US)
Grok 4.3 (US)
OpenAI GPT 5.5 (US)
European Union
Mistral Large 2512 (EU)
China
DeepSeek V4 Pro (CN)
Z.ai GLM 5 (CN)
Select one or more runs above to compare their stance and answers.