PodcastIntelligence Snacks · Weekly conversations about AI, software and business, from the team behind Normie Mode.Listen →
Tests explained · People's votes

People's votes on legal questions

People compared two anonymous answers to legal and government questions.

Run by

LMArena

How to read it

A rating from thousands of head-to-head votes. Higher means people preferred its answers more often. The thin line is the likely range.

Used for

1 task on this site.

How the models did

Rating from blind votes · Higher is better

Claude Fable 5.1 and Claude Opus 5.5 are neck and neck at the top, followed by Kimi K3 and Gemini 3.8 Flash.

Claude Fable 5.1 · Anthropic1st
Claude Opus 5.5 · Anthropic2nd
Kimi K3 · Moonshot AI3rd
Gemini 3.8 Flash · Google DeepMind4th
Gemini 3.1 Pro · Google DeepMind5th
Gemini 3.6 Flash · Google DeepMind6th
GPT-5.6 Sol · OpenAI7th
DeepSeek V4 Pro · DeepSeek8th
GPT-6 Astra · OpenAI9th
Claude Sonnet 5 · Anthropic10th
GPT-5.6 Luna · OpenAI11th
DeepSeek V4.1 Flash · DeepSeek12th
Gemini 3.5 Flash-Lite · Google DeepMind13th
Claude Haiku 4.5 · Anthropic14th
GPT-6 Luna · OpenAINot tested yet
GPT-6 Sol · OpenAINot tested yet
Source: LMArena · votes as of 25 Sept 2026

Where we use it

Each task page labels this test by how closely it matches the task.

TaskHow close
Best AI for legal questionsTests this task
Intelligence Snacks newsletter

The big AI ideas each week, in your inbox