The bench

Every seat has a name and a model.

Dr Moot does not hide which model answered. Each seat is pinned to a specific model on each plan, the same models the evaluations were run on. This page is the whole list.

Light

Solo

One model, asked once. Light answers with the chair seat and does not convene a panel.

ChairAnthropic logo
Claude Sonnet 4.6anthropic/claude-sonnet-4.6

Weighs the arguments still standing and delivers one verdict.

Reads text, image, pdf · knows to 31 August 2025

Configured, not convened

Light answers with one model, so these seats do not sit. They are the panel Light would convene if council modes reached this plan: DeepSeek V4 Pro (generator), GLM 5.2 (sceptic), Mistral Large 3 (specialist).

Plus

Solo, Review, Debate, Moot

The full panel. Four seats, three labs, one verdict.

GeneratorOpenAI logo
GPT-5.6 Terraopenai/gpt-5.6-terra

Commits to the strongest complete answer and states the assumptions it rests on.

Reads text, image, pdf · knows to 16 February 2026

ScepticAnthropic logo
Claude Sonnet 5anthropic/claude-sonnet-5

Hunts factual errors, unsupported assumptions and the gaps that would change the answer.

Reads text, image, pdf · knows to 31 January 2026

SpecialistxAI logo
Grok 4.5xai/grok-4.5

Checks facts, figures, definitions and sources, separating evidence from inference.

Reads text, image, pdf

ChairAnthropic logo
Claude Sonnet 5anthropic/claude-sonnet-5

Weighs the arguments still standing and delivers one verdict.

Reads text, image, pdf · knows to 31 January 2026

Pro

Solo, Review, Debate, Moot, Studio

The same shape, at the frontier of each lab's line-up.

GeneratorOpenAI logo
GPT-5.6 Solopenai/gpt-5.6-sol

Commits to the strongest complete answer and states the assumptions it rests on.

Reads text, image, pdf · knows to 16 February 2026

ScepticAnthropic logo
Claude Sonnet 5anthropic/claude-sonnet-5

Hunts factual errors, unsupported assumptions and the gaps that would change the answer.

Reads text, image, pdf · knows to 31 January 2026

SpecialistxAI logo
Grok 4.5xai/grok-4.5

Checks facts, figures, definitions and sources, separating evidence from inference.

Reads text, image, pdf

ChairAnthropic logo
Claude Opus 4.8anthropic/claude-opus-4.8

Weighs the arguments still standing and delivers one verdict.

Reads text, image, pdf · knows to January 2026

Where each seat shows up

The mode decides who sits.

The models above do not all run on every question. The mode you choose decides which seats are convened.

Solo
Chair answers directly.
Review
Generator drafts and revises; Chair critiques and judges.
Debate
Generator argues the Affirmative, Sceptic the Negative; Chair decides.
Moot
Generator proposes, Sceptic challenges, Specialist checks; Chair rules.

The specifications

What each model brings.

Context is how much of your question, documents and the panel’s own working a model can hold at once. Latency is how long it waits before speaking, throughput how fast it speaks once it starts.

CategoryCapabilities
GPT-5.6 TerraOpenAI
plus / generator1.05M1.8s147tps
  • Extended reasoning (to max)
  • Reads images
  • Reads PDFs and documents
  • Calls tools
  • Searches the web
GLM 5.2Z.ai
light / sceptic1M2.6s139tps
  • Extended reasoning (to xhigh)
  • Calls tools
Claude Sonnet 5Anthropic
plus, pro / sceptic, chair1M3.4s112tps
  • Extended reasoning (to xhigh)
  • Reads images
  • Reads PDFs and documents
  • Calls tools
  • Searches the web
Claude Opus 4.8Anthropic
pro / chair1M3.5s92tps
  • Extended reasoning (to xhigh)
  • Reads images
  • Reads PDFs and documents
  • Calls tools
  • Searches the web
GPT-5.6 SolOpenAI
pro / generator1.05M4.2s70tps
  • Extended reasoning (to max)
  • Reads images
  • Reads PDFs and documents
  • Calls tools
  • Searches the web
DeepSeek V4 ProDeepSeek
light / generator1.05M2.5s69tps
  • Extended reasoning (to xhigh)
  • Calls tools
Grok 4.5xAI
plus, pro / specialist500K1.6s59tps
  • Extended reasoning (to high)
  • Reads images
  • Reads PDFs and documents
  • Calls tools
  • Searches the web
Claude Sonnet 4.6Anthropic
light / chair1M1.5s56tps
  • Extended reasoning (to high)
  • Reads images
  • Reads PDFs and documents
  • Calls tools
  • Searches the web
Mistral Large 3Mistral
light / specialist256K0.6s56tps
  • Reads images
  • Calls tools

Context and capabilities come from the Vercel AI Gateway model catalogue. Latency is p50 time to first token and throughput is p50 output tokens per second, both as published by Vercel from live traffic on the same gateway every Dr Moot run goes through. Both sources were read on 12 August 2026; they are measurements of a live system and they move. A dash means nothing is published, which we prefer to an estimate. Speed is not quality: the slowest seat on this page is often the one that catches the error.

Straight from the makers

Every lab says how its model should be driven.

Each maker publishes its own guidance on how to ask - what to spell out, what to leave alone, where it wants structure instead of prose. We apply those conventions strictly, so a seat holds its persona whichever model is sitting in it: one job per seat, carried in that seat’s own system prompt and restated every time it speaks; ballots and critiques returned as structure rather than prose about structure; and anything a model cannot accept stripped before the call instead of left for it to trip over.

Line-ups change.

Seats move when the evidence says they should - a seat that times out or reviews the debate instead of answering the question gets replaced. When a seat changes, this page changes with it. Several models agreeing is not proof that they are right, and Dr Moot does not present it as such.

Questions about a specific seat? Ask us.