Compare models on the same prompt
Send one prompt to every configured model and compare the answers side by side, along with what each one cost and how long it took.
Roster
| Model | Provider | Default seats | Speed | Tier | State |
|---|---|---|---|---|---|
openai/gpt-5.5 | openai | Claims | balanced | premium | ready |
openai/gpt-5.6-sol | openai | deliberate | premium | ready | |
openai/gpt-4.1 | openai | balanced | standard | ready | |
openai/gpt-4o-mini | openai | fast | economy | ready | |
anthropic/claude-opus-4.8 | anthropic | deliberate | premium | ready | |
anthropic/claude-sonnet-4.6 | anthropic | balanced | standard | ready | |
google/gemini-3.1-flash-lite | fast | economy | ready | ||
google/gemini-2.5-flash | Pro sideCon sideJudgeFraming & evidence | fast | economy | ready | |
deepseek/deepseek-v4-pro | deepseek | deliberate | standard | ready | |
deepseek/deepseek-v4-flash | deepseek | fast | economy | ready | |
x-ai/grok-4.5 | xai | balanced | premium | ready | |
z-ai/glm-5.2 | zai | balanced | standard | ready |