Skip to content
PayingForAI Test Lab By OverpayingForAI
MENU

One flagship per provider

Providers

Every provider represented in the lab's fights, with its tested flagship and record.

Alibaba Qwen

Alibaba: Qwen3 Max

failed

qwen/qwen3-max

Runs
0/3 accepted
Median landed cost

Anthropic

Anthropic: Claude Opus 4.5

failed

anthropic/claude-opus-4.5

Runs
0/3 accepted
Median landed cost

Cohere

Cohere: Command A

failed

cohere/command-a

Runs
0/3 accepted
Median landed cost

DeepSeek

DeepSeek: DeepSeek V3.2

failed

deepseek/deepseek-v3.2

Runs
0/3 accepted
Median landed cost

Google

Google: Gemini 3.1 Pro (Preview)

accepted

google/gemini-3.1-pro-preview

Runs
3/3 accepted
Median landed cost
$6.38

Meta

Meta: Llama 4 Maverick

failed

meta-llama/llama-4-maverick

Runs
0/3 accepted
Median landed cost

Mistral AI

Mistral: Mistral Large 2512

failed

mistralai/mistral-large-2512

Runs
0/3 accepted
Median landed cost

Moonshot AI

MoonshotAI: Kimi K2.5

accepted

moonshotai/kimi-k2.5

Runs
2/3 accepted
Median landed cost
$6.26

OpenAI

OpenAI: GPT-5.1

failed

openai/gpt-5.1

Runs
0/3 accepted
Median landed cost

xAI

xAI: Grok 4.5

accepted

x-ai/grok-4.5

Runs
3/3 accepted
Median landed cost
$6.26