Skip to the list
whichmodel

Updated

1 of 6 sources stale

Which one do I point my harness at?

Best value

Nothing beats it on both score and price.

5 of 8 models · ordered by cost per task · as of Jul 11, 2026

C

Qwen3 72B

Alibaba·$0.32/M tokens

$0.02

/task

A

DeepSeek V4

DeepSeek·$0.85/M tokens

$0.03

/task

A

Claude Sonnet 4.5

Anthropic·$11.40/M tokens

$0.21

/task

S

Gemini 3 Pro

Google·$9/M tokens

$0.28

/task

S

Claude Opus 4.6

Anthropic·$19/M tokens

$0.42

/task

Everything else, ranked

Works with

Capability score 0-100 · cuts from weights 2026.08.1

measured on Aider Polyglot, as of Jul 11, 2026

  1. S80-1003

    Gemini 3 ProBest value

    Google

    88.7

    $0.28

    $9/M tokens·Tool use 89·Gemini CLI, Cursor, Aider

    GPT-5.1

    OpenAI

    87.9

    $0.39

    $15.20/M tokens·Tool use 85·Codex CLI, Cursor, Aider

    Claude Opus 4.6Best value

    Anthropic

    91.4

    $0.42

    $19/M tokens·Tool use 86·Claude Code, Cursor, Aider

    No model in this list works with that one yet.

  2. A65-792

    DeepSeek V4Best value

    DeepSeek

    58.4

    $0.03

    $0.85/M tokens·Tool use 54·Aider, Cursor

    Manual call

    Manual call BA

    Benchmarks place it in B, but at 1/20th the price it is the only credible self-host path. Promoted with the reason on the record.

    Claude Sonnet 4.5Best value

    Anthropic

    78.2

    $0.21

    $11.40/M tokens·Tool use 74·Claude Code, Cursor, Aider

    No model in this list works with that one yet.

  3. B45-642

    Mistral Large 3

    Mistral

    51.7

    $0.11

    $4.80/M tokens·Tool use 52·Aider

    Grok 4.1

    xAI

    58.3

    No data

    $11.40/M tokens·Tool use No data·Cursor

    No model in this list works with that one yet.

  4. C0-441

    Qwen3 72BBest value

    Alibaba

    38.9

    $0.02

    $0.32/M tokens·Tool use 42·Aider

    No model in this list works with that one yet.

Not on this list meta/llama-4.2-405b, cohere/command-r-plus-2
Named in a source, not matched to a model moonshot/kimi-k3 (terminal_bench), Qwen3-Max-Preview (livebench)