Skip to content

Which model should you pick?

Every AI tool now hands you a model menu. The names are noise. This page sorts the models I actually use by the job you are trying to do, and tells you where you will find each one.

Direct

The architect tier: plan the work, design the approach, then direct cheaper models to do it, picking the right one for each piece. Right now Claude Opus does this job for almost everything; Fable only adds a little edge on very big, long-running tasks. I expect that to flip again with the next Fable release.

3 models

Think

Strong reasoning for hard problems: strategy, analysis, the calls where the quality of the thinking decides everything downstream. Claude Opus is my default here and for almost everything else. GPT-6 belongs here at its Sol tier.

3 models · Why "think expensive, execute cheap" →

Execute

The workhorses. Drafting, coding, formatting, summarising, agents that run all day. I rarely pick these by hand anymore: they do their best work as subagents, when a stronger model hands them a focused job. GPT-6 comes in two tiers here, Sol for the harder work and Luna for volume.

8 models

Anthropic

Claude Sonnet

Fabian's pick

Focused coding, drafting, and research jobs handed to it as a subagent by Opus or Fable, and high-volume API work.

Context
1M tokens
Price
$2 / $10
Speed
Fast
In Claude, Poe, Perplexity, Claude Code
OpenAI

GPT-6

Tiers: Sol · Luna

Execution work at scale: Sol when the task is hard or mistakes are expensive, Luna for high-volume work where speed and price matter most.

Context
1.05M tokens
Price
Sol $2 / $10 · Luna $0.10 / $0.50
Speed
Fast
In ChatGPT, OpenAI Codex, Poe, OpenRouter
Zhipu AI (Z.ai)

GLM-5.3-Flash

Running an assistant or agent that works all day on tool calls, where cost per task matters more than personality.

Context
1M tokens
Price
$0.15 / $0.50
Speed
Fast
In OpenClaw, OpenRouter
Alibaba Cloud

Qwen3.7-Max

Complex, multi-step agentic coding tasks.

Context
1M tokens
Price
~$0.65 / $3.25
Speed
Moderate
In Poe, OpenRouter
DeepSeek

DeepSeek V4 Flash

High-volume reasoning tasks where cost efficiency is critical.

Context
1M tokens
Price
$0.14 / $0.28
Speed
Fast
In OpenRouter
Google

Gemini Flash

High-volume execution work: coding, agents, copy, and anything where speed and price beat peak intelligence.

Context
1M tokens
Price
$0.75 / $3.75
Speed
Instant
In Gemini, Poe, Google AI Studio, Antigravity
xAI

Grok 4.7

Running inside Grok Bot, and reading what people on X are saying about a topic right now.

Context
500K tokens
Price
$2 / $6
Speed
Moderate
In Grok Bot, Poe, OpenRouter
Anthropic

Claude Haiku

A subagent that searches large document collections for a stronger model.

Context
200K tokens
Price
$1 / $5
Speed
Instant
In Claude, Poe

Run privately on your own machine

Open-weights models small enough for a laptop or a workstation. Nothing leaves the building. Weaker than the cloud tier, and that is the trade.

4 models

Make images

No context window here, and prices per image or per month. What matters is what each one is good at.

5 models

Make video

Two entries so far, both priced per second of footage.

2 models

Voice & music

Speech that carries emotion, and full songs with a structure you control.

2 models

Missing a model you rely on? Tell me and I will put it through its paces.