MODEL CHOICE LAB

What do you want to do?

Choose a task for a shortlist of models with reasons, reviews and caveats.

Your priority?

Editorial choices informed by tests and reviewed user reports · 2026-09-08. Candidates for your trial task, not a universal ranking.

Everyday tasks 2

GoogleLower-cost option

Gemini 3.8 Flash

Start with low reasoning for short tasks; increase it when verification matters.

API input / output per 1M tokens
$0.75 / $3.75
Access
Google AI Studio and Gemini API · GA

ConsiderQuota experiences conflict. Output speed is not time to a finished answer.

Evidence behind this choice1
r/google_antigravity · user discussion

Praise for design and complex edits, alongside complaints about latency and quota use. Commenters report different experiences.

OpenAIComplex work

GPT-5.6 Sol

For requests with several constraints and follow-up verification.

API input / output per 1M tokens
$4 / $20
Access
ChatGPT, Codex and API

ConsiderMaximum effort and the listed API tier are excessive for a short translation.

Evidence behind this choice1
Artificial Analysis · independent evaluation

Astra is stronger in terminal use and automation; Fable 5.1 in AA-Briefcase and SciCode. Both round to 53 on v4.3.

Compare on your own material

Give two candidates identical inputs. Check facts and voice for writing, tests and change size for coding. Measure time to an accepted result and total task cost. Prices are AA API snapshots; cache, long context, taxes and subscriptions differ.

Open benchmarks