Compare text, image, speech, music and video models by price, cost per task, licence and features. Each one is reviewed by hand, with its real drawbacks listed. Updated every two weeks.
73 models tracked. Last updated 2 Oct 2026. Next update 16 Oct 2026.
What one job costslower is better
The 8 cheapest of 14 models with a published price. Bars are logarithmic so the cheap end stays readable; every row carries the exact figure.
Claude Opus 4.5
Anthropic · LanguageStrongest of the Claude line for long reasoning and code. Priced accordingly — use Sonnet unless you actually need the extra depth.
200K context
Quality —
Claude Sonnet 4.6
Anthropic · LanguageThe sensible default: most of Opus's quality at a fraction of the cost.
1M context
Quality —
GPT-5.5
OpenAI · LanguageHuge context window and strong general reasoning. Output tokens cost six times input, so verbose prompting gets expensive fast.
1.1M context
Quality —
GPT-5 mini
OpenAI · LanguageThe cheap tier for classification and cleanup work that doesn't need a frontier model.
272K context
Quality —
Gemini 2.5 Pro
Google · LanguageA million-token context window makes it the pick for whole-codebase or whole-document work.
1M context
Quality —
Gemini 2.5 Flash
Google · LanguageBuilt for speed and volume rather than depth. Good for bulk tasks, weaker at hard reasoning.
1M context
Quality —
DeepSeek V4
DeepSeek · LanguageNear-frontier quality at a fraction of Western API pricing, and the weights are public. Check where your data goes before putting client work through the hosted API.
131K context
Quality —
Grok 4
xAI · LanguageCompetitive on reasoning benchmarks. Smaller ecosystem and fewer integrations than the big three.
Quality —
Mistral Large
Mistral AI · LanguageEU-hosted, which matters if data residency is a hard requirement for you.
262K context
Quality —
GPT Image 1
OpenAI · ImageBest-in-class at following a detailed prompt and rendering legible text in an image.
Quality —
DALL·E 3
OpenAI · ImageFlat per-image pricing makes budgeting simple. Superseded by GPT Image 1 on quality.
Quality —
FLUX.1 dev
Black Forest Labs · ImageThe open-weights favourite — the most liked image model on Hugging Face. Non-commercial licence on the dev variant; read it before client work.
Quality —
Stable Diffusion XL
Stability AI · ImageRuns on consumer hardware and has the deepest ecosystem of LoRAs and fine-tunes. Older, and weaker than current models on prompt adherence.
Quality —
Midjourney
Midjourney · ImageStill the best pure aesthetic quality out of the box. No free tier, and no real API — it is a subscription product, not infrastructure.
Quality —
Veo 3.1
Google · VideoGenerates native audio alongside video. Billed per second of output, which adds up quickly on iteration.
Quality —
Veo 2
Google · VideoCheaper per second than 3.1 and fine for silent b-roll and motion tests.
Quality —
Runway Gen-4
Runway · VideoThe most complete editing suite around the model, not just raw generation. Credit-based pricing is hard to predict month to month.
Quality —
HunyuanVideo
Tencent · VideoThe leading open-weights video model. Needs a serious GPU — this is not a laptop workload.
Quality —
ElevenLabs v3
ElevenLabs · Speech & VoiceThe quality benchmark for synthetic voice and cloning. Billed per character, so long-form narration costs real money.
Quality —
Whisper
OpenAI · Speech & VoiceTranscription, not synthesis. Open weights mean you can run it locally for free instead of paying per minute.
Quality —
Kokoro 82M
Hexgrad · Speech & VoiceTiny enough to run on a laptop and startlingly good for its size. Fewer voices and languages than the commercial options.
Quality —
Suno
Suno · Music & AudioFull songs with vocals from a text prompt. Commercial rights only on paid tiers — check before using anything in client work.
Quality —
Stable Audio Open
Stability AI · Music & AudioOpen weights, trained on licensed audio, which makes provenance far less murky. Built for samples and loops rather than finished songs.
Quality —
ACE-Step
ACE Studio · Music & AudioFast open-weights music generation you can self-host. Rougher output than Suno.
Quality —
gpt-6-sol
openai · LanguageQuality —
gpt-6-luna
openai · LanguageQuality —
gpt-6-astra
openai · LanguageQuality —
gpt-6.1-sol
openai · LanguageQuality —
xai/grok-4.7
xai · LanguageQuality —
gpt-5.5-cyber
openai · LanguageQuality —
gpt-image-2.5-flare
openai · ImageQuality —
gpt-image-2.5-sunburst
openai · ImageQuality —
xai/grok-imagine-video
xai · VideoQuality —
xai/grok-imagine-video-1.5
xai · VideoQuality —
xai/grok-voice-transcribe-1.0
xai · Speech & VoiceQuality —
xai/grok-voice-transcribe-2.0
xai · Speech & VoiceQuality —
DeepSeek-R1
deepseek-ai · Languagemit
Quality —
Llama-3.1-8B-Instruct
meta-llama · Languagellama3.1
Quality —
Meta-Llama-3-8B
meta-llama · Languagellama3
Quality —
DeepSeek-V4-Pro
deepseek-ai · Languagemit
Quality —
gpt-oss-120b
openai · Languageapache-2.0
Quality —
Meta-Llama-3-8B-Instruct
meta-llama · Languagellama3
Quality —
GLM-5.2
zai-org · Languagemit
Quality —
gpt-oss-20b
openai · Languageapache-2.0
Quality —
stable-diffusion-xl-base-1.0
stabilityai · Imageopenrail++
Quality —
stable-diffusion-v1-4
CompVis · Imagecreativeml-openrail-m
Quality —
FLUX.1-schnell
black-forest-labs · Imageapache-2.0
Quality —
Z-Image-Turbo
Tongyi-MAI · Imageapache-2.0
Quality —
stable-diffusion-3-medium
stabilityai · Imageother
Quality —
stable-diffusion-3.5-large
stabilityai · Imageother
Quality —
OrangeMixs
WarriorMama777 · Imagecreativeml-openrail-m
Quality —
Sulphur-2-base
SulphurAI · VideoQuality —
Wan2.1-T2V-14B
Wan-AI · Videoapache-2.0
Quality —
mochi-1-preview
genmo · Videoapache-2.0
Quality —
MiniMax-H3-Turbo-Lora
larryvrh · Videoapache-2.0
Quality —
HunyuanVideo-1.5
tencent · Videoother
Quality —
AnimateDiff-Lightning
ByteDance · Videocreativeml-openrail-m
Quality —
pyramid-flow-sd3
rain1011 · Videoother
Quality —
XTTS-v2
coqui · Speech & Voiceother
Quality —
Dia-1.6B
nari-labs · Speech & Voiceapache-2.0
Quality —
VibeVoice-1.5B
microsoft · Speech & Voicemit
Quality —
csm-1b
sesame · Speech & Voiceapache-2.0
Quality —
Qwen3-TTS-12Hz-1.7B-CustomVoice
Qwen · Speech & Voiceapache-2.0
Quality —
chatterbox
ResembleAI · Speech & Voicemit
Quality —
VoxCPM2
openbmb · Speech & Voiceapache-2.0
Quality —
stable-audio-open-1.0
stabilityai · Music & Audioother
Quality —
ChatTTS
2Noise · Music & Audiocc-by-nc-4.0
Quality —
MiniMax-Music3
MiniMaxAI · Music & AudioQuality —
YuE2-3B
m-a-p · Music & Audiocc-by-nc-4.0
Quality —
Ace-Step1.5
ACE-Step · Music & Audiomit
Quality —
ACE-Step-v1-3.5B
ACE-Step · Music & Audioapache-2.0
Quality —
riffusion-model-v1
riffusion · Music & Audiocreativeml-openrail-m
Quality —
musicgen-large
facebook · Music & Audiocc-by-nc-4.0
Quality —
Remove a filter, or submit the model you are looking for.
Prices checked hourly against the LiteLLM dataset · last updated 1 second ago. API rates, not seat subscriptions. Watching 320 models for new releases — 49 added in the last 30 days, marked not reviewed yet.
Don’t see a model?
Submit it with its pricing page. We review and add it within a few days.
More free software, every discipline: open source tools.