Ollama
One-command local LLM runtime.
OPEN SOURCECPU-CAPABLE
OpenAI's cheapest tier — $0.20/$1.25 for bulk workloads.
GPT-5.4 Nano is OpenAI's high-volume budget model: $0.20/$1.25 per 1M tokens with a 400K context window. It targets classification, routing, extraction and other bulk tasks where per-call cost dominates.
Bulk pipelines: tagging, moderation, routing, short summaries.
This is exactly the tier where local models win — a 7–14B model on Ollama does most Nano jobs at zero marginal cost.
One-command local LLM runtime.