Directory
Ollama
Local model runner
Free tier + paidModel APIs, routers & observabilityollama.com
What it is
An open-source runtime that downloads and serves open-weight models locally with one command, exposing both its own and an OpenAI-compatible API on localhost. A newer hosted cloud tier runs larger models on Ollama's servers through the same client.
Why it earns a slot
Zero marginal cost and zero data egress — point your app at
localhost:11434 and iterate on prompts all day for free. Ideal for offline dev, privacy-sensitive features and CI tests that must not call a paid API.Know before you commit
Local quality is capped by your RAM — a 7B quantised model is not Claude or GPT, and prompts tuned on it often break when moved to a frontier model. Cloud models do leave your machine.
Getting started
curl -fsSL https://ollama.com/install.sh | sh
Neighbourhood
How Ollama sits against the rest of the atlas — swaps, companions and the tools builders actually save alongside it.
Also in Model APIs, routers & observability
Anthropic API / Claude Platform
platform.claude.comNot yet swept
Frontier model provider API
PaidModel APIs, routers & observability
Per token. Opus 5 $5/$25 per M in/out; Sonnet 5 $2/$10; Haiku 4.5 $1/$5; Fable 5 and Mythos 5 $10/$50. Batch API 50% off; cache reads 0.1x input.
Open dossier
Braintrust
braintrust.devNot yet swept
LLM eval and experiment platform
Free tier + paidModel APIs, routers & observability
Starter free ($10/mo model credits, 1GB data, 10k scores, 14-day retention), Pro $249/mo ($249 credits, 5GB, 50k scores), Enterprise custom. Overage $3-4/GB and $1.50-$2.50 per 1k scores.
Open dossier
Fal.ai
fal.aiNot yet swept
Generative media inference host
PaidModel APIs, routers & observability
Per output or per second for media models: images ~$0.02-$0.04 at 1MP, video $0.05-$0.40 per second. Serverless GPU rates $1.10-$8.50/hr with committed-use discounts.
Open dossier
Fireworks AI
fireworks.aiNot yet swept
Open-model inference host
Free tier + paidModel APIs, routers & observability
Per token serverless across Standard/Priority/Fast tiers; embeddings $0.008-$0.10/M. On-demand GPUs billed per second: H100/H200 ~$7-8/hr, B200 ~$10-13/hr. Fine-tuning $0.50-$40 per M training tokens. $1 free credit.
Open dossier