Directory
LM Studio
Local model desktop app
Free tier + paidModel APIs, routers & observabilitylmstudio.ai
What it is
A desktop GUI for running open models locally on Mac, Windows and Linux, built on llama.cpp and MLX. It includes a model browser, chat UI and a local OpenAI-compatible server, plus an
lms CLI and Python and TypeScript SDKs for wiring apps to it.Why it earns a slot
The friendly version of local inference — it tells you which quantisation fits your RAM before you download 40GB. Good for comparing open models side by side, then serving the winner to your app on localhost.
Know before you commit
GUI-first design makes it awkward to run headless on a server — use Ollama or vLLM for that. Commercial use inside a company may require the separate enterprise licence.
Neighbourhood
How LM Studio sits against the rest of the atlas — swaps, companions and the tools builders actually save alongside it.
Also in Model APIs, routers & observability
Anthropic API / Claude Platform
platform.claude.comNot yet swept
Frontier model provider API
PaidModel APIs, routers & observability
Per token. Opus 5 $5/$25 per M in/out; Sonnet 5 $2/$10; Haiku 4.5 $1/$5; Fable 5 and Mythos 5 $10/$50. Batch API 50% off; cache reads 0.1x input.
Open dossier
Braintrust
braintrust.devNot yet swept
LLM eval and experiment platform
Free tier + paidModel APIs, routers & observability
Starter free ($10/mo model credits, 1GB data, 10k scores, 14-day retention), Pro $249/mo ($249 credits, 5GB, 50k scores), Enterprise custom. Overage $3-4/GB and $1.50-$2.50 per 1k scores.
Open dossier
Fal.ai
fal.aiNot yet swept
Generative media inference host
PaidModel APIs, routers & observability
Per output or per second for media models: images ~$0.02-$0.04 at 1MP, video $0.05-$0.40 per second. Serverless GPU rates $1.10-$8.50/hr with committed-use discounts.
Open dossier
Fireworks AI
fireworks.aiNot yet swept
Open-model inference host
Free tier + paidModel APIs, routers & observability
Per token serverless across Standard/Priority/Fast tiers; embeddings $0.008-$0.10/M. On-demand GPUs billed per second: H100/H200 ~$7-8/hr, B200 ~$10-13/hr. Fine-tuning $0.50-$40 per M training tokens. $1 free credit.
Open dossier