KBAISE/ for lovable
Directory

LM Studio

Local model desktop app

Free tier + paidModel APIs, routers & observabilitylmstudio.ai

What it is

A desktop GUI for running open models locally on Mac, Windows and Linux, built on llama.cpp and MLX. It includes a model browser, chat UI and a local OpenAI-compatible server, plus an lms CLI and Python and TypeScript SDKs for wiring apps to it.

Why it earns a slot

The friendly version of local inference — it tells you which quantisation fits your RAM before you download 40GB. Good for comparing open models side by side, then serving the winner to your app on localhost.

Know before you commit

GUI-first design makes it awkward to run headless on a server — use Ollama or vLLM for that. Commercial use inside a company may require the separate enterprise licence.

Field notes

0 from the community

Sign in to add a field note. Reading is always open.

Loading notes…

Also in Model APIs, routers & observability

Anthropic API / Claude Platform
platform.claude.comNot yet swept

Frontier model provider API

PaidModel APIs, routers & observability
Per token. Opus 5 $5/$25 per M in/out; Sonnet 5 $2/$10; Haiku 4.5 $1/$5; Fable 5 and Mythos 5 $10/$50. Batch API 50% off; cache reads 0.1x input.
Open dossier
Braintrust
braintrust.devNot yet swept

LLM eval and experiment platform

Free tier + paidModel APIs, routers & observability
Starter free ($10/mo model credits, 1GB data, 10k scores, 14-day retention), Pro $249/mo ($249 credits, 5GB, 50k scores), Enterprise custom. Overage $3-4/GB and $1.50-$2.50 per 1k scores.
Open dossier
Fal.ai
fal.aiNot yet swept

Generative media inference host

PaidModel APIs, routers & observability
Per output or per second for media models: images ~$0.02-$0.04 at 1MP, video $0.05-$0.40 per second. Serverless GPU rates $1.10-$8.50/hr with committed-use discounts.
Open dossier
Fireworks AI
fireworks.aiNot yet swept

Open-model inference host

Free tier + paidModel APIs, routers & observability
Per token serverless across Standard/Priority/Fast tiers; embeddings $0.008-$0.10/M. On-demand GPUs billed per second: H100/H200 ~$7-8/hr, B200 ~$10-13/hr. Fine-tuning $0.50-$40 per M training tokens. $1 free credit.
Open dossier