Together Fireworks

Use when calling open-weight LLMs on Together AI or Fireworks AI's OpenAI-compatible endpoints — `base_url` plus namespaced model id, the cheapest model that clears the bar, per-1M-token cost math, serverless vs batch vs dedicated. NOT renting GPUs to self-host weights (that is `runpod`), NOT running a model locally for free (that is `ollama`).

ericrisco a550b67 6 files · 33.9 KB Updated

File contents

ericrisco/rsc-harness/tree/main/skills/together-fireworks commit a550b67cb5

Frequently asked questions

npx skillmds@latest add ericrisco/together-fireworks