Truefoundry LLM Deploy

Deploys ML and LLM models on TrueFoundry with GPU inference servers (vLLM, TGI, NVIDIA NIM). Uses YAML manifests with `tfy apply`. Use when serving language models, deploying Hugging Face models, or hosting GPU-accelerated inference endpoints.

truefoundry 3c2d04e 18 files · 159.9 KB Updated

File contents

truefoundry/tfy-deploy-skills/tree/main/skills/llm-deploy commit 3c2d04e505

Frequently asked questions

npx skillmds@latest add truefoundry/truefoundry-llm-deploy