Minicpm5 Deploy Llama Cpp

Run MiniCPM5-1B or MiniCPM5-2B with llama.cpp using the released GGUF artifacts (F16 / Q8_0 / Q4_K_M). Use when the user wants CPU-only / consumer-GPU / cross-platform native deployment, asks for "llama.cpp", "llama-cli", "llama-server", "GGUF", or has no Python available.

OpenBMB fea681a 3.5 KB Updated

File contents

openbmb/minicpm/tree/main/skills/minicpm5-deploy-llama-cpp commit fea681a556

Frequently asked questions

npx skillmds@latest add openbmb/minicpm5-deploy-llama-cpp