Intel Neural Speed Gguf Inference

Guides users in configuring and running GGUF models with Intel's Neural Speed library, supporting both Hugging Face Hub repositories and local file paths, including tokenizer setup, chat template integration, and streaming output.

ECNU-ICALK Updated 559 repo stars

File contents

ECNU-ICALK/AutoSkill/tree/main/SkillBank/ConvSkill/english_gpt4_8_GLM4.7/intel_neural_speed_gguf_inference commit c3e4df06f4

Frequently asked questions

npx skillmds@latest add ecnu-icalk/intel-neural-speed-gguf-inference