LLM Instruction Finetuning

How to fine-tune a pre-trained LLM to follow instructions and respond to tasks like a chatbot. Use this skill whenever the user wants to train an LLM on instruction-response pairs, format datasets for instruction tuning, evaluate fine-tuned model responses, or understand the complete instruction fine-tuning workflow. Make sure to use this skill when users mention instruction tuning, chatbot training, Alpaca format, Phi-3 format, or any scenario where they need to make an LLM respond to specific prompts rather than just generate text.

abelrguezr Updated

File contents

abelrguezr/hacktricks-skills/tree/main/skills/AI/AI-llm-architecture/7.2.-fine-tuning-to-follow-instructions commit 6734f8a210

Frequently asked questions

npx skillmds@latest add abelrguezr/llm-instruction-finetuning