Relay Gen Model Switching

Reduce inference cost by dynamically switching from large to small LLMs during reasoning generation. Large model handles demanding reasoning phases; small model completes consolidation and answer stages triggered by discourse cues. Achieves 2.2× speedup with minimal accuracy loss.

adu2021 a307d1d 8.5 KB Updated

File contents

adu2021/skillxiv/tree/main/skills/skillxiv-v0.0.2-claude-opus-4.6/relay-gen-model-switching commit a307d1d9c6

Frequently asked questions

npx skillmds@latest add adu2021/relay-gen-model-switching