Omni Moe Atomic Experts

Scale mixture-of-experts models efficiently by decomposing experts into atomic vector pairs with Cartesian product routing and expert-centric scheduling. Achieves 10.9× speedup and 50% fewer parameters versus fine-grained baselines through system-algorithm codesign that converts scattered memory access into contiguous batched operations.

adu2021 Updated

File contents

adu2021/skillxiv/tree/main/skills/skillxiv-v0.0.2-claude-opus-4.6/omni-moe-atomic-experts commit 4ac0ce92e7

Frequently asked questions

npx skillmds@latest add adu2021/omni-moe-atomic-experts