Text2QA

build concept, process, and case-application supervision datasets from markdown books or long markdown documents. use when generating training data from many .md files or precomputed chunk files and when full chunk coverage, resumable batch processing, status tracking, validation, and coverage auditing are required. use for book-to-sft pipelines where every chunk must end in exactly one final status and where answer-only qa is not sufficient because the dataset should also teach grounded reasoning patterns and rule application.

opendcai 240b8e1 8 files · 53.6 KB Updated

File contents

opendcai/data-preparation-bench/tree/main/md_to_qa/SKILL commit 240b8e1fc8

Frequently asked questions

npx skillmds@latest add opendcai/text2qa