# Stata Data Provenance

> Track dataset lineage, transformation steps, merge logic, and reproducibility risks in Stata workflows. Use when the user needs to explain where data came from, how it changed, or why a pipeline can be trusted.

- Skill: `tmonk/stata-data-provenance` (Agent Skill, multi-file: 4 files)
- Install (CLI): `npx skillmds@latest add tmonk/stata-data-provenance`
- Raw SKILL.md: https://api.skillmd.com/api/skills/tmonk/stata-data-provenance/raw
- Safety review: pending
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: DevOps & Infra
- Author: tmonk (https://skillmd.com/u/tmonk)
- Updated: 2026-09-17
- Page: https://skillmd.com/skills/tmonk/stata-data-provenance

---


# Data Provenance

Use this skill when lineage and reproducibility matter.

1. Map the sequence of source files and transformations.
2. Flag untracked merges, overwrites, and silent sample restrictions.
3. Produce a concise provenance narrative a coauthor can audit.

Read `references/lineage.md` for the provenance checklist.

