Track dataset lineage, transformation steps, merge logic, and reproducibility risks in Stata workflows. Use when the user needs to explain where data came from, how it changed, or why a pipeline can be trusted.
Scanned 9/3/2026
Install to Claude Code
npx -y skills add brycewang-stanford/Auto-Empirical-Research-Skills --skill stata-data-provenance --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Stata Data Provenance?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/brycewang-stanford-stata-data-provenance)More formats (shields.io, HTML) on the badges page.
---
name: stata-data-provenance
description: Track dataset lineage, transformation steps, merge logic, and reproducibility risks in Stata workflows. Use when the user needs to explain where data came from, how it changed, or why a pipeline can be trusted.
---
# Data Provenance
Use this skill when lineage and reproducibility matter.
1. Map the sequence of source files and transformations.
2. Flag untracked merges, overwrites, and silent sample restrictions.
3. Produce a concise provenance narrative a coauthor can audit.
Read `references/lineage.md` for the provenance checklist.
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!