skill
Full-empirical-analysis-skill
Classical end-to-end empirical analysis workflow in the traditional Python econometric stack — pandas + numpy + scipy + statsmodels + linearmodels + pyfixest + rdrobust + econml + causalml + matplotlib/seaborn. **Defaults to economics empirical-paper style** (AER / QJE / AEJ) — every run produces a publication-ready output set with a multi-column regression table (M1→M6 progressive controls/FE) as the centerpiece, plus Table 1 (descriptives), mechanism / heterogeneity / robustness tables, and event-study + coefficient + trend figures. Covers the full 8-step pipeline an applied economist or quantitative social scientist runs on every paper — (1) data cleaning, (2) variable construction & transformation, (3) descriptive statistics & Table 1, (4) statistical diagnostic tests, (5) baseline empirical modeling, (6) robustness battery, (7) further analysis (mechanism, heterogeneity, mediation, moderation), (8) publication-ready tables & figures. **Also covers two parallel domain modes that share the same 8-step scaffolding** — **Mode A — Epidemiology / public health** (target-trial emulation via `zepid` / hand-rolled `pandas`, IPTW + g-formula + TMLE doubly-robust triplet via `zepid` / `econml` / `lifelines`, Mendelian randomization via `pymr` / `mrtool` (or `rpy2` → `MendelianRandomization`/`TwoSampleMR`), KM / AFT / Cox survival via `lifelines`, E-value sensitivity, principal stratification — STROBE / TRIPOD reporting), and **Mode B — ML causal inference** (DML via `econml.dml` / `doubleml`, S/T/X/R/DR meta-learners via `econml.metalearners` / `causalml`, causal forest via `econml.grf` / `causalml`, Dragonnet / TARNet / CEVAE neural causal via `causalml`, BCF via `pymc-bart` / `bcf-py`, matrix completion, CATE distribution + policy tree via `econml.policy` / `policytree-py`, off-policy evaluation, conformal causal via `mapie`, fairness audit via `fairlearn`, DAG learning via `causal-learn` / `cdt` / LLM-assisted). Prescribes which library to reach for at each step, shows the canonical code, and links to deeper `references/` files for variant-specific patterns. Use when the user asks for a **complete empirical analysis** in Python, wants to replicate an applied-economics paper from scratch, needs a reproducible workflow that is NOT opinionated on any single vertical package (contrast with StatsPAI), wants explicit control over every estimator and diagnostic, or asks "how do I write a full empirical pipeline in Python?". Also triggers when the user names a specific classical step in isolation — "winsorize at 1/99%", "run Breusch-Pagan", "build a Table 1 balance table", "do a placebo test", "event study plot", "mediation analysis" — and wants it wired into the broader pipeline. Mode A triggers on "target trial emulation", "IPTW", "TMLE", "Mendelian randomization", "STROBE", "公共健康", "流行病学". Mode B triggers on "DML", "double machine learning", "causal forest", "meta-learner", "Dragonnet", "BCF", "policy tree", "conformal causal", "fairness audit", "因果机器学习".
Description as published by the maintainer. Source
- active
active — Most recent push to the repository was 2026-08-06.
Signals
These are separate measurements of different things. They are deliberately not combined into one score, because a popularity number that mixes website traffic with saves and stars cannot be checked or acted on.
| Signal | Value | What it measures | Window | Observed | Source |
|---|---|---|---|---|---|
| GitHub stars | 3,289 | Stars on the repository that contains this skill, not on the skill itself. A collection of fifty skills shares one number, so it says nothing about this particular skill. | cumulative, all time | GitHub | |
| Last commit | 2026-08-06 | Most recent push to the containing repository. It may reflect work on a different skill in the same collection. | point in time | GitHub | |
| repository status | active | The repository holding this skill exists and is not archived. | as of fetch | GitHub |
Will this work with your setup?
No harness stated by the author and no install path convention detected. Compatibility is untested.
We have not run this skill against a task with and without it enabled, so we cannot tell you whether it improves anything, what it costs in tokens, or whether it duplicates behaviour your harness already has. When we have run that test, the result will appear on this page with the task, the versions and the budget it used.
The skill definition lives at plugins/empirical-analysis-python/skills/pipeline/SKILL.md in https://github.com/brycewang-stanford/Auto-Empirical-Research-Skills.
Where to get it
Related, by what their authors tagged them
-
brycewang-stanford-auto-empirical-research-skills-academic-paper-composer
— last commit 2026-08-06, shares academic-research, awesome-list, communication
Systematic writing framework for philosophy and interdisciplinary academic papers from optimized outline to submission-…
-
brycewang-stanford-auto-empirical-research-skills-academic-paper-strategist
— last commit 2026-08-06, shares academic-research, awesome-list, communication
Systematic strategic planning framework for philosophy and interdisciplinary academic papers targeting preprint platfor…
-
brycewang-stanford-auto-empirical-research-skills-answering-research-questions
— last commit 2026-08-06, shares academic-research, awesome-list, communication
Main orchestration workflow for systematic literature research - search, evaluate, traverse, synthesize
-
brycewang-stanford-auto-empirical-research-skills-auto-empirical-research-skills
— last commit 2026-08-06, shares academic-research, awesome-list, communication
Route empirical-research requests through the Auto-Empirical Research Skills catalog when this whole repository is inst…
-
brycewang-stanford-auto-empirical-research-skills-building-paper-screening-rubri
— last commit 2026-08-06, shares academic-research, awesome-list, communication
Collaboratively build and refine paper screening rubrics through brainstorming, test-driven development, and iterative…
-
brycewang-stanford-auto-empirical-research-skills-checking-chembl-for-structured
— last commit 2026-08-06, shares academic-research, awesome-list, communication
Check if medicinal chemistry papers are in ChEMBL database to access curated bioactivity data
-
brycewang-stanford-auto-empirical-research-skills-citation-management
— last commit 2026-08-06, shares academic-research, awesome-list, communication
Comprehensive citation management for academic research. Search Google Scholar and PubMed for papers, extract accurate…
-
brycewang-stanford-auto-empirical-research-skills-full-empirical-analysis-skill-59df41
— last commit 2026-08-06, shares academic-research, awesome-list, communication
Classical end-to-end empirical analysis workflow in the modern tidyverse + econometrics R ecosystem — dplyr + tidyr + h…
-
brycewang-stanford-auto-empirical-research-skills-full-empirical-analysis-skill-5b6c79
— last commit 2026-08-06, shares academic-research, awesome-list, communication
Classical end-to-end empirical analysis workflow in the traditional Python econometric stack — pandas + numpy + scipy +…
-
brycewang-stanford-auto-empirical-research-skills-full-empirical-analysis-skill-78bf74
— last commit 2026-08-06, shares academic-research, awesome-list, communication
Classical end-to-end empirical analysis workflow in the traditional Stata ecosystem — native Stata + reghdfe + ivreg2 +…
These share tags the maintainers applied themselves, such as academic-research, awesome-list, communication, copaper. Common tags like "mcp" or "ai" are ignored for this: agreeing with six hundred other projects is not a similarity.
This is not a recommendation and not a test result. It is a map of what the authors said their work is about.
Also from brycewang-stanford
-
brycewang-stanford-auto-empirical-research-skills-full-empirical-analysis-skill-a37523
— last commit 2026-08-06
Classical end-to-end empirical analysis workflow in the modern tidyverse + econometrics R ecosystem — dplyr + tidyr + h…
-
brycewang-stanford-auto-empirical-research-skills-full-empirical-analysis-skill-af5eb1
— last commit 2026-08-06
Classical end-to-end empirical analysis workflow in the traditional Stata ecosystem — native Stata + reghdfe + ivreg2 +…
-
brycewang-stanford-auto-empirical-research-skills-getting-started-with-research
— last commit 2026-08-06
Introduction to literature search & review skills - systematic paper finding, screening, extraction, and citation trave…
-
brycewang-stanford-auto-empirical-research-skills-hypothesis-generation
— last commit 2026-08-06
Structured hypothesis formulation from observations. Use when you have experimental observations or data and need to fo…
-
brycewang-stanford-auto-empirical-research-skills-hypothesis-generation-80cd32
— last commit 2026-08-06
Generate testable hypotheses. Formulate from observations, design experiments, explore competing explanations, develop…
-
brycewang-stanford-auto-empirical-research-skills-literature-review
— last commit 2026-08-06
Conduct comprehensive, systematic literature reviews using multiple academic databases (PubMed, arXiv, bioRxiv, Semanti…
-
brycewang-stanford-auto-empirical-research-skills-literature-review-3e1315
— last commit 2026-08-06
Conduct comprehensive, systematic literature reviews using multiple academic databases (PubMed, arXiv, bioRxiv, Semanti…
-
brycewang-stanford-auto-empirical-research-skills-paper-slide-deck
— last commit 2026-08-06
Generate professional slide deck images from academic papers and content. Creates comprehensive outlines with style ins…
-
brycewang-stanford-auto-empirical-research-skills-peer-review
— last commit 2026-08-06
Systematic peer review toolkit. Evaluate methodology, statistics, design, reproducibility, ethics, figure integrity, re…
-
brycewang-stanford-auto-empirical-research-skills-research-grants
— last commit 2026-08-06
Write competitive research proposals for NSF, NIH, DOE, DARPA, and Taiwan NSTC. Agency-specific formatting, review crit…
How the author describes it
Topics the maintainer set on GitHub: academic-research, agent-skills, ai-agent, awesome-list, communication, copaper, economics, education, empirical-research, international-relations, political-science, psychology, public-administration, reproducible-research, skills-library, social-science, sociology.
This record as data
Every field on this page, with its source and observation date, is in the catalog JSON. Fetch the whole kind at once instead of parsing this HTML.
GET /api/v1/entries/skill.json