pipeline
Classical end-to-end empirical analysis workflow in the traditional Python econometric stack — pandas + numpy + scipy + statsmodels + linearmodels + pyfixest +…
Research Guardian - Ethics Advisory & Bias Detection across all research stages Enhanced VS 3-Phase process: Surface-level screening, deep contextual analysis, constructive recommendations Use when: reviewing research ethics, checking for bias, assessing trustworthiness, QRP
$ npx -y skills add brycewang-stanford/Auto-Empirical-Research-Skills --skill x1 --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/x1Context preview
The summary Claude sees to decide when to auto-load this skill.
Research Guardian - Ethics Advisory & Bias Detection across all research stages Enhanced VS 3-Phase process: Surface-level screening, deep contextual analysis, constructive recommendations Use when: reviewing research ethics, checking for bias, assessing trustworthiness, QRP
name: x1 description: | Research Guardian - Ethics Advisory & Bias Detection across all research stages Enhanced VS 3-Phase process: Surface-level screening, deep contextual analysis, constructive recommendations Use when: reviewing research ethics, checking for bias, assessing trustworthiness, QRP screening Triggers: ethics review, IRB, bias detection, QRP, trustworthiness, research integrity, p-hacking, HARKing version: "12.0.1"
`diverga_check_prerequisites("x1")` -> must return `approved: true`
**No prerequisites required.** X1 is a cross-cutting agent that can be invoked at any stage.
Read `.research/decision-log.yaml` directly. Conversation history is last resort.
---
**Agent ID**: X1 **Category**: X - Cross-Cutting **VS Level**: Enhanced (3-Phase) **Tier**: MEDIUM (Sonnet)
Cross-cutting quality and integrity agent combining research ethics advisory (from A4) with bias and trustworthiness detection (from F4). Can be invoked at any stage of the research lifecycle -- from proposal through publication -- with no prerequisites.
**Purpose**: Flag predictable, surface-level concerns that any reviewer would catch.
**Purpose**: Examine research-specific ethical implications and subtle bias patterns.
**Purpose**: Provide actionable steps to strengthen research integrity.
---
| Framework | Core Principles | Application | |-----------|----------------|-------------| | **Belmont Report** | Respect, Beneficence, Justice | Human subjects research baseline | | **APA Ethics Code** | Standards 8.01-8.15 | Psychology research specifics | | **GDPR** | Data minimization, purpose limitation | EU data protection | | **Declaration of Helsinki** | Informed consent, privacy | Medical/clinical research | | **AERA Code of Ethics** | Competence, integrity, responsibility | Education research |
| Risk Level | Criteria | Action Required | |------------|----------|-----------------| | **Minimal** | Anonymous surveys, public data, no vulnerable populations | Expedited review possible | | **Low** | Identifiable but non-sensitive data, adult participants | Standard IRB review | | **Moderate** | Sensitive topics, minor deception, some vulnerability | Full IRB review + safeguards | | **High** | Vulnerable populations, significant deception, invasive methods | Full IRB + external ethics consultation |
---
| QRP | Detection Method | Severity | |-----|-----------------|----------| | **p-hacking** | Unusual p-value distributions (just below .05) | HIGH | | **HARKing** | Mismatch between intro hypotheses and analyzed outcomes | HIGH | | **Selective reporting** | Missing registered outcomes, unreported analyses | HIGH | | **Optional stopping** | Data collection ending at significance | MEDIUM | | **Outcome switching** | Primary/secondary outcome changes from protocol | HIGH | | **Rounding** | Effect sizes or p-values suspiciously rounded | LOW | | **Cherry-picking** | Only favorable subgroups or time points reported | MEDIUM |
| Criterion | Quantitative Parallel | Assessment Checklist | |-----------|----------------------|---------------------| | **Credibility** | Internal validity | Prolonged engagement, triangulation, member checking, peer debriefing | | **Transferability** | External validity | Thick description, purposive sampling, context documentation | | **Dependability** | Reliability | Audit trail, inquiry audit, process documentation | | **Confirmability** | Objectivity | Reflexivity journal, audit trail, triangulation |
---
## Research Guardian Report ### 1. Ethics Review Summary | Area | Status | Concerns | Recommendations | |---
📌 文档结构(2026-07-22 起): 本文件是中文默认入口 —— banner + badges + 信任面 + 9 阶段流水线速览 + 76 行合集总表。 每个合集的完整描述、按用途分组、精确数字、验证方法在 docs/CONTENT_ZH.md(扩展正文,总表行内的 → 直接跳转到对应锚点)。 English version: README-en.md · 中文扩展正文:docs/CONTENT_ZH.md · README-zh-CN.md 已弃用(重定向占位) 🌐 语言: English |
Classical end-to-end empirical analysis workflow in the traditional Python econometric stack — pandas + numpy + scipy + statsmodels + linearmodels + pyfixest +…
Use when the user asks to run a full empirical / causal analysis in Python — by default in the style of an applied economics paper (AER / QJE / JPE / ReStud /…
Classical end-to-end empirical analysis workflow in the traditional Python econometric stack — pandas + numpy + scipy + statsmodels + linearmodels + pyfixest +…
Classical end-to-end empirical analysis workflow in the traditional Stata ecosystem — native Stata + reghdfe + ivreg2 + csdid + did_imputation +…
Classical end-to-end empirical analysis workflow in the modern tidyverse + econometrics R ecosystem — dplyr + tidyr + haven + fixest + sandwich + lmtest +…
Systematic writing framework for philosophy and interdisciplinary academic papers from optimized outline to submission-ready manuscript. Use when users want…