aaai-artifact-evaluati…
Use when packaging AAAI code, data, multimedia appendices, technical appendices, reproducibility evidence, and post-acceptance artifact releases without…
Use when appraising the cumulative evidence and ensuring balance in an Academy of Management Annals (Annals) review — weighing conflicting findings by credibility, steelmanning rival schools, and handling the author's own work even-handedly. Audits evidence quality and fairness;
$ npx -y skills add brycewang-stanford/Awesome-Journal-Skills --skill amann-evidence-standards --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/amann-evidence-standardsContext preview
The summary Claude sees to decide when to auto-load this skill.
Use when appraising the cumulative evidence and ensuring balance in an Academy of Management Annals (Annals) review — weighing conflicting findings by credibility, steelmanning rival schools, and handling the author's own work even-handedly. Audits evidence quality and fairness;
name: amann-evidence-standards description: Use when appraising the cumulative evidence and ensuring balance in an Academy of Management Annals (Annals) review — weighing conflicting findings by credibility, steelmanning rival schools, and handling the author's own work even-handedly. Audits evidence quality and fairness; it does not design the framework (amann-organizing-framework) or build exhibits (amann-tables-figures).
An Annals review carries no estimates of its own — but it is not therefore neutral. The "attitude" is **disciplined critical appraisal**: you judge how good the cumulative evidence is, reconcile conflicts by credibility, and state where the field's confidence is and is not warranted. This is the review-craft replacement for primary-research robustness checks: you are appraising *other people's* designs, not defending your own.
A review must be **comprehensive in coverage** yet **selective in emphasis** — long (~50 pages) but not an inventory. Resolve the tension by **tiering** the corpus:
| Tier | Treatment | |------|-----------| | **Foundational / field-defining** | discussed in text — what it established and its limits | | **Important contributions** | grouped and weighed within the framework's cells; cited with their finding | | **Confirmatory / incremental** | cited in clusters ("see also …") to show coverage without bloating prose | | **Tangential** | cited only where it bears on a specific claim |
Comprehensiveness is proven by the *citation set* (the saturation log from `amann-literature-synthesis`); selectivity is exercised in the *prose*. Equal-length summaries of every paper abdicate the editorial judgment that is the review's value.
Management findings conflict constantly. Reconcile them by **why they differ**, never by tally:
Occasionally a credibility judgment cannot be settled by reading alone: the review's account of a controversy hinges on whether a staggered-adoption TWFE estimate survives modern corrections, or an apparent consensus may melt once publication bias is priced in. When a load-bearing magnitude comes with a replication package, audit it rather than adjudicate by prose — the shared playbook [`execution-with-mcp`](../../../shared-resources/empirical-methods/execution-with-mcp.md) maps each design family to callable StatsPAI / Stata MCP tools (`bacon_decomposition` to expose bad-comparison weighting, `callaway_santanna` to re-estimate, `honest_did_from_result` for pre-trend fragility). Any number produced this way must come from an actual run and be labelled as the review's own re-analysis. This is the exception, not the Annals default: most appraisal here stays qualitative.
Annals is the **review-of-the-field**: its account of a debate becomes the field's shared reference, and **the surveyed authors often referee the review**. Balance is therefore both ethical and strategic.
Stanford REAP × CoPaper.AI · 由斯坦福实证方法论团队精选与维护 访问 copaper.ai 微信:CoPaper.AI 按 11 个主流学科板块覆盖 经管与商科 社会科学 人文学科 数学与物理科学 生命科学 医学与健康 工程与技术 计算机科学与 AI 体育科学 点击任一学科名可跳转到对应说明;每类下的代表子领域在正文总览中完整列出。下方封面墙按 venue 导航,完整分类见覆盖一览。 🧭 布局指南 · 📚 Skill Pack 一览 · ⚡ 如何使用 · 🧪 自动实证
Use when packaging AAAI code, data, multimedia appendices, technical appendices, reproducibility evidence, and post-acceptance artifact releases without…
Use when drafting an AAAI author response (rebuttal) under the single short character-limited author-feedback window, the no-URL rule, no-new-results guidance,…
Use when preparing an accepted AAAI paper for camera-ready source submission to AAAI Press, including proceedings page limits, two-column template compliance,…
Use when designing or auditing AAAI experiments for the broad-AI program committee, including baselines, ablations, statistical significance, robustness, human…
Use when positioning an AAAI paper's novelty against archival work, contemporaneous arXiv or workshop papers, and AAAI/IJCAI/NeurIPS/ICML/ICLR neighbors across…
Use when strengthening an AAAI paper's reproducibility checklist (placed after references), experimental traceability, seed and hyperparameter reporting,…