aaai-artifact-evaluati…
Use when packaging AAAI code, data, multimedia appendices, technical appendices, reproducibility evidence, and post-acceptance artifact releases without…
Use when shaping the prose and structure of an ASE (IEEE/ACM Automated Software Engineering) research paper, covering the automation-first first-page arc, stating the automated task precisely, keeping the tool runnable and the model-swap test in mind, threats-as-argument, and
$ npx -y skills add brycewang-stanford/Awesome-Journal-Skills --skill ase-writing-style --agent claude-codeHow it fires
How this skill gets triggered: by you, by Claude, or both.
/ase-writing-styleContext preview
The summary Claude sees to decide when to auto-load this skill.
Use when shaping the prose and structure of an ASE (IEEE/ACM Automated Software Engineering) research paper, covering the automation-first first-page arc, stating the automated task precisely, keeping the tool runnable and the model-swap test in mind, threats-as-argument, and
name: ase-writing-style description: Use when shaping the prose and structure of an ASE (IEEE/ACM Automated Software Engineering) research paper, covering the automation-first first-page arc, stating the automated task precisely, keeping the tool runnable and the model-swap test in mind, threats-as-argument, and the 10+2 page discipline on the ACM acmart template.
Write the paper so an automated-SE reviewer sees, on the first page, **what task you automate, how you automate it, and that it runs on real subjects**. ASE rewards a clearly stated automation with evidence proportional to the claim — not a systems win, not a leaderboard, and not a broad finding that would read better at FSE. The worked example in `resources/worked-examples/01-introduction.md` shows the arc before → after.
Lead with the automation, in this order:
1. **The automatable task** — a software-engineering task the reader recognizes (detect X, generate Y, repair Z, comprehend W), named precisely enough that a reviewer knows what "success" is. 2. **Why current automation (or manual practice) is inadequate** — what existing tools do *not* do, stated as a gap, not a literature tour. 3. **The contribution as an automation design** — the technique (analysis, generator, synthesizer, repair, learned model) and, ideally, the **tool** that embodies it. 4. **Evidence on real subjects** — real systems, credible *tool* baselines, and the metric that matches the task (not a proxy). 5. **What it automates for practice + threats posture** — the payoff, with the central threat named where it lives, not deferred to a closing paragraph.
Put the automation and the first evidence within the first three pages; the early-rejection gate means a weak opening can end the process before rebuttal.
code quality" is not a task; "given a flaky test, synthesize a patch that makes it deterministic without weakening its assertions" is.
what it **produces** — reviewers map this straight to feasibility.
If a learned component is involved, write so the **automation design is the contribution**, not the model. Report an **ablation** that isolates the learned part from the analysis/oracle, and phrase claims so the software-engineering lesson survives a model swap. A paper whose lesson evaporates when the model changes reads as an ML re-route (see `ase-topic-selection`).
hold the property?), not just similarity to a reference.
Pair every claim in the abstract with a table or figure it points to. Match evidence to claim shape; see `ase-experiments`.
Argue construct, internal, and external validity where they arise. For automated-SE tools the usual suspects: the **oracle** (how do you know a "repair" is correct?), **subject selection** (are the systems representative or self-selected?), **baseline fairness** (equal budgets/tuning?), and **overfitting** to the evaluation set. Name the residual threat plainly and bound it (an audited subsample, a held-out subject set) rather than reciting a checklist.
references only. The mandatory **Data Availability Statement** after Conclusions counts inside the 10 pages.
configs, extra tables, proofs) to the artifact, but nothing that *decides acceptance* may live outside the body (see `ase-supplementary`).
per-root-cause breakdown), not by decorating it.
vague ones like *leverages* or *explores*.
| Failure | Why it hurts at ASE | Fix | |---|---|---| | Model/leaderboard framing | Reads as ML, not automated SE | Foreground the automation design; add the ablation | | Vague task statement | Reviewer cannot judge success | Give input/output/success criterion in ¶1 | | Proxy-metric evaluation | Evidence does not match a repair/synthesis claim | Verify the produced artifact directly (re-run/oracle) | | Threats as a closing recital | Construct/oracle validity never engaged | Argue each threat where it arises; bound it | | Body over 10 pages | Desk-reject-grade | Move reproducible detail to the artifact |
[Task] input -> output -> success criterion (one sentence, on page 1?) [Automation-first arc] task / inadequacy / technique+tool / real-subject evidence / payoff+threats — all present? [Model-swap] ablation isolating the learned component present? lesson survives a model swap? [Evidence-claim pairing] each abstract claim -> a table/figure with a matching metric [Budget] content pages / reference pages / Data Availability inside 10pp?
Stanford REAP × CoPaper.AI · 由斯坦福实证方法论团队精选与维护 访问 copaper.ai 微信:CoPaper.AI 按 11 个主流学科板块覆盖 经管与商科 社会科学 人文学科 数学与物理科学 生命科学 医学与健康 工程与技术 计算机科学与 AI 体育科学 点击任一学科名可跳转到对应说明;每类下的代表子领域在正文总览中完整列出。下方封面墙按 venue 导航,完整分类见覆盖一览。 🧭 布局指南 · 📚 Skill Pack 一览 · ⚡ 如何使用 · 🧪 自动实证
Use when packaging AAAI code, data, multimedia appendices, technical appendices, reproducibility evidence, and post-acceptance artifact releases without…
Use when drafting an AAAI author response (rebuttal) under the single short character-limited author-feedback window, the no-URL rule, no-new-results guidance,…
Use when preparing an accepted AAAI paper for camera-ready source submission to AAAI Press, including proceedings page limits, two-column template compliance,…
Use when designing or auditing AAAI experiments for the broad-AI program committee, including baselines, ablations, statistical significance, robustness, human…
Use when positioning an AAAI paper's novelty against archival work, contemporaneous arXiv or workshop papers, and AAAI/IJCAI/NeurIPS/ICML/ICLR neighbors across…
Use when strengthening an AAAI paper's reproducibility checklist (placed after references), experimental traceability, seed and hyperparameter reporting,…