diff options
| author | sillylaird <sillyfanboy@gmail.com> | 2026-09-03 00:33:59 +0000 |
|---|---|---|
| committer | sillylaird <sillyfanboy@gmail.com> | 2026-09-03 00:33:59 +0000 |
| commit | 898b52edcb47bcb3e9d6106e74ca73e74ea01e70 (patch) | |
| tree | 85c6ee5ad58b860144551184d4cf86b560c62b91 /.agents/skills/results-report/references | |
| download | www-898b52edcb47bcb3e9d6106e74ca73e74ea01e70.tar.gz www-898b52edcb47bcb3e9d6106e74ca73e74ea01e70.zip | |
Diffstat (limited to '.agents/skills/results-report/references')
6 files changed, 194 insertions, 0 deletions
diff --git a/.agents/skills/results-report/references/EVIDENCE-PROPAGATION.md b/.agents/skills/results-report/references/EVIDENCE-PROPAGATION.md new file mode 100644 index 0000000..bd46910 --- /dev/null +++ b/.agents/skills/results-report/references/EVIDENCE-PROPAGATION.md @@ -0,0 +1,25 @@ +# Evidence Propagation + +Use this file to keep `results-analysis` outputs aligned with the final report. + +## Mapping rule + +- `analysis-report.md` -> main findings and narrative summary +- `stats-appendix.md` -> test choice, uncertainty, effect size, correction rule +- `figure-catalog.md` -> figure purpose and per-figure interpretation scaffolding +- figure files -> visual evidence cited in `Figure-by-Figure Interpretation` + +## Minimum statistical carry-over + +Every strong claim in a results report should preserve: +- sample size or run/seed count, +- metric definition, +- uncertainty summary, +- test name, +- effect size when relevant, +- multiple-comparison handling when relevant. + +## Unsupported claim rule + +If the analysis bundle does not support a claim strongly enough, keep the claim tentative and say why. +Do not upgrade a suggestive result into a decisive conclusion during report writing. diff --git a/.agents/skills/results-report/references/decision-oriented-analysis.md b/.agents/skills/results-report/references/decision-oriented-analysis.md new file mode 100644 index 0000000..a63a0c3 --- /dev/null +++ b/.agents/skills/results-report/references/decision-oriented-analysis.md @@ -0,0 +1,17 @@ +# Decision-Oriented Analysis + +The purpose of a post-experiment report is not only to record what happened. +It should change the project's next decision. + +## Required final questions +- What should stop? +- What should continue? +- What should be tested next? +- What should be promoted into a durable result note? +- What, if anything, is ready for manuscript-facing writing? + +## Good closing pattern +- “This round supports X.” +- “It does not yet resolve Y.” +- “The main blocker is Z.” +- “Therefore the next concrete action is A.” diff --git a/.agents/skills/results-report/references/figure-interpretation.md b/.agents/skills/results-report/references/figure-interpretation.md new file mode 100644 index 0000000..1156d74 --- /dev/null +++ b/.agents/skills/results-report/references/figure-interpretation.md @@ -0,0 +1,17 @@ +# Figure Interpretation in Results Reports + +A results report should not dump figures. + +For each major figure, write four blocks: +- **Why this figure exists** +- **What to notice** +- **What interpretation is supported** +- **What this changes in the project decision** + +## Example micro-structure + +### Figure X +- Purpose: compare adapter vs freezing under the same transfer setting. +- Observation: adapter improves mean WER and reduces variance. +- Interpretation: subject-specific adaptation likely resolves part of the transfer mismatch. +- Decision implication: prioritize adapter ablations before expanding frozen-only variants. diff --git a/.agents/skills/results-report/references/report-naming.md b/.agents/skills/results-report/references/report-naming.md new file mode 100644 index 0000000..7132468 --- /dev/null +++ b/.agents/skills/results-report/references/report-naming.md @@ -0,0 +1,53 @@ +# Report Naming Standard + +## Filename + +Use: + +```text +YYYY-MM-DD--{experiment-line}--r{round}--{purpose}.md +``` + +Rules: +- date must be the report date, +- `experiment-line` should be short and stable, +- `round` should be zero-padded only if that is already the project convention; otherwise `r3` / `r03` are both acceptable if used consistently, +- `purpose` should describe why the report exists, not a vague label like `summary` unless that is truly the purpose. + +Recommended purpose values: +- `transfer-summary` +- `ablation-report` +- `failure-analysis` +- `robustness-check` +- `round-review` + +## Title + +Use: + +```text +{Experiment Line} / Round {N} / {Purpose} / {YYYY-MM-DD} +``` + +## Frontmatter fields + +Required: +- `type: results-report` +- `date` +- `experiment_line` +- `round` +- `purpose` +- `status` +- `source_artifacts` +- `linked_experiments` +- `linked_results` + +## Placement in Obsidian + +Internal reports go to: + +```text +Results/Reports/{filename} +``` + +Do not put internal experiment reports in `Writing/` unless they are already manuscript/slides/rebuttal material. diff --git a/.agents/skills/results-report/references/report-structure.md b/.agents/skills/results-report/references/report-structure.md new file mode 100644 index 0000000..e0eb247 --- /dev/null +++ b/.agents/skills/results-report/references/report-structure.md @@ -0,0 +1,63 @@ +# Results Report Structure + +## 1. Executive Summary +Answer: +- what was tested, +- what the highest-confidence conclusion is, +- what decision this changes. + +## 2. Experiment Identity and Decision Context +Answer: +- which experiment line this belongs to, +- why this round was run, +- what prior uncertainty or decision it was meant to resolve. + +## 3. Setup and Evaluation Protocol +Answer: +- datasets / subjects / splits, +- methods compared, +- primary metrics, +- repeated-run structure, +- any deviations from prior protocol. + +## 4. Main Findings +Answer: +- what changed most, +- which comparison matters most, +- where the largest gains or failures appear. + +## 5. Statistical Validation +Answer: +- what evidence supports the major claims, +- what tests were used, +- where the evidence is weak. + +## 6. Figure-by-Figure Interpretation +For each main figure: +- why it is shown, +- what to notice, +- what is supported, +- what remains uncertain. + +## 7. Failure Cases / Negative Results / Limitations +Answer: +- what did not work, +- what instability appeared, +- what limits the current conclusion. + +## 8. What Changed Our Belief +Answer: +- which prior hypothesis is strengthened, +- weakened, +- or still unresolved. + +## 9. Next Actions +Answer: +- stop / continue / ablate / scale / write. + +## 10. Artifact and Reproducibility Index +List: +- source artifacts, +- figure paths, +- scripts/logs, +- linked Obsidian notes. diff --git a/.agents/skills/results-report/references/statistical-completeness.md b/.agents/skills/results-report/references/statistical-completeness.md new file mode 100644 index 0000000..4c8756c --- /dev/null +++ b/.agents/skills/results-report/references/statistical-completeness.md @@ -0,0 +1,19 @@ +# Statistical Completeness for Results Reports + +A report may summarize statistics, but it must not silently weaken them. + +## Always carry forward +- sample size / seed count +- metric direction +- descriptive statistics +- uncertainty estimate +- test choice +- effect size +- correction rule when relevant +- evidence boundary + +## Never do this in the report +- upgrade a trend to a conclusion +- omit the sample size +- replace effect size with adjectives like “large” without numbers +- cite a figure without saying what the uncertainty represents |
