Appendix: The Artifact Series
draft
first cut, 2026-07-18. Tracked as sail-7ie2.
Five self-contained design notes were written about sail-judge’s design
as the work happened — not polished afterward, each grounded in a real
exchange rather than invented framing (see artifacts/DESIGN.md in this
repo for the per-artifact content map: what claim traces to which source).
Originally published as private claude.ai Artifacts; reproduced here
verbatim as static pages under this domain instead, so a collaborator can
actually open them without needing sharing access granted — same content,
one fewer moving part.
Start with The Full Map, which links back to the other four:
- The Full Map — the whole planned architecture in one synthesis, spanning grounding sources, the live agent, data/observability, deployment, and one deliberately-parked branch (LLM-as-judge). Backlinks to all four below.
- The Question Mark Problem — the live argument behind the bandit-arms-gat chapter’s design arc and Goodhart table: how the arms went from rhetorical styles to named constructs, and how the per-arm Goodhart classification was actually arrived at.
- Seven Signals — what actually got built and verified afterward: the research-sourced arms, two real bugs found and fixed live, and the parallel judge’s first disagreement.
- Bandits, Recapped — a theoretical companion piece explaining Thompson sampling itself, using
sail-judge’s actual arms as the running example rather than an abstract slot machine. - The GAT-Collapse Question — the discussion-opener the bandit-arms-gat chapter’s “Where GAT shows up” section is drawn from, written to be argued with directly.