References for the agentic engineering guide
What this draws on, and what has actually been read
The two original sources have not been read in full. Both were retrieved on 2026-09-02 through an automated fetch that returns a summary rather than the page, so every quotation and section heading below reached this project second-hand. The working document’s Section 1 and Section 2 characterize both sources, and those characterizations are what needs checking. The observations in Sections 3, 4 and 5 come from this repository’s own history and do not depend on either paper. The context-engineering tutorial added on 2026-09-17 has been checked for the specific claims recorded below.
The working document is the specification, and the folder’s documents are on the project index.
Status markers
- ❌ Not checked. No claim in the working document has been verified against this source.
- ⚠️ Transcribed, unverified. A quotation, heading or claim in the working document came from this source but has not been checked against it.
- ✅ Checked. The claim the working document draws from it has been verified against the source.
Change a marker when the check is done, not when the page is opened.
Reading queue
The two original sources remain in the queue, with the playbook first.
⚠️ The AI-Native SDLC Playbook. Read first. Section 2 of the working document maps this repository’s existing artifacts onto its vocabulary, so if the mapping is wrong the whole section is wrong. Read against: what does it put in
intent.mdas againstspec.md, what does it gate on, and what does it say a human must still do at each of the six stages?⚠️ Silva, 2026. Read second, and quickly. It is the piece that prompted the project and it argues a position rather than supplying a procedure. Read against: what does owning “the specifications, the tests, the quality assurance process” mean concretely, and does it name a verification mechanism the working document’s Section 5 does not have?
Sources
✅ Pritchard-Bell A, Lin C-W, Holmes W, Doshi S. Context Engineering for AI-Assisted Pharmacometrics: A Practical Tutorial. CPT: Pharmacometrics & Systems Pharmacology, 2026;15:e70317. Publisher article · Open full text
Checked against the full-text passages on 2026-09-17: task-specific rules, verification and examples; fresh context with workspace-file handoffs; the ΔOFV > 3.84 covariate-inclusion threshold in Section 1; and forward selection and backward elimination in Section 4.3. Used in the project index’s related reading. The preference there for a full covariate model, and concern about overfitting noise through stepwise selection, are Andy’s methodological reservation, not a conclusion of the tutorial.
⚠️ Anthropic. The AI-Native SDLC Playbook. claude.com/blog/the-ai-native-sdlc-playbook
Six stages: plan, design, build, test, deploy, maintain. Version-controlled markdown as the thread through them, with intent.md holding the problem statement and constraints, spec.md the requirements and design, plan.md the implementation plan, and CLAUDE.md the standing conventions. Humans approve at gates and remain accountable for decisions that need judgement; agents draft the spec from the intent, the plan from the spec, and the implementation from the plan.
Used in Section 2 for the artifact vocabulary and in Section 3 for the intent-then-spec split. The stage names and the artifact list came from the fetch summary rather than the page.
⚠️ Silva DW. Why vibe code when you can engineer? Substack, 31 August 2026. davidwsilva.substack.com
Contrasts vibe coding, which accepts suggestions without reading them, with agentic engineering, where the developer owns the specifications, the tests and the quality assurance, and requirements are engineering artifacts that a human and an agent can both consume: architecture constraints, security requirements, coding standards, compliance policies, stated before implementation. Agents generate; humans inspect, direct and govern.
Used in Section 1 for the framing and the phrase “requirements written so that a human and an agent can both consume them”. The section headings recorded from the fetch were “There is a better way”, “Go beyond markdown files” and “An important distinction”; those are unverified.
✅ Vincent J, Prime Radiant. Superpowers. GitHub repository, MIT licence. github.com/obra/superpowers
Checked against the README on 2026-09-30, read in full rather than through a fetch summary: the Claude Code install command, the skills triggering automatically, the workflow from brainstorming through finishing a branch, tasks of two to five minutes, deletion of code written before its test, and the telemetry opt-out. Used in Section 8.
✅ GitHub. Spec Kit. GitHub repository, MIT licence. github.com/github/spec-kit · Documentation · Command reference
Checked against the README and the command reference on 2026-09-30: the specify command-line tool, the constitution once per project, the specify, plan, tasks, implement and converge sequence, what converge does, the three optional gates, and the bug-fixing and idea-assessment extensions with their verdicts. Used in Section 8. Which integration key selects Claude Code was not checked; the integrations page lists them.
Not yet read
Candidates, none of them retrieved.
Jesse Vincent’s release announcement for Superpowers, which the README links for the reasoning behind the design.
Spec Kit’s full methodology, the argument for spec-driven development behind the commands.
Karpathy’s original description of vibe coding, quoted second-hand by Silva.
Anthropic’s published guidance on skills and on
CLAUDE.md, which Section 2 treats as the constraint mechanism without citing it.Anything on specification practice from outside software: a statistical analysis plan is the closest analogue in Andy’s own field, and Section 7’s open question about whether this generalizes is really a question about how a SAP and a
spec.mddiffer.