Lab 1 — Prompt and Verification

Lab 1 — Prompt and Verification

Key jargon

Term Plain-language meaning
Hypothesis A testable expectation about how a prompt change will affect results.
Test case One defined input with expected properties or outcome.
Control prompt The unchanged baseline used for comparison.
Failure taxonomy Named categories used to classify observed errors.

Key concepts

Concept map

flowchart LR
    A["Define hypothesis and cases"] --> B["Run baseline and variant"]
    B --> C["Verify claims and classify failures"]
    C --> D["Write evidence-backed conclusion"]

Goal

Produce a one-page comparison using three supplied authoritative sources and demonstrate which source supports each material claim.

Steps

  1. Define audience, decision, scope, date, and output contract.
  2. Supply three short documents plus source metadata.
  3. Require claim-level citations and explicit unknowns.
  4. Run three times; record model/version/settings.
  5. Open every citation and label support as full, partial, or absent.
  6. Revise the prompt once and compare failure classes.

Pass criteria

Artifacts

Task contract, source manifest, prompt versions, outputs, claim audit, decision, and lessons learned.