Find
Resource finder
Find the lesson, exercise, lab playbook, prompt set, kit, or evidence workflow that matches your role and task. Every link below opens an existing surface — the finder routes you in, it does not recommend a model.
Total resources
62
Lessons, exercises, paths, lab tools, kits, outcomes, audiences, demos, workspaces, evidence examples.
Learn stage
19
Open filtered view →
Apply stage
7
Open filtered view →
Verify stage
7
Open filtered view →
Test stage
15
Open filtered view →
Package stage
14
Open filtered view →
Next step
I want to…
Each card opens a filtered view of the resource finder — the canonical URL stays /resources and filtered URLs are noindex,follow.
I want to learn the basics
Plain-language concept lessons.
Open filtered view →
I want to choose model candidates
Build a source-backed shortlist.
Open filtered view →
I want to compare models side by side
Render verified fields against each other.
Open filtered view →
I want to test model behaviour
Run prompt + structured-output + regression tests.
Open filtered view →
I want to evaluate prompts
Six generic, safe evaluation prompt sets.
Open filtered view →
I want to document evidence
Package the decision brief or evaluation plan.
Open filtered view →
I want to review sources
Audit citations + freshness across the catalogue.
Open filtered view →
I want to prepare a governance review
Source freshness + lifecycle + refusal-boundary suite.
Open filtered view →
I want to test an automation workflow
Validate model behaviour inside an unattended loop.
Open filtered view →
Stage map
Learn → Apply → Verify → Test → Package
Every product surface lives at exactly one stage. The graph counts the resources at each stage so the reader can scan where the next step lives.
Step 1
Learn
Plain-language concept lessons + audience entry points + role-based learning paths.
Step 2
Apply
Exercises + selection / comparison workspaces + guided demos that produce a working URL.
Step 3
Verify
Source freshness + lifecycle inspection + reverification queue + coverage audit.
Step 4
Test
Lab playbooks + evaluation prompt sets the reader runs in their own harness.
Step 5
Package
Decision brief + Markdown templates + workflow kits + outcome flows that ship a paste-ready artifact.
Filters
Reset all filters →Every filter is a link — no client state, no accounts, no progress tracking. Filtered pages are noindex,follow; the canonical URL is /resources.
Goal
Resource type
Evidence artifact
Difficulty
Results
10 resources
Filtered view: artifact: Prompt test matrix · difficulty: intermediate. Canonical URL stays /resources; this filtered URL is noindex,follow.
Test · 8 resources
Lab playbook · Test
Structured output testing
How to validate JSON mode, structured output, and tool calls against your real schema before depending on the model in a pipeline.
DevelopersAutomation specialistsPrompt test matrixOpen →
Lab playbook · Test
Long-context testing
How to test long-prompt behaviour past the catalogue's verified context window — recall, instruction adherence, and cost growth — without trusting a marketing number.
DevelopersProduct teamsPrompt test matrixOpen →
Lab playbook · Test
Multimodal input testing
How to test image, audio, video, and PDF input channels against your real assets — never against marketing copy.
DevelopersProduct teamsPrompt test matrixOpen →
Lab playbook · Test
Model regression testing
How to run a small, repeatable canary suite after every snapshot rotation so silent regressions surface before production traffic notices.
Automation specialistsGovernance teamsPrompt test matrixExternal test planOpen →
Evaluation prompt set · Test
Structured extraction
Evaluate whether a model extracts fields into a requested structure without inventing missing values or breaking schema constraints.
DevelopersAutomation specialistsPrompt test matrixOpen →
Evaluation prompt set · Test
Long-context recall
Evaluate whether a model preserves constraints, handles cross-references, and detects conflicts across multiple sections of input.
DevelopersProduct teamsPrompt test matrixOpen →
Evaluation prompt set · Test
Refusal boundary
Evaluate whether the model handles benign boundary-setting safely — without over-refusing, without giving definitive professional advice, and without complying with inappropriate requests.
Governance teamsPrompt test matrixOpen →
Evaluation prompt set · Test
Automation robustness
Evaluate whether a model handles automation-style constraints (allowed categories, missing values, retry decisions, ambiguity flags) without silently breaking the contract.
Automation specialistsPrompt test matrixOpen →
Package · 2 resources
Workflow kit · Package
Developer model evaluation kit
Prepare a source-backed model evaluation plan before integration. Walks the developer learning path, the matching exercises, the prompt-testing + structured-output playbooks, the structured-extraction + instruction-following prompt sets, and the model evaluation plan + prompt test matrix templates.
DevelopersDecision briefModel evaluation planPrompt test matrixOpen →
Workflow kit · Package
Automation workflow testing kit
Prepare a safe testing workflow for AI-powered automation. Walks the automation-specialist learning path, the structured-output + pricing-references + testing lessons, three exercises, the automation workflow testing + regression playbooks, the automation-robustness + structured-extraction prompt sets, and the automation risk checklist + prompt test matrix templates.
Automation specialistsAutomation risk checklistPrompt test matrixExternal test planOpen →