Prompt Eval Handbook
A queryable handbook for golden sets, LLM judges, pass@k, leakage, and prompt regression — written for people who ship prompts.
Portable policy knowledge for local retrieval.
Prompt Eval Handbook is original evaluation guidance. It is not a vendor cookbook.
A prompt change without a frozen golden set is an anecdote, not a release.
The included agent metadata is never executed by Knolo Hub.
Intended use
Grounding prompt and model evaluation. Lookup when building a golden set, using a judge, or blocking a prompt change.
Out of scope
This pack does not call a model, score production traffic, or store API keys. It does not run a shell.
Sources
Sample questions
This Knowledge Image contains prompts, tool policy, namespace grants, or other agent registry metadata.
Mounting the pack for retrieval does not execute tools. Applying these policies affects your application only if your host explicitly honors them.
Credentials must never be stored inside a pack.