Prompt Evaluations¶
The evaluation system uses a Prompt Quality Rubric with eight dimensions:
- Objective clarity.
- Context sufficiency.
- Instruction completeness.
- Constraint clarity.
- Output contract quality.
- Grounding and uncertainty handling.
- Eval readiness.
- Efficiency and safety.
The examples in catalog/evaluations.json show before-and-after prompt changes for
support routing, RAG policy answers, and coding-agent repository changes. Each
example records before scores, after scores, related patterns, and acceptance
criteria.
Example files are also available in examples/evaluations/.