15. Evaluation and Rubrics
Judge AI output with measurable criteria instead of personal feeling.
By Jacques Botte, founder of Toptronic®. Last updated 12 September 2026.
The lesson
Professional AI users define what good means before asking for work. A rubric can score correctness, completeness, safety, clarity, cost, maintainability, and test coverage.
For coding, require compile results, tests, edge cases, security impact, and a concise change summary.
For business, require assumptions, risks, alternatives, numbers, and decision criteria.
Check yourself
Question 1: What does a rubric provide?
- A window size
- Measurable criteria for judging output — correct
- A secret API key
- A random style
Answer: Measurable criteria for judging output
Rubrics define quality before work begins.
Question 2: For code, which gate belongs in a rubric?
- Compile and test results — correct
- Favorite color
- No acceptance criteria
- Ignore edge cases
Answer: Compile and test results
Code should prove itself with builds and tests.
Question 3: For business output, what should be checked?
- Only grammar
- Only emojis
- Nothing
- Assumptions, risks, alternatives, numbers, and decision criteria — correct
Answer: Assumptions, risks, alternatives, numbers, and decision criteria
Business decisions need evidence and trade-offs.
← Previous lesson · All 83 lessons · Next lesson →
The full course — 83 lessons and 249 quiz questions — ships inside the app. Get TPEE to study it offline.