15. Evaluation and Rubrics

Judge AI output with measurable criteria instead of personal feeling.

By Jacques Botte, founder of Toptronic®. Last updated 12 September 2026.

The lesson

Professional AI users define what good means before asking for work. A rubric can score correctness, completeness, safety, clarity, cost, maintainability, and test coverage.

For coding, require compile results, tests, edge cases, security impact, and a concise change summary.

For business, require assumptions, risks, alternatives, numbers, and decision criteria.

Check yourself

Question 1: What does a rubric provide?
  1. A window size
  2. Measurable criteria for judging output — correct
  3. A secret API key
  4. A random style

Answer: Measurable criteria for judging output

Rubrics define quality before work begins.

Question 2: For code, which gate belongs in a rubric?
  1. Compile and test results — correct
  2. Favorite color
  3. No acceptance criteria
  4. Ignore edge cases

Answer: Compile and test results

Code should prove itself with builds and tests.

Question 3: For business output, what should be checked?
  1. Only grammar
  2. Only emojis
  3. Nothing
  4. Assumptions, risks, alternatives, numbers, and decision criteria — correct

Answer: Assumptions, risks, alternatives, numbers, and decision criteria

Business decisions need evidence and trade-offs.

← Previous lesson · All 83 lessons · Next lesson →

The full course — 83 lessons and 249 quiz questions — ships inside the app. Get TPEE to study it offline.