Palestra
Curriculum/Eval Systems & Quality Gates/5.1 Why Code-Gen Evals Are Uniquely Hard

5.1 Why Code-Gen Evals Are Uniquely Hard

Understand why AI code generation cannot use traditional assertions — non-determinism requires semantic evaluation — and map the eight-level correctness spectrum from syntax parsing to aesthetic judgment that defines the full eval infrastructure Make must build.

Lesson Locked

Complete the previous lesson to unlock this one.

Your balance: 1,000