the.ai

Code / Verification

verified

Test Generation

If execution is how code gets verified, then tests are the bottleneck — and writing them is exactly the tedious work people want a model for. The circularity is obvious and worth stating: a model that misunderstands the specification writes tests that encode the misunderstanding, and then the wrong code passes them confidently.

Viz primitive · budget-splitgenerated-tests = 4

generated-tests holds 12% of the budget; rest holds the remaining 88%.

Behaviours a generated suite exercises, against the ones it never reaches, in behaviours. Drag the test count up to watch coverage saturate below the specification — the inputs left over are the empty, enormous and adversarial ones a model writing typical cases does not think of.

4

Reviewed by opendroid · 2026-08-18