skip to content

Testing AI-Powered Features

Testing the product built around a model, not the model itself: acceptance criteria for variable output, suite seams and repeats, and the ship call. Interviewers probe where that line falls.

on this pageshow

questions

page 2 of 2

When a feature's generative step is swapped and its wording shifts, which test cases should fail and which must not?

level: seniorimportance: nice to knowfreq 30%

basics

~20 s

Nothing on the exact-assertion surface should move: dispatch, response fields, citation rendering, states and permissions keep passing. Only the wording-reading checks may fail, and their output must name the text as what changed rather than the surrounding behaviour.

open as a page

A generative product feature's test case met its criteria in four of five repeats - what does a single pass mark hide?

level: seniorimportance: nice to knowfreq 32%

basics

~20 s

A single pass mark collapses a rate into a certainty. It hides which repeat missed and on which criterion, how badly it missed, whether four-in-five is normal for this case, and whether the misses cluster on one input.

open as a page

showing 31–32 of 32