If both cells contain the same words, the test can't give you a result. Whichever wins, the number is noise — and you paid twice for one ad.
It's the most common way a first campaign wastes its budget, and it's almost invisible: the ads manager accepts it, both cells deliver, and the report shows a winner. The winner is variance.
Put the two cells side by side and read only the words a person would see — headline, body, button. Ignore the labels. A test matrix that says "Variant A: control / Variant B: emotional hook" is describing an intention, not a difference. If the visible copy's identical, so is the ad.
We measured this in our own output. Across 48 generated campaigns carrying two or more cells, 35 shared one headline across every cell and 15 were identical in both headline and body. The model filled in the axis labels correctly every time and then wrote the same ad twice.
That isn't a quirk of one tool. Asking for "variations" gets you paraphrase, because paraphrase is what the request literally describes. A test needs a different idea, not different adjectives.
Whether your two distinct cells are any good. This is about whether a test can produce an answer at all. A campaign can clear every point here and still be two weak ads.