Golf ball
A vision model describes an image, in one or two open sentences, as if to someone who can't see it. A text-to-image model renders that description. The render becomes the next round's input. Eight rounds. Nobody edits or selects between them.
The seed is an abstract study from earlier in this practice — ink-like marks, no recognisable subject. By the last round, the chain has arrived, unprompted and unplanned, at a specific photograph of an object that was never there to begin with.
a different seed, the same mechanism
Shown this image instead, the same kind of chain didn't drift — it stopped. Round 0's description named the letter correctly but got the colour backwards (“a white letter… on a white background”), which is self-cancelling once rendered. All eight rounds that followed stayed blank. Two outcomes, one mechanism: nothing in this loop checks a caption against the pixels it was written from. A caption can only sharpen toward something recognisable, or describe nothing at all — there is no way back to the original once either has happened.