I’ve found a core misconception is persistent... people use the CLIP interrogator model expecting it to recover the original prompt from an image. It cannot do this, and if you look at the architecture it becomes clear why. The mapping from prompt to image is non-injective - many different prompts produce nearly identical outputs, and some visual featur…
Substack is the home for great culture


