This works best when the hard part is explaining what is already on screen, not making a new image. The useful paths are the practical ones: accessibility, OCR, and batch export, where the text can go straight into another system or workflow.
It also helps that the tool does not force every image task into one generic output. Plain descriptions, decorative alt text, prompt extraction, social captions, and OCR are split apart. The part that still needs proving is trust at scale, because the product story is ahead of the amount of outside validation behind it.