@essjax> it would take an enormous effort to match the descriptive power that LLMs can generate in a second.
That just isn't my experience of
#LLM alt text for images.
The LLM can *emit many words*, but do those words adequately describe the image? Do the words omit the irrelevant parts, describe only the salient parts, connect the description to the *intention* of posting that image? Far too often, they don't.
The LLM has no understanding of the image. The LLM has no understanding of the words it emits. They have no intention. Far too often, they don't describe the image in any way worth the words extruded.
And so, IMO they are not valuable as image
#AltText. Human description is far superior; if there's no human willing to describe it, I wonder whether it's worth posting the image.
#ArtificialStupidity #WriteWithIntention