logoalt Hacker News

xpct • yesterday at 11:10 PM • 0 replies • view on HN

I experimented with something similar with LLMs. They can kinda do it for some stuff.

More interesting, you can ask a vision model to color parts of speech in img2img and it works OK for frontier models.