logoalt Hacker News

daft_pinktoday at 3:58 PM1 replyview on HN

I'm really looking for a multi-modal image capable version of Jev.

If we could get machine learning type results on images without training, that would be fantastic.


Replies

oreoftwtoday at 5:06 PM

Jev’s context should have signals/features to operate on. Same way LLMs can use CV & code to analyze an image.