I trained a vision-language model that answers typed questions about an image. It uses also choice, score and noul, like Jev. I thought, why Jev only processes text?

Source: [Hacker News](https://github.com/bykof/peekaboolean)

Sponsored