← r/LocalLLaMA
▲
0
 
11👁
r/LocalLLaMA · u/opUserZero · 10d ago

Jev mode for images!

post image

So Codacus created Jev mode for Lllama.cpp , and I thought Why not extend this concept further and ask questions about images and have the constrained answer be an image selection? So i spun up an agent and added image support and a harness. Now you can use images as your prompt without the decode step, no caption pause, just a decision based on an image or group of images. Ask the same question for a batch of images, like clasification. OR hand 1 context a whole group of images and ask it to pick on. like which of these 20 images has a ruber duck?
https://github.com/thecodacus/llama.cpp/pull/17

Youtube explainer using Codacus own RenderDiv framework to create the video.
https://youtu.be/Xuw3la2zVpg?si=rtSAydhuF9n3SYWV

1 0 0 10/3 06:30 10/7 07:59 UTC
scorecomments11 sightings
first seen 2026-10-03 06:30 UTClast seen 2026-10-07 07:59 UTCscore then 0score now 0gained 0sightings 11
open on reddit ↗ 💬 21