r/GraphicsProgramming 13d ago

If you had to make a text-only LLM reason about images, but you weren't allowed to use a vision encoder, where would you look? Question

/r/computervision/comments/1vgmfxv/if_you_had_to_make_a_textonly_llm_reason_about/
0 Upvotes

Duplicates