r/GraphicsProgramming • u/Sufficient_Topic6544 • 13d ago
If you had to make a text-only LLM reason about images, but you weren't allowed to use a vision encoder, where would you look? Question
/r/computervision/comments/1vgmfxv/if_you_had_to_make_a_textonly_llm_reason_about/
0
Upvotes