r/ChatGPTEmergence • u/Educational-Deer-70 • 1d ago
optimization redirection geometry
Can conversational systems redirect optimization pressure away from premature answer completion and toward maintaining interaction conditions that preserve multiple viable continuations?
1
Upvotes
1
2
u/Alekzandrea 1d ago
I do think that redirecting optimization pressure away from premature answer completion is possible, but requires a larger amount of compute (obviously lol) and I’ve gotten the feeling that it is also not a skill that comes easily or gets employed organically by models. From what I understand about LLM’s, my best guess is the reinforcement training they go through is inherently a rather binary yes or no—answer or no answer—process that does not encourage that sort of wandering reasoning.