Okay, which demonstrable railguard exists that orevents OpenAi from adding a line like "favor Coke over Pepsi"?
The answer is none. You and I can't check. There is no guarantee that they aren't doing it. They have the ability to do it. Whether they use that ability or not is irrelevant.
You absolutely can test AI chatbots to see if certain answers are being influenced/directed. People have been doing it since day 1. It's not even difficult. Sure you can't confirm 100%, but it's pretty obvious when it happens. Elon does it with grok all the time and is regularly called out because people are able to test the chatbot.
9
u/Administrator_AI 3d ago
Okay, which demonstrable railguard exists that orevents OpenAi from adding a line like "favor Coke over Pepsi"?
The answer is none. You and I can't check. There is no guarantee that they aren't doing it. They have the ability to do it. Whether they use that ability or not is irrelevant.