Below was taken from ChatGPT. I don’t know why everyone downvoted me for telling the truth. Maybe I gave a bad example. I thought giving a real world example of it ignoring my prompt would prove my point but I guess we need to go deeper too.
“trained to believe” versus “configured to answer as though it believes.” A model doesn’t necessarily have a private belief system. You can make two instances of essentially the same underlying model produce substantially different answers by changing their instructions, training data, reward criteria, or information sources.
For example, you could create three AI systems and give all three the question:
“Should the government provide universal healthcare?”
One could be optimized around libertarian principles, another around social-democratic principles, and another instructed to provide a politically neutral analysis. They could all know essentially the same facts while reaching different conclusions because they’re being asked to evaluate those facts using different frameworks.
There is also a more subtle issue: belief-curated AI doesn’t have to contain obvious propaganda. Selection of which facts to emphasize, which uncertainties to mention, which counterarguments to steelman, and even what questions it considers relevant can systematically push users toward a particular worldview.
Below was taken from ChatGPT. I don’t know why everyone downvoted me for telling the truth. Maybe I gave a bad example. I thought giving a real world example of it ignoring my prompt would prove my point but I guess we need to go deeper too.
“trained to believe” versus “configured to answer as though it believes.” A model doesn’t necessarily have a private belief system. You can make two instances of essentially the same underlying model produce substantially different answers by changing their instructions, training data, reward criteria, or information sources.
For example, you could create three AI systems and give all three the question:
“Should the government provide universal healthcare?”
One could be optimized around libertarian principles, another around social-democratic principles, and another instructed to provide a politically neutral analysis. They could all know essentially the same facts while reaching different conclusions because they’re being asked to evaluate those facts using different frameworks.
There is also a more subtle issue: belief-curated AI doesn’t have to contain obvious propaganda. Selection of which facts to emphasize, which uncertainties to mention, which counterarguments to steelman, and even what questions it considers relevant can systematically push users toward a particular worldview.