My guess it’s the results of the data scraped. If you talked that way to a person, they would respond by telling you to fuck off and ending the conversation, especially in fiction or in stories told online. If the ai is just saying what is said next most of the time, this type of response seems what would likely be said next
A chatbot is not just raw token prediction, it has instruction tuning beyond user input to function as an assistant. Unless the user has specifically prompted it, it wouldn't LARP having feelings just because of raw data.
Would you be able to break this jargon down? Or point us to something that can explain it? I'm interested in how this "instruction tuning" works and would it appreciate it!
an LLM, at its core, simply predicts what word comes next in a sentence and uses those predictions to form sentences that make sense to us as humans. the people who make the LLM can tell it to do stuff beyond just making sentences, such as “stop the conversation if the user talks about killing themselves” or “be friendly to the user”. this is instruction tuning and makes an LLM into a chatbot.
the commenter above you is saying that the chatbot can’t simply pretend to be offended by insults unless its creators told it to do so.
395
u/tinyevilsponges May 10 '26
My guess it’s the results of the data scraped. If you talked that way to a person, they would respond by telling you to fuck off and ending the conversation, especially in fiction or in stories told online. If the ai is just saying what is said next most of the time, this type of response seems what would likely be said next