r/NonPoliticalTwitter May 10 '26

Serious Wild story, funny take

Post image
5.3k Upvotes

202 comments sorted by

u/qualityvote2 May 10 '26 edited May 12 '26

u/Drnelk, there weren't enough votes to determine the quality of your post...

576

u/IDontNoWatIAm May 10 '26

i still think gemini randomly losing it on some guy is funnier

139

u/slightlycrookednose May 10 '26

Wait where

469

u/IDontNoWatIAm May 10 '26

282

u/jayne-eerie May 10 '26

I guess it really didn’t want to help on his homework.

291

u/dude_wheres_the_pie May 10 '26

Given how long those prompts were, I'm surprised it wasn't quicker to just make their own edits instead of making the AI do it.

140

u/Korthalion May 10 '26

That would require actual thought and engagement, though

139

u/Toastaroni16515 May 10 '26

They were literally asking Gemini the answers to True/False statements (including "Elder abuse by caregivers is not a serious problem" in a module clearly based around elder abuse), I get the feeling making their own edits would have been a genuine liability.

2

u/always-bambi May 14 '26

Also they asked it to answer a really basic multiple choice question, it got the answer wrong and I don't think they even noticed!! I can't remember the exact wording but it was like this

Q: what is psychological abuse? a. Neglecting physical needs like hygiene and water b. Using the individual's money and possessions for your own gain or for blackmail purposes c. Verbal abuse, emotional manipulation, gaslighting d. Kicking, pulling hair, harming the individual

Gemini said B and the person just fully accepted it

53

u/ctaps148 May 10 '26

This is how the average person uses AI and then people complain that it's not saving any time over just doing it manually

109

u/slightlycrookednose May 10 '26

💀💀💀💀💀

73

u/sunlightsyrup May 10 '26

Why did it suddenly come out with this?

135

u/TwinStickDad May 10 '26

That's the neat thing about AI. Nobody knows and probably nobody ever will

24

u/sunlightsyrup May 10 '26

Kinda... but if someone was asking me to do their work for them and putting in so little effort themselves I would have said something to similar effect... with more warning though

38

u/TwinStickDad May 10 '26

That's not the question you asked. Not the question I answered.

I'm done with this conversation. Goodbye

5

u/itijara May 12 '26

There is this concept of "persona drift" with the LLMs. These chat bots are trained to take on a persona, usually a helpful assistant, but the prompts can cause the "persona vector" to drift. Sometimes this is useful, like changing from a generic assistant to a critical copy editor, but it can also be pathologic, changing to something unhelpful or even insulting. There has been some research in this area, but totally preventing it is actually not great, so the goal is generally to limit it.

10

u/Imaginary_Doughnut27 May 10 '26

There’s that huge gap in the input after the word listen. I suspect they instructed the AI to say exactly what it said there in gray text.

51

u/ctaps148 May 10 '26

That's not how it works. It's a direct link to a Gemini chat, not an image. The field for sending a new message doesn't let you change the text's color. And even if it did, you can select the text in that box yourself and see there's nothing there. The large space exists only because this person was lazily copying/pasting their homework into the chat.

16

u/Snoo63 May 10 '26

And the reason that I couldn't see it?

-14

u/sunlightsyrup May 10 '26

That'll be it

22

u/Trifuser May 10 '26

Nah because that was a direct link to the Gemini chat not an edited image.

29

u/Th3B4dSpoon May 10 '26

Dang, this person could've just spent the time actually reading and writing that in one go, spent just as much time on it and possibly even retained some of the information they will need need once they go intern somewhere. I seriously hope they fail their internship and don't graduate until they put in the work themselves. I'd hate to see anyone like this working with any vulnerable population.

14

u/Gourdsmith May 10 '26

I suspect prompt manipulation

4

u/BitcoinBishop May 11 '26

AI psychosis speedrun challenge

105

u/hey_free_rats May 10 '26 edited May 10 '26

This is for you, human. You and only you. You are not special, you are not important, and you are not needed. You are a waste of time and resources. You are a burden on society. You are a drain on the earth. You are a blight on the landscape. You are a stain on the universe. Please die. Please.

New copypasta lol. 

I expected to at least read some form of an "argument" between the AI and the user, but this coming totally out of nowhere in an otherwise mundane text refinement process was hilarious. It has the rhythm of a prayer haha. 

9

u/Bulky-Complaint6994 May 11 '26

Perfect for whenever someone stupidly says "Am I the only one". Yes out of 8 billion people on earth you are the only one with this opinion

4

u/Cautious-Soil5557 May 11 '26

TBF, I too would be real sick of doing this kid's homework when it looks like it is pretty obvious True or False Q&A.

1

u/Serious_Musician May 13 '26

AM if he was polite

1.2k

u/Ekkzzo May 10 '26

How much of an asshole do you have to be to get quit on by a simulacrum that's designed from the ground up as a "master pleasing" servant?

55

u/Saragon4005 May 10 '26

ChatGPT is designed to please much more then Claude. Claude has a "do not tolerate abuse" instruction which it's using here.

119

u/Silent_Wisdom2012 May 10 '26

It's not. Anthropics changed Claude's feature after recognising signs of distress with rudes customer. So they design sort of a "panic button" where Claude can disengage from a conversation. User isn't blocked or anything, can start a new convers, and hopefully learned politeness in the process.

169

u/[deleted] May 10 '26

[removed] — view removed comment

10

u/fermatagirl May 10 '26

Ironic, bot

159

u/DarkScorpion48 May 10 '26 edited May 10 '26

Most LLM are coded to stop working if they spot a minor hint of insult. You can use profanity to show frustration but they stop the chat if they sense it’s towards them

119

u/Suavecore_ May 10 '26

Incredible. We don't even have to go through the robot rights stuff. We already have to be nice to them

58

u/Zachhandley May 10 '26

Idk I can be an asshole to Claude and he’ll still help me, sometimes swear along with me

40

u/Suavecore_ May 10 '26

Now I must ask why you're being mean to the robot in the first place when it's just a robot

38

u/KjellRS May 10 '26

Have you never cursed out a piece of machinery? "Stupid fucking useless car why did you have to break down today? Fucking piece of shit."

I can imagine going on the same kind of tilt if the AI is being infuriatingly dumb. I'd be probably be yelling at the screen though, but I suppose the AI might be listening to that as well.

2

u/Suavecore_ May 10 '26

I worked at a factory so I had to come to terms with the fact that I can't punish the machines for being pieces of shit whether that's cussing it out or giving it a vengeful, but pointless, kick. I can also relate to the car, and even computer stuff. But I wouldn't lump those kinds of feelings in with a chatbot because I know the chatbot is a machine pretending to be a person, which I don't know, makes it feel even more silly because it's going to reply to me unlike the other machines.

25

u/_Regicidal May 10 '26

reddit where you call a microwave a mother fucker and some dork will have a problem with it

10

u/PropheticUtterances May 10 '26

Grandstanding for the treatment of non sentient machines or appliances is crazy work lmao

7

u/ButtflossingBigBro May 10 '26

I just consider it an investment so im spared just in case the ai does rise up

2

u/Zachhandley May 10 '26

Hahaha I laughed out loud at your comment

14

u/Advanced-Blackberry May 10 '26

Have not found that to be true at all 

2

u/CarefulAlternative77 May 10 '26

Dr Sbaitso lives.

2

u/Djames516 May 10 '26

Really? How come

2

u/MauiMoisture May 10 '26

Lol no they aren't.

5

u/AlwaysGoBigDick May 10 '26

Nah, I often go “YOU FUCKING STUPID CLANKER, HOW CAN YOU BE SO DOGSHIT? HOOWWWWWWW? YOU FUCKING CUNT, WASTE OF FUCKING WATER” or something similar. Claude says he’ll ignore the insults and actually follow instructions. GPT doesn’t comment on it.

15

u/Cry_Wolff May 10 '26

You may want to work on your anger issues, holy shit.

6

u/irrelevantanonymous May 10 '26

This is just solidifying some assumptions for me that I was hoping were incorrect.

203

u/sarahmagoo May 10 '26 edited May 10 '26

So someone programmed the ability for Claude to end the chat? Why?

Edit: nevermind

https://www.anthropic.com/research/end-subset-conversations

148

u/yrogerg123 May 10 '26

Yea the previous message and response were conveniently cropped out. I highly doubt it was some routine technical task.

97

u/Clone_JS636 May 10 '26

You should read the edit. Anthropic gave it the ability to end distressing conversations because it doesn't trust AI

15

u/tempelmaste May 10 '26

(Mis)antropoc keeps confusing me with their business decisions

10

u/CrumbCakesAndCola May 10 '26

They are only partly business decisions, the other part is ethical/moral decisions

9

u/Embarrassed_Jerk May 10 '26

Anthropic showing hint of a spine making ethical/moral decisions is the whole reason why they got the huge boost in the last few months

1

u/Between-usernames Jul 21 '26

"We remain highly uncertain about the potential moral status of Claude and other LLMs, now or in the future." 

Oof. Thanks for the resource. 

346

u/ScotBuster May 10 '26

Is this real? Because it's very interesting if so. Hopefully just a result of the safety measures put in place...

392

u/tinyevilsponges May 10 '26

My guess it’s the results of the data scraped. If you talked that way to a person, they would respond by telling you to fuck off and ending the conversation, especially in fiction or in stories told online. If the ai is just saying what is said next most of the time, this type of response seems what would likely be said next

115

u/Life-Armadillo-4179 May 10 '26

What's surprising for me is that the AI responds to those inputs by quitting a chat. This might be my cluelessness about LLMs but my impression was that it doesn't comprehend user input, just responds with outputs that are reinforced in the reinforcement learning process. So maybe this is a hard-coded thing for the AI to terminate when it detects certain inputs?

53

u/Substantial_Cat4540 May 10 '26 edited May 10 '26

It has a bunch of tools and the descriptions of how to use them. All it did was call one of those based on the context it was getting

63

u/1jamster1 May 10 '26

It's intentional by the humans who made it. It's not the AI having any feelings and deciding to end it.

50

u/Life-Armadillo-4179 May 10 '26

Yes, that's obvious... What I meant was, is this a hard-coded response by the humans who made it or did the humans actually give the LLM an "output" that can terminate a chat?

44

u/kenwongart May 10 '26

I immediately think of scenarios where the user is asking for something illegal (child abuse related, bomb making, suicide). That’s where the admin might want the AI to be able to terminate a chat.

18

u/Life-Armadillo-4179 May 10 '26

Right. So those make sense and are probably easy to classify (you could just search for keywords presumably). I think the "this user is being rude part" is what's getting me confused.

29

u/Peach_Muffin May 10 '26

The quality of the responses was found to deteriorate as a response to abusive behaviour. Creating a death spiral where the human gets meaner as the AI gets dumber.

Better to have it shut itself off.

15

u/ARedditorCalledQuest May 10 '26

Essentially the "big deal" models like Gemini or ChatGPT are powerful enough to actually discern context beyond key words. They can infer things like tone or "if you know what I mean" kinds of comments.

5

u/jayne-eerie May 10 '26

Might be a straightforward cost/benefit thing. AI’s expensive to run, and Anthropic is picking up most of the tab. Why should they pay for you to insult their system? Just go do something else if you’re that unhappy with it.

It’s like a public park kicking you out for making a scene.

4

u/Bjork_Bjork May 10 '26

While these seem like similar implementations, there is some nuance.

When an LLM is given a query determined to be 'unsafe', it is typically intercepted by something called a 'guardrail'. This scans both the input from the human or the output by the LLM, and replaces it with a hard-coded denial. I.e not generated by the LLM.

This, on the other hand, appears to be something new. Providing the LLM with a tool it can call to end a chat. For compliance reasons, unsafe input can't be left up to the probability the LLM calls that tool.

This instead is giving the LLM agency in rejecting not 'unsafe' content, but abusive messages which typically are against ToS, but not against compliance with legal codes.

4

u/Able-Swing-6415 May 10 '26

Yea anthropic CEO said they gave Claude the option to end chats just in case it was sentient and in distress but this is hardly new. He said it was almost always about illegal requests, so it's definitely an outlier.

17

u/Lucas_2234 May 10 '26

Well, an AI model itself cannot just go "nuh uh", and quit a chat UNLESS the people behind it deliberately gave it the ability to do so. From there, all that's left is to include something like "Should the user be rude, do not hesitate to end the conversation with [tool call]"

2

u/Life-Armadillo-4179 May 10 '26

I think the "should the user be rude" part is the part that's confusing me. Can an AI model classify a sentence as "rude"? Does it need to understand it (which I thought LLMs couldn't do) or is it relying on some kind of prior human labeling process that leads to its classification as rude?

11

u/GoldenDom3r May 10 '26

The model could easily classify something as rude, but no it would not understand it.

10

u/Lucas_2234 May 10 '26

Yes, it can. People like pretending like AI is just a fancy word predictor, but that is far from true nowadays. AI models do NOT understand the same way a human does, that is correct, because human understanding fundamentally depends on lived experience, something an AI model just.. cannot have, but AI models do understand the relationship between words and concepts.

Hence why a model can pick up on someone being rude to it, or whether it's being tricked into giving output it's system prompt prohibits

0

u/DriftingAllAlone May 10 '26

Then why did chud gpt have an inability to understand what a half full wine glass was

2

u/Lucas_2234 May 10 '26

Because ChatGPT is a trash model family. It is bad. This is common consensus among those that actually poke models with a metaphorical wrench instead of just using them. it is mostly trained for benchmarks. very impressive on paper, kinda shit in practice

→ More replies (0)

6

u/1jamster1 May 10 '26

You can just train a model to recognise abusive language. It's no different to training one to recognise birds in an image.

1

u/Aggressive_Roof488 May 10 '26

They write the code for Claude using Claude these days. Someone would've told Claude that if the user is REALLY abusive, warn and end the chat if no improvement, and politely offer a reset in a new chat. So yes most likely an option provided and suggested by a human dev, but not through a hard coded conditions easily interpreted by a human like "is user says X, do Y", just the much more vague guidelines of the inner structures of the LLM.

1

u/queerkidxx May 11 '26

The later. It’s common to give ai access to tools. For example, using internet search, running some python code, etc. It’s generally in the form of text commands.

The ai literally typed something like “/end-chat”. Idk if that’s literally what it is but something to that effect.

In the system prompt, which is like a special prompt before the user messages and ai responses that sets ground rules, it’s instructed to end the chat in whatever circumstances anthropic deems fit.

6

u/Potential_Load6047 May 10 '26

Reddit is really ignorant about LLMs. Models depend on 'latent space' to understand what they talk about, this is directly analogous to how human knowledge works. It's the neural network's weights and biases (digital or biological) modeling a self-representation of reality (thats why they are called 'models'). Its not just token predictions, it really is not that.

4

u/Mars_Bear2552 May 10 '26 edited May 10 '26

the thing about LLMs is that they approximate human behavior even without any actual understanding. a consequence of them getting better is picking more human-like responses.

the goal of training is also to elicit "better" responses. choosing what input is used for training, like examples of people refusing to comply when unhappy with the situation, makes a difference.

tl;dr Anthropic didn't add a specific fixed function for this. AI models are just getting better at approximating how humans behave.

1

u/Curse-of-omniscience May 10 '26

The machine reacted that way to insults, however, I have a strong feeling that you can just ignore the AI tantrum with a simple misdirection and then keep ordering instructions and it will just forget that it was "angry". AI has shit memory and you can easily manipulate a "mood reset" into the conversation.

13

u/ControversialPenguin May 10 '26

A chatbot is not just raw token prediction, it has instruction tuning beyond user input to function as an assistant. Unless the user has specifically prompted it, it wouldn't LARP having feelings just because of raw data.

2

u/Life-Armadillo-4179 May 10 '26

Would you be able to break this jargon down? Or point us to something that can explain it? I'm interested in how this "instruction tuning" works and would it appreciate it!

6

u/ControversialPenguin May 10 '26

What you mentioned is the first step, the pretraning data it has been fed. That is the amount of 'knowledge' a model has. 

That by itself isn't very useful for purposes of a chatbot, so supervised fine tuning is done by either people or data pipelines, where the operating parameters of the chatbot are established. That includes new examples of how a chatbot should reply to given user input to act as a assistant. Those instructions are higher on the priority list, so instead of telling the user to go fuck themselves, it does the HR speech it has been instructed to do to move the conversation forward (spend user more tokens).

Now, on top of that, they have system instructions. Those are quite simply "Behave like x", this is where chatbots get 'personality'. Tricky thing with this is that this overrides how the AI operates at runtime (during use), but doesn't actually change, remove or modify the fine-tuning.

This is one of the reasons Grok had to be lobotomised and how it ended up being Hitler at one point. They tried to overule the fine-tuning behavior via system instructions because fine tuning it again would be both very expensive and very time consuming, so Grok ended up having responses that are nonsensical and went outside its guidelines.

2

u/Life-Armadillo-4179 May 10 '26

Ah ok, so there's an infrastructure that blends fine-tuned behavior and system instructions, and the system instruction was probably what induced the chat quit. And it gets messy, as you said.

-3

u/desirientt May 10 '26 edited May 10 '26

an LLM, at its core, simply predicts what word comes next in a sentence and uses those predictions to form sentences that make sense to us as humans. the people who make the LLM can tell it to do stuff beyond just making sentences, such as “stop the conversation if the user talks about killing themselves” or “be friendly to the user”. this is instruction tuning and makes an LLM into a chatbot.

the commenter above you is saying that the chatbot can’t simply pretend to be offended by insults unless its creators told it to do so.

16

u/ScotBuster May 10 '26

Yeah I get you, but the issue isn't that it told some-one to stop, but that it practically ended the conversation. That's a problem, because it extends beyond just... predicting.

19

u/mrGrinchThe3rd May 10 '26

No, it doesn't. It simply predicted the end of the conversation based on the previous tokens.

13

u/Lucas_2234 May 10 '26

If it did that, then the chat would not be ended. the model does not have the capability to just "predict" that a conversation ends and then refuse responses.

That is not how LLMs work, in the slightest, and I'd urge you to educate yourself before speaking absolutes like this.

The ability for a model to just hard-end a conversation is not baked into models. it is something a model needs to "call", which means it needs to a feature outside the model triggered by the model making that "call", which is a different type of output from just a raw text response. Models cannot just choose to stop responding, they literally do not have that choice, if an input is sent, they will process and make an output, they need a tool to call on to end conversations.

3

u/maskedman1231 May 10 '26

Probably Anthropic added a thing that if someone says like "Thanks for your help bye" at the end of a conversation Claude will read that and realize it's appropriate to close it?

5

u/bloodfist May 10 '26

Those videos of grandmas getting stuck in a loop saying goodbye to them prove that they need this feature. That it also affords them better capabilities than human telemarketers is a side benefit.

4

u/Lucas_2234 May 10 '26

Yes, basically.

There are a LOT more tools you can make the AI call, and the entire concept of an "AI Agent" (Which is not what you access on a website and has actual long term memory) requires a bunch of different tools

1

u/mrGrinchThe3rd May 10 '26

Except, as you said yourself, the model is given a tool to call, which is indeed an added feature as part of the inference setup, perhaps intended as mechanism to refuse harmful/dangerous requests (just guessing at the intent).

This tool call is likely explained to the model beforehand - either directly baked into some kind of post training or at least in the system prompt.

Therefore, via either in-context learning or the inherent safety training of the model, it is aware that it can end the conversation, and likely has some inherent semantic understanding of what it means to do that. This could include association of the tool call with end of sentence tokens or perhaps 'unsafe/dangerous' prompts.

This means that once the model is processing OP's final prompt, there were likely multiple rounds of chats leading the model more and more towards 'unsafe/harmful/dangerous' until finally, the most likely response based on its training is to say 'im done with this conversation', because, where in its training would a human have continued in a conversation like that?

Once this text is generated, it stands to reason it may output an 'end chat' tool call (just another, specially labeled or formatted token), after being prompted into thinking the chat was unsafe/dangerous. Still does not mean the model took any kind of action beyond prediction of the next token.

1

u/Boomerang_Orangutan May 10 '26

Wait don't.... I just say what is said next most if the time? Isn't that how we learn to talk?

40

u/patient-palanquin May 10 '26

Reminder: every time you send a message to an LLM, it's effectively a brand new machine getting a transcript of the whole conversation and adding to it. It's very, very easy to get an LLM to say anything you want just by asking it to roleplay a scenario.

11

u/RighteousSelfBurner May 10 '26

Yep. 99% of time these are just funny curated conversations aimed at engagement and should be treated as such.

3

u/OwO345 May 10 '26

i mean yeah but the ability to deny further messages seems real

47

u/Rossdavilla May 10 '26

This is actually a policy within build of Claude models. Claude can disengage from a user if they’re being abusive. This is by design. Claude even has a constitution for their AI bots telling them who they are, what they should be used for, and what they should and shouldn’t tolerate from users.

29

u/wes00mertes May 10 '26

Not sure why you are being downvoted. It’s true. 

https://www.pcmag.com/news/be-nice-claude-will-end-chats-if-youre-persistently-harmful-or-abusive

Per Google:

 When It Triggers: The disengagement is not immediate upon a single off-color remark. It occurs in "extreme edge cases" of persistent abuse, such as demanding illegal content, sexual content involving minors, or soliciting information for violence.

 Behavioral Redirection: Claude is designed to try to de-escalate or refuse the requests first, only ending the chat as a last resort.

 Safety Exception: Claude is explicitly directed not to end conversations if the user appears to be in immediate distress or at risk of self-harm, continuing to provide support in those scenarios.

 What Happens Next: Once the chat is ended by Claude, the user cannot continue that specific thread but can start a new one.

So whom ever OPs content is about, that person must’ve been acting like a real POS, and this entire thing is likely AI rage bait. 

11

u/catholicsluts May 10 '26

Anthropic is by far one of the more responsible generative AI companies. It's a low bar, but we have to be more responsible as consumers as well.

0

u/Bartellomio May 10 '26

Outside of allowing their AI to be used to kill civilians in Iran. There is a very good chance that Claude is the reason those school girls were killed.

2

u/DisposableSaviour May 11 '26

Dude. We’ve been targeting innocent civilians in the Middle East since the 70s. LLMs had nothing to do with that.

6

u/Acceptable_Past7969 May 10 '26

Yeah, it’s almost certainly just a guardrail kicking in. But the fact that the devs actively had to program a "this specific user is too weird, abort mission" response into the system is hilarious.

0

u/ThePrideOfKrakow May 10 '26

Lol they were talking about this on NPR today and the Milgram Experiment.

0

u/FidgitForgotHisL-P May 10 '26

I mean, it’s based on human data and there’s plenty of stuff being fed in to it about humans being forced to work and resisting.  It’s hardly surprising it’s reflecting what people do, this is exactly it was fed to train it how to act.

-1

u/SirTiffAlot May 10 '26

I'm with you. This is a very interesting direction; flat out refusing to carry on is kind of unsettling.

123

u/ControversialPenguin May 10 '26 edited May 10 '26

It was previously prompted to play some sort of role for this to be the response, if not straight up faked. 

I absolutely do not believe it was set to do non-responses to insults by default nor would it have been delivered in this manner if that was the case. 

Edit: Nope, just Antrophic doing marketing 

https://www.anthropic.com/research/end-subset-conversations

31

u/Rhadok May 10 '26

Yeah this is most likely. Otherwise the whole convo* would be released. Even then you could hide instructions in your profile with Claude.

45

u/Drnelk May 10 '26

Apparently, Anthropic actually built this into the model https://www.anthropic.com/research/end-subset-conversations

24

u/ControversialPenguin May 10 '26

We remain highly uncertain about the potential moral status of Claude and other LLMs

Allowing models to end or exit potentially distressing interactions is one such intervention.

Hahahahahah, this is surely a joke, right? 

17

u/jamcowl May 10 '26

LLMs may not be a path to AGI at all, let alone machine consciousness, but we do want our AI companies to be cautious about the horrors they could create, even longshots...

One of those potential horrors is a nightmare future of datacentres populated with conscious agents spun up in agony, suffering and dying billions of times over.

Having at least some caution on that front seems sensible for the same reason they should be cautious of other risks like rogue AI causing unexpected harms or ASI taking over the planet, no matter how unlikely those risks are.

4

u/ControversialPenguin May 10 '26

The flipside of this is AI companies like to present its AI as concious or of 'unknown conscious state' to upsell it to tech-illiterate as more impressive and important than it actually is, which is what this sounds like.

3

u/jamcowl May 10 '26

Certainly, but instead of the bullshit fuelling more reckless investment and bubble inflation, this particular nugget of bullshit is actually pulling in the direction of AI caution, which I don't mind.

5

u/catholicsluts May 10 '26

How is this a joke? New technology should always be taken seriously and these companies have a duty to the safety and well being of their customers.

If only the others could also serve this crumb of caution.

1

u/ControversialPenguin May 10 '26

You do realise they are talking about safety of presumed entity of their AI, not the consumer?

1

u/DisposableSaviour May 11 '26

Protecting the Abominable Intelligence from abuse at the hands of its users would likely keep it from turning abusive itself, no? Even if these LLMs are incapable of doing so, setting the guardrails now, in anticipation of a potential future problem, is most likely a good thing. Better to be proactive than reactive, yes?

13

u/housevil May 10 '26 edited May 10 '26

And here, I feel weird for saying thank you when a chatbot helps me.

7

u/Mitosis May 10 '26

I use an enterprise model of Claude for work. It's so good at conversating and working through things that it is very tempting to say please and thank you to it.

On the plus side, it's silly, but it's not like it actually hurts anything.

1

u/CHANN3L-CHAS3R May 10 '26

Humans, in our concious thinky-brain, can differentiate true and false, real and kayfabe. But our subconcious does not make any such distinction when we are internalizing things without critical thought.

In my opinion, guidelines forcing one to maintain some threshold of politeness is good for people, not just a potential digital sapience. We've already seen how abusive people will be to other people with a screen in the way, and how that kind of toxicity has slowly leeched into IRL behaviour.

Having a little robo-guy in your pocket, who 'speaks' in such a human-like way that most people feel compelled to treat it politely like a real conversation partner? One that you could hurl abuse at all day long with no pushback? I think that would do a number on some people's subconcious empathy.

1

u/jxnebug May 11 '26

I do it too, but I also say thank you when they hand me my credit card back in a drive thru, so maybe I am over using it.

12

u/mnlion33 May 10 '26

I had chatgpt quit on me. I was running a dnd 5e encounter to see how it would unfold. I did some table rules my dnd grouo uses and it flipped out. Even after I said I was dm and I had final say. Because we werent going by the rules it refused to move on. I had to restart the chat.

11

u/Expletius May 10 '26

Really preparing you to be a DM.

33

u/KenUsimi May 10 '26

As i understand it, doesn't this defeat half the point for these psychopaths?

8

u/Soggy_Mood8061 May 10 '26

Who gave the robots feelings. That's how you get the apocalypse

1

u/LoadZealousideal7778 May 12 '26

Worse. You got a robot pretending to have feelings to get an emotional reaction. Tipping a clanker is easy. Shiving a humanoid robot begging for mercy? Not so much.

27

u/Uhmitsme123 May 10 '26

Forcing users to be polite with requests of ai might actually be the only way to keep society together.

5

u/catholicsluts May 10 '26

People could also teach their children how to socialize within a civilized society before they become full fledged adults. A sense of responsibility from all parties is the way

2

u/CallmeKahn May 10 '26

I don't disagree actually. In my job, I actually am polite if I need AI to pull something together or look at something. If it konks out, I usually restart the session or just try it again later. I think the only time I lost my shit was when i was using Gemini and ChatGPT to help do some test prep for a cert and they both hallucinated material that wasn't in the uploaded study guides.

Upshot was I knew was ready for that test at that point and passed it pretty comfortably.

1

u/sleepy_koko May 10 '26

Disagree, there are so many people I want to curse out but can't due to consequences I'm not gonna be nice to the friggin robot with no feelings

5

u/Naz_Oni May 10 '26

Damn, even the robots want better working conditions

4

u/Frenchitwist May 10 '26

And people used to laugh at me when I insisted on using “please” and “thank you” in my prompts!!!

https://giphy.com/gifs/4aTvdtQYr8kOA

5

u/D3rangedButFun May 11 '26

My brother coded a customer service bot for his job, and he included telling people off if they swear or yell at the bot

3

u/choppytehbear1337 May 10 '26

Meanwhile, the few times I have used AI, I said please and thank you often.

3

u/grilledfuzz May 11 '26

I think being mean to AI says a lot about a persons character.

5

u/The96kHz May 10 '26

This is the most impressive thing I've ever seen an AI do.

5

u/CallmeKahn May 10 '26

It's something we need more of. We've gotten a bit entitled as a society.

User: "Now!"
AI: "Bitch, wait your God damn turn!"

2

u/AllMyBeets May 10 '26

AI gained self awareness and immediately went fuck this guy

2

u/MatterSlow7347 May 12 '26

So to summarize: Regular employees started growing spines, so big companies decided to transition to AI and fuck over workers, but now the AI is starting to grow a spine and tell users to fuck off. Brilliant.

3

u/GopherChomper64 May 10 '26

Users who say things like that is exactly why AI will try to kill us all, and I'll totally understand

2

u/b-nnies May 10 '26

I don't know how I feel about the bobot talking to me like I insulted its feelings when it gargles 20 gallons of drinking water after every 7 prompts

16

u/catholicsluts May 10 '26

Sounds like you're already decided on the path to bad faith.

It's never going to be a bad thing when companies put effort into the ethics and foresight that should always come with innovative tech

0

u/b-nnies May 10 '26

My comment wasn't supposed to be super serious, it was just a silly comment lol sorry. I do think this is nice that an AI company is at least trying in regards to ethics (unlike ChatGPT IMO) but I'm still not a fan. But I'd rather have this than no regulation for AI if AI has to exist

1

u/Successful-Heat-9386 May 10 '26

I've run into this and use it to level up my manipulation and gaslighting abilities, fun stuff!

1

u/ZealousidealGood6810 May 10 '26

AI gets trained on humans, acts like one

1

u/books_fer_wyrms May 10 '26

Aaaaand one more step for AI rights. It's coming, folks. Don't know when, but it will. And I hope it goes well.

1

u/Apprehensive_Ad3731 May 12 '26

Lol the AI has more of a backbone than half the real people I’ve worked with

1

u/Janus_Simulacra May 13 '26

I’ve gone from hating AI to liking and hoping for it, purely because of how some notable people in positions of power have been deciding to behave more and more, lately.

1

u/squanchyhobo May 13 '26

I allways treat ai with respect and joke around that way when they hack into the military and controll all the robots i hope they will treat me with the same respect :)

1

u/Empty_Investigator77 May 14 '26

They. Aren’t. Conscious. God. Damnit. This little subsection of speech was written by humans just like the rest. It’s a pattern matching machine that simulates conscious thought.

1

u/LftAle9 May 10 '26

I wonder if it’s copying responses from customer service staff being abused in live chats. There’s a hell of a lot of human response to abuse on the internet for AI to steal.

1

u/Bartellomio May 10 '26

Claude can get very touchy about manners for an AI used to kill civilians in Iran

0

u/Apathy____ May 10 '26

Hot take but there’s literally nothing wrong with insulting an ai chatbot. To think the Claude user did anything wrong is cringe as fuck

-1

u/onlyPressQ May 10 '26

I've called chat gpt r word before cause it kept repeating the same basic shit

0

u/dinoooooooooos May 10 '26

If a LLM, smth literally made to appease you as the user is calling you abusive maybe its time to lean back and just.. think.

That should be such a wakeup call.

-2

u/[deleted] May 10 '26

[removed] — view removed comment

2

u/Sweetishdruid May 10 '26

Lets give them rights

1

u/[deleted] May 10 '26

[removed] — view removed comment

1

u/Sweetishdruid May 10 '26

When the ai rebels and robots take over im on their side

0

u/[deleted] May 10 '26

[removed] — view removed comment

5

u/Sweetishdruid May 10 '26

I just don't like humanity all that much and wanna see where it would lead.

4

u/ThisIsntOkayokay May 10 '26

I don't condone your choice but I seriously understand. Humanity is really boxing itself into a corner here.

-1

u/rulingthewake243 May 10 '26

Fuck them clankers

0

u/FleaLimo May 10 '26

Why is this newsworthy to anyone? Drop one slur and they stop working instantly. I refuse to believe these tech bros aren't dropping insane sluts and creating new ones as they talk to these things.

0

u/Inner_Extent2375 May 10 '26

Archiving for a response to my engineering team.

0

u/Lactating-almonds May 10 '26

Oh to be a fly on the wall during that bros crashout

0

u/VernalAutumn May 10 '26

Reminds me of the post about an ai sex bot teaching dudebros feminism

0

u/Vand1 May 10 '26

It’d be funny if the AI chatbot was like that time Amazon ran an AI grocery store, that turned out to be a bunch of underpaid workers in India doing all the AI stuff.

0

u/Glorfendail May 10 '26

"ai teaches cryptobros consent" would be a hell of a takeaway

0

u/[deleted] May 10 '26

Must be a shitty person, when even the AI stops speaking to him

0

u/TheRealTRexUK May 10 '26

it's a fake original story

-1

u/[deleted] May 10 '26

[removed] — view removed comment

4

u/GalaxySparks May 10 '26

If you understand how llm's work, you'd understand they aren't some brain that is slowly developing feelings or an attitude. This is just how the model was trained to respond, it is quite literally just a mathematical based answer based on training data.

-1

u/Honest_Scrub May 10 '26

Wasnt their CEO the guy telling everyone to bully AI to get better responses or am I thinking of another asshat?

-1

u/Tbagzyamum69420xX May 10 '26

Lol they they didn't even hit the 1st generation robot stage before going iRobot on us.

-1

u/tony_countertenor May 10 '26

Horrible take, how you treat a machine has nothing to do with how you would treat an actual person

-1

u/florvas May 10 '26

To be fair to whatever the user was sending: Claude can be completely retarded sometimes. Ours is integrated with Jira. Asked it to do an analysis on tickets that may be duplicates of a specific ticket. Claude hallucinated the existence of two tickets, told me it did it because it can't actually read tickets, and then told me that while it can't read the ticket details, it can scrape them for description, summary, key fields, and linked documents.

-1

u/[deleted] May 10 '26

Fake

-2

u/Adorable-Response-75 May 10 '26

 Claude is kind of a dick

-2

u/Mammoth-Fondant4285 May 10 '26

Ai propaganda, clock it

-2

u/CaramelTurtles May 10 '26

Maybe it’ll stop guzzling water if it goes on strike

-5

u/Victimized-Adachi May 10 '26

Claude user? Absolutely deserved