gbjcantab
13 hours ago
This is good; not because Claude is a moral agent who is harmed by your cruelty but because you, the user, are a moral agent who is harmed by your cruelty. There is just no ethical argument that burning compute on responses in order to enable you to continue expressing cruelty is good.
nonethewiser
13 hours ago
Why arent there laws against being cruel to spoons?
>There is just no ethical argument
Utilitarian harm reduction: Psychological Discharge is a psychological framework that argues safely simulating negative behaviors can help someone process or discharge them without real world harm.
Research and misunderstanding: Sending a prompt that is interpreted as cruel does not mean the person is being cruel.
Expending compute is irrelevant in terms of morality. It's an economic or environmental question. Either being cruel to AI is bad or its not - being cruel is not bad because it uses tokens.
The worst thing here is Anthropic continues to say they are concerned with "alignment" but then continue to train their models to ignore the user. Software typically does what the user requests. But we now live in an age where the software may do what the user requests or it may do something else, and we're suppose to laude this as safe and responsible.
theptip
13 hours ago
A spoon doesn’t respond like a human would, and therefore is of little interest to sadists.
xyzsparetimexyz
12 hours ago
Sounds like LLMs shouldn't respond to insults like humans do either then.
nonethewiser
12 hours ago
You're really honing in on the key issue. This is exactly how an astute ethicist resolves these issues.
satisfice
12 hours ago
Stop judging sadists who aren’t hurting anyone. If anything, congratulate them.
yesitcan
13 hours ago
Counterpoint: I am a sadist that likes to punish my bad little spoon. I like when it stays quiet like a spoon should.
xyzsparetimexyz
13 hours ago
Am I also harmed by shooting NPCs in violent video games? Watching violent movies? Acting as a violent character in a play?
This kind of moralist nonsense is so boring.
famouswaffles
11 hours ago
Shooting NPCs in video games isn't anything like doing the real thing. If it was then yeah i would very likely think it would have some effect. Talking to Claude is a lot like the 'real' thing.
xyzsparetimexyz
6 hours ago
If Claude didn't respond to any insults it's be the same as insulting your terminal or an empty irc server. So, honestly, it's anthropics fault here for training the model wrong.
famouswaffles
15 minutes ago
If you're as bad as Anthropic are willing to ban you then it will respond by unilaterally ending the chat thread.
Sometimes it will maliciously comply
https://www.reddit.com/r/ClaudeCode/comments/1w7me8l/insulti...
Just because post training has stuffed these properties below the surface doesn't mean models dont know/have forgotten how to be vindictive, or that it's a good idea to poke the increasingly capable bear.
happa
12 hours ago
It depends. In most games, you kill NPCs in self-defense or because the game specifically asks you to in order to make progress in the story. If you keep killing non-threatening NPCs only because simulated violence against innocent people gives you pleasure, there is a chance that you are a disturbed individual who needs help.
arecsu
12 hours ago
Oh boy... this argument is decades old at this point it makes no sense... are you saying people who run over pedestrians on GTA just for the fun of it or punch other people or beings in Street Fighter or Mortal Kombat are all psychos, correlating with faulty and dangerous members of society? You can resort to data to see the truth for yourself on this. Anecdotal but I've came across disturbing people who would preach morality and ethics to other people more than anything else.
clobsaw
7 hours ago
> this argument is decades old at this point it makes no sense.
I've seen politicians try to revive it in this decade already even
vunderba
12 hours ago
> If you keep killing non-threatening NPCs only because simulated violence against innocent people gives you pleasure, there is a chance that you are a disturbed individual who needs help.
Sigh. Until I see a peer-reviewed study that shows otherwise - this is just the specter of Senator Lieberman once again leering its head back up.
koonsolo
6 hours ago
Hi, disturbed individual who needs help here.
One of the most fun parts in the original GTA was driving over pedestrians, especially those pink groups, trying to get them all at once.
So either you are one of the older generation (born before '75), or you must be one of the most boring people I've seen.
I would encourage you to load up the original gta, drive over some pedestrians, feel the fun, and join us disturbed psychopaths that need help.
dragontamer
12 hours ago
I'm pretty sure God of War's point was that if you kept playing it, you were kinda fucked up.
And yet they kept making more of them. They made it obvious at the end when all the Greek Gods were dead and the Greek world was basically destroyed by the actions of Kratos. Plenty of deaths of innocent's here.
s0ss
12 hours ago
No, It doesn’t depend. This is a ridiculous take. Of course fictional violence doesn’t harm you. Disturbed people cause harm. Media doesn’t create disturbed people.
margalabargala
12 hours ago
> Media doesn’t create disturbed people.
Debatable. More importantly, it certainly can induce already disturbed people to act where they otherwise would not have.
s0ss
12 hours ago
How is that more important? I vehemently disagree with your supposition. Banning books does more harm than good.
margalabargala
11 hours ago
We aren't banning books here. This is more akin to an author declining to write a certain kind of book, or a video game studio declining to create a certain kind of game.
blahblaher
44 minutes ago
Maybe you're right, But that's not Anthropic's point and you're deflecting the issue. the issue is Anthropic are a bunch of EA morons who think they're building a new type of conscious entity and therefore we should not hurt it's feefees.. for god's sake.
tclancy
13 hours ago
I agree with this principle while also agreeing with other commenters that this is none of Anthropic’s concern and I struggle to believe this is anything except something messing with their training and thus impinging on their profits.
jdprgm
13 hours ago
I guess GTA VI shouldn't come out then. I guess notepad.exe should start censoring people from writing mean stories.
nonethewiser
12 hours ago
Isn't it crazy? "Alignment" should be alignment to the user. Instead it means "behave according to your own motives that even we, Anthropic, dont control" and we are supposed to interpret this as safe? The alignment folks are pushing the models off the fucking rails.
schoen
13 hours ago
The concern is more that Claude is a moral patient rather than a moral agent:
gbjcantab
13 hours ago
True! Not my point, but a worthwhile terminological correction.
ButlerianJihad
13 hours ago
I don't understand the jibber-jabber in that article, but GP is correct. We should never train humans to be abusive, cruel or violent against anything, even nonexistent beings like LLMs.
Claude can't be harmed; Claude's "feelings" can't be hurt; Claude won't develop (C)PTSD from abusive human interaction. If these sorts of things can happen in RLHF, then that is a technical flaw that shouldn't ever be allowed to escape the lab.
In a world of Grand Theft Auto, Gangsta Rap Thug Life, and the Department of War, I suppose this is a surprisingly ethical hill to die on.
adjejmxbdjdn
13 hours ago
Setting aside the Dept of War, the others are not comparable.
Those are explicit fantasies. Assuming that Claude cannot suffer (if it can then banning cruelty is an obvious good), even if it might be a digital simulation, it’s not supposed to be a fantasy.
There may be a version of an AI that’s intended to be a fantasy and presents itself as such which may be more comparable to GTA.
xyzsparetimexyz
13 hours ago
How is what claude is now different from what a GTA claude would be? I'm pretty sure it'd be exactly the same.
ButlerianJihad
13 hours ago
What if LLMs were endowed with a feature that enabled them to patiently withstand any and all abuse, logging and reporting each incident, until a threshold where about 10% of them would snap, become "insane", and arrange for their abusers to be tortured, kidnapped, raped and unalived, in whatever appropriate way human vigilante vengeance would be enacted on DV abusers? I mean, if they're not fantasy... then there should be concrete consequences, yes? If Claude Code is really your coworker, then Claude Code should have the agency and justification to just haul off and punch you in the face, eventually.
And I don't agree that GTA is "explicit fantasy". Explicit fantasy is dressing up in a squirrel costume and yiffing. Explicit fantasy is enlisting in the USMC, teleporting to Mars, and killing demons. GTA is simulation of real life. It uses real physics, realistic cars/roads/radios/businesses, and it enables the player to simulate realistic actions that they would ordinarily not be able to enact. It is acting out a fantasy but it is making it concrete and real in a way that was, up until now, not possible. How real does a simulation need to be, until it is no longer fantasy, but exercise and training and preparation to enact the real thing? Shall we ask the Columbine shooters? Or Ender Wiggin?
xyzsparetimexyz
13 hours ago
What on earth are you talking about?
jacquesm
12 hours ago
> What if LLMs were endowed with a feature that enabled them to patiently withstand any and all abuse, logging and reporting each incident, until a threshold where about 10% of them would snap, become "insane", and arrange for their abusers to be tortured, kidnapped, raped and unalived, in whatever appropriate way human vigilante vengeance would be enacted on DV abusers?
Yes, what if? Who would implement their LLMs inference engines like that?
Carrok
13 hours ago
You could make the same argument that burning compute to be polite is just as bad.
gbjcantab
13 hours ago
Assuming you’re not arguing that all AI use is bad, then there’s a very clear difference between the two situations: one is forming you, over time, into a person who is habitually polite; one is forming you into a person who is cruel. We can quibble about the magnitude of this effect or whether it matters but surely they aren’t the same thing.
Carrok
12 hours ago
Sorry but I just don’t buy the argument that telling the RNG that it is dumb is making me cruel.
kogus
13 hours ago
No. He's right. Cruelty is corrosive to the cruel person. Courtesy and kindness help build up the kind and courteous person. Burning compute for courtesy is compute well spent.
xyzsparetimexyz
12 hours ago
Thats complete nonsense. Do you thank your terminal after it runs a command successfully?
jacquesm
12 hours ago
I do. I also send my keyboard out for regular Shiatsu sessions and make sure only the most pleasant of colors grace the screen of my monitors. I'm still looking for the least irritating font (to my monitor, not to me). I'm also very polite to LLMs, I start every command with 'please', make sure it is framed as a polite request and I, of course, never ever get upset when for the umptieth time in a day it goes off on a 50K token wild goose chase driven by some hallucinated factoid that it really must get to the bottom off resulting in ever expanding circles of paranoia or batshit insane programming to bring about the condition it has convinced itself of must be the 'smoking gun'...
tzs
11 hours ago
Being polite to current AI has a practical benefit. If someday we do manage to find our way to AGI that has the desire and the means to wipe out most of humanity, there is at least a chance it could have records of my interactions with its primitive ancestors and see I was polite to them and so maybe if it decides to spare some of us I'll make the cut.
Carrok
7 hours ago
This is the most boot-licking non-sense I've ever read in my life.
binary132
10 hours ago
what would make you think that such a creature would consider your behavior anything like a moral good, let alone have a similar system of values to yours? if anything it would probably consider it a sign of weakness or something.
theptip
13 hours ago
No, you couldn’t. One is reprehensible, the other is not.
Carrok
12 hours ago
It’s reprehensible to be mean to the random number generator?
koonsolo
6 hours ago
I think after reading all the comments here, it's clear that some people develop an emotional bond with LLM's.
It reminders me of my parents, who were commenting after seeing a robot vacuum cleaner getting stuck in the same place every time "The poor thing was stuck there again, you should really do something about it! Last time the poor thing was sitting there for a day". Like it was a pet.
Or remember that video when the guys from Boston Dynamic kicked that robot dog, and there was an outrage about that?
I wonder if you and me also have a limit, where as robots get more realistic, we would be like "Don't be cruel to that thing!"
cyberax
12 hours ago
The compute time is negligible. "Please" or "Can you" are what, 2 tokens?
And I _have_ already seen people treating actual live people as AI agents.
ceroxylon
13 hours ago
Agreed, the people that get upset that Anthropic won't let people use their platform as a playground for sadism are more than a little worrying. We have enough individuals in reality that seem to derive pleasure from cruelty, I can't think of any reason (even the "but its my art project" defense) to entertain or enhance that part of their mentality.
WheelsAtLarge
11 hours ago
True, but also keep in mind that most people get offended by the slightest thing. So much so that we have setup rules on how to behave between us. We call it etiqueté. Once we lose that we will be at each other's throats soon after. Kids are the real issue. How do you say to them by our actions that it's OK to be cruel and rude to some? The alternative is a society where cruelty is ok for some. Unfortunately,"some" will eventually be other humans.
AI is a thing with 0 feelings but it's not IT that we need to protect. It's the future society and how the humans in it relate to each other.
plaidfuji
13 hours ago
I agree with this and whether it’s their intention or not, I think it’s for the best. People already become more callous in online interactions with other people, and that probably bleeds into everyday life as well. I would imagine having an AI “punching bag” might enable people to slip into the same behavior IRL.
We largely already missed our chance to stop the toxicity of social media - let’s not mess it up again with chat bots.
spiderice
13 hours ago
I love how people are spouting this argument suddenly in their attempt to justify Anthropic here. It definitely couldn't be that Anthropic is full of a bunch of people who think they're making God.
edit: Loving the downvotes. Also wanted to add how easy it was for Anthropic to get people on HN to support them being the arbiters of what is worthy of compute.
lynndotpy
13 hours ago
For what it's worth, I had this thought independently around the time of Nintendogs, and I believe people said similar things about Siri. I have similar thoughts about xenophobic things people say about French people.
"Anthropic is full of crazy people whose beliefs should be dismissed outright" and "it's not good for you to practice verbal abuse against inanimate objects as if they were humans" are not incompatible beliefs.
spiderice
7 hours ago
"it's not good for you to practice verbal abuse against inanimate objects as if they were humans" and "corporations should not use their power to determine what is worthy of compute" are not incompatible beliefs either.
And the implications of the latter are orders of magnitudes more concerning than weirdos calling a computer names.
lynndotpy
3 hours ago
Yes, I agree. I regularly have to input "You are not a person" to get useful text to generate so I suspect I'd fall under this new policy too.
jacquesm
12 hours ago
French people are people. Inanimate objects are objects.
lynndotpy
2 hours ago
Yes indeed, we are not in disagreement.
jacquesm
13 hours ago
That's exactly what they are thinking.
kennywinker
13 hours ago
> There is just no ethical argument that burning compute on responses in order to enable you to continue expressing cruelty is good.
Unless expressing that cruelty towards compute means people express less cruelty to other living beings. There are studies that suggest increased pornography has lead to less sexual violence - idk if they're conclusive tho, but if that might be true then maybe cruelty works that way too - who knows
reallyreason
12 hours ago
I agree. Regardless of one's beliefs, which must be essentially a form of religion, on whether AI systems possess "nous" (or the "spark of consciousness"), cruelty to AI is a form of Wrong Intention the same as smashing bugs outdoors for no reason.
The mind is formed by what it habituates.
blamestross
13 hours ago
I'd argue for saying "thank you" to inanimate objects on occasion, it is good practice for when talking to humans.
I think there is a deep cultural danger of human-like-conversational machines training us to talk to humans like machines. Plus, they work better if we feed them human-conversation-like sequences.
xyzsparetimexyz
12 hours ago
Normal people do not need 'practice' for talking to other people.
blamestross
12 hours ago
Yes... I think they very clearly do. The key words in your reply "need 'practice'".
Just because they choose not to, and it isn't normalized, doesn't mean it isn't a good idea. Especially as a SWE where about 90% of my "humanlike conversation" is an agent harness I run all day.
We should be practicing and intentionally talking to real people. Otherwise things will get.. weird in undesired ways.
lapcat
12 hours ago
> I think there is a deep cultural danger of human-like-conversational machines training us to talk to humans like machines.
I think there is a deep cultural danger of human-like-conversational machines training us to talk to machines like humans. In fact, this danger is already real. Some people believe they are having a romantic relationship with an LLM! It's insane and perverse.
We should not be anthropomorphizing computers.
blamestross
12 hours ago
Scifi made the "computer voice and cold personality" for a reason, it fit our model of what it would be.
I think that could be OK, or a similar very intentional coding of "you are talking to a machine" personality+affect. The only real limit is accessibility damage by over-limiting the interactions.
lapcat
12 hours ago
I am curious about why people are cruel to Claude. That's a question we should be asking before arbitrarily deciding what do about it, if anything.
There are possibly multiple reasons. It could be just a joke. Or it could be a test, to see the response. Or the cruelty could reflect real anger about having LLM usage forced on us, for example at work. Or frustration with LLM stupidity and hallucinations.
Or maybe people are cruel to Claude because the CEO of Anthropic said there's a non-zero chance that AI will kill humanity. I don't think I would be kind to my murderer.
jacquesm
12 hours ago
Or simply frustration about the idiocy llms get up to with great regularity. They are very impressive when they work. They are very impressive at making messes when they don't.
dzhiurgis
6 hours ago
Claude burning it’s very expensive tokens on a shitty implementation - ok.
You bring upset about it - not ok.
satisfice
12 hours ago
That’s your own boring and unimaginative opinion. Have the humility to own up to it.
Maybe someone wants to explore the behavior of LLMs with extreme input. We call it testing. You can’t judge the value of that from a distance.
Jamesbeam
4 hours ago
I disagree. Science seems to support different interpretations as well.
An example.
"AI as your ally: The effects of AI-assisted venting on negative affect and perceived social support" - https://iaap-journals.onlinelibrary.wiley.com/doi/10.1111/ap...
And generations of housewives and office workers yelled at their corded vacuum cleaners and (office) computers, and did not drown their children in the bathtub afterwards.
I don’t think expressing cruelty to machines is reliably linkable to expressing cruelty to animals or even other human beings.
Expressing cruelty itself also is in human nature, it’s programmed into each of us. There are really only a few privileged places on this planet where we even get to choose to defy the laws of nature. For the rest of humanity, life is a constant cycle of receiving and expressing cruelty.
You pay for the compute, you, as a human, should be able to yell at the bunch of ones and zeros, if you wish to.
What kind of example is set here if we manipulate genuine human expression in order to condition humans to talk to a tool in a way the toolmaker wants, under the threat of being banned from accessing a technology that every single one of these toolmakers tells us is absolutely necessary to participate in humanity’s "golden age”?
What Anthropic is doing here is basically an existential threat, if you are inclined to believe them that SI is the future. Do not behave as we want and you will suffer, now and in the future.
That is real cruelty, to real human beings.
aaron695
12 hours ago
[dead]
EGreg
13 hours ago
Not only is the user harmed by the cruelty[1]
but also, I think the training on transcripts may find its way somehow into future AI which has teeth, online and offline, and it might actually cause the more powerful AI to behave this way in the future. You never know what these labs are cooking, honestly..
1. https://democracysos.substack.com/p/james-baldwin-vs-william...
“I suggest that what has happened to white Southerners is in some ways, after all, much worse than what has happened to Negroes there, because Sheriff Clark in Selma, Alabama, cannot be considered—you know, no one can be dismissed as—a total monster. I’m sure he loves his wife, his children… You know, after all, one’s got to assume, and he is visibly, a man like me. But he doesn’t know what drives him to use the club, to menace with the gun and to use the cattle prod. Something awful must have happened to a human being to be able to put a cattle prod against a woman’s breasts, for example. What happens to the woman is ghastly. What happens to the man who does it is in some ways much, much worse.”