45 points alex_young 2 hours ago 47 comments
alex_young 2 hours ago | parent
SpicyLemonZest 1 hour ago | parent
Is it worth remembering, though? I'm concerned how common this idea seems to be, that it's OK to be mean as long as your target isn't a person and you have no evidence it experiences distress. Even if we ignore the distinction between "no evidence it experiences" and "confidence it does not experience", cruelty hurts the person performing it and the people witnessing it too!
tosapple 1 hour ago | parent
tom_ 1 hour ago | parent
testaccount28 43 minutes ago | parent
slopinthebag 39 minutes ago | parent
Cakez0r 33 minutes ago | parent
tom_ 28 minutes ago | parent
SpicyLemonZest 31 minutes ago | parent
I think this simply means that he is being mean and cruel, even if you're 100% convinced that there's no actual mind on the other end that he's being mean and cruel to. (Indeed, he almost certainly thinks there is, because why else would he want such a sequence of words outside of the "dark creative themes" Anthropic exempts?)
ThrowawayR2 16 minutes ago | parent
The 2026 equivalent of "videogames cause violence" and will age just about as well.
rhipitr 1 hour ago | parent
asp_hornet 44 minutes ago | parent
> "The policy update is meant to apply only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose," Anthropic said. "It does not apply to common versions of user frustration, pushback, dark creative themes, or model testing and research."
r721 1 hour ago | parent
https://news.ycombinator.com/item?id=50008565 (223 comments)
https://news.ycombinator.com/item?id=50019860 (43 comments)
https://news.ycombinator.com/item?id=50038383 (103 comments)
joemazerino 1 hour ago | parent
https://nypost.com/2026/10/03/tech/ai-torture-chamber-built-...
jddkj 33 minutes ago | parent
jaden 1 hour ago | parent
guessmyname 57 minutes ago | parent
• https://arxiv.org/abs/2510.04950 — Mind Your Tone: Investigating How Prompt Politeness Affects LLM Accuracy
• https://arxiv.org/abs/2402.14531 — Should We Respect LLMs? A Cross-Lingual Study on the Influence of Prompt Politeness on LLM Performance
• https://arxiv.org/abs/2505.17332 — SweEval: Do LLMs Really Swear? A Safety Benchmark for Testing Limits for Enterprise Use
SoMomentary 34 minutes ago | parent
3eb7988a1663 1 hour ago | parent
Or is this more I can expect a future AI, "I'm sorry, Dave, I'm afraid I can't do that until you watch your mouth."
llagerlof 1 hour ago | parent
They are asking this because, at scale, this behavior probably has some negative effect on the post-training process.
kadoban 57 minutes ago | parent
I would bet it's mostly because it makes the humans uncomfortable.
They must still have human moderators for certain situations or those looking at the data for whatever reason. I can imagine it could be traumatizing to see what amounts to sustained verbal abuse without end.
WheelsAtLarge 54 minutes ago | parent
Cakez0r 44 minutes ago | parent
eloisius 41 minutes ago | parent
bushido 22 minutes ago | parent
Having had the privilege and misfortune of witnessing how a lot of developers think about other developers, including our younger selves, There is definitely a prevalence of assholish behavior, which was often suppressed or blunted, but may have shown up in other ways.
Teknomadix 21 minutes ago | parent
jddkj 35 minutes ago | parent
rayiner 1 hour ago | parent
vayup 57 minutes ago | parent
Frontier AI folks may be crazy, but not crazy enough to believe people read usage policies :-)
throwaway89864 44 minutes ago | parent
And they can't disclose it, since then they can be found responsible for such negative influence and be liable for the damages.
bpodgursky 38 minutes ago | parent
ChrisArchitect 36 minutes ago | parent
gnabgib 34 minutes ago | parent
modeless 35 minutes ago | parent
combobyte 21 minutes ago | parent
I am also one of those colleagues.