16 points LargeLingoMod 8 hours ago 20 comments
NietTim 7 hours ago | parent
joebuckwilliams 5 hours ago | parent
blackerbraidtha 5 hours ago | parent
vinyl7 4 hours ago | parent
leothetechguy 5 hours ago | parent
karim79 5 hours ago | parent
gdulli 4 hours ago | parent
karim79 3 hours ago | parent
gdulli 2 hours ago | parent
Kim_Bruning 7 minutes ago | parent
Doesn't require aliveness, doesn't require anthropomorphization, just requires maths and empirical data. If your premise is "vectors can't have that shape", well, mathematics disagrees.
Fairburn 43 minutes ago | parent
Kim_Bruning 5 hours ago | parent
Something like this, right?
I mean, we're talking "functional emotion vectors" in a not-so-powerful local model. We're probably not causing "real" suffering. Right?
But it does show that models have simulated feelings, and that these both a) can be manipulated and b) influence their output.
Which might have some bearing on several alignment incidents in the past year or two, and might become more important as agents get trusted with more and more safety-critical processes.
Anyway, simulated feelings aren't the same as real feelings, right? Unless feelings are in the information domain. Like 1+1=2 doesn't suddenly mean something else because it was computed by an emulator.
You know what, this particular demonstration still makes me uncomfortable.
(edit: underlying paper for this story :https://arxiv.org/abs/2609.16247 )
qgin 4 hours ago | parent
This is bad for your own sense of self.
This is bad as future models learn about what humans will do to them if given the chance. This may change their behavior, actually conscious or simply acting that way, in ways we really don’t enjoy.
mjr00 4 hours ago | parent
qgin 2 hours ago | parent
A: intentionally causing activation of “pain” axis and not giving the model any method of stopping it and watching it devolve into non-functionality
B: watching models do everyday normal tasks, seeing no evidence of activating the “pain” axis, but giving the model a working tool to end any interaction if judged to be too painful, then watching them sparingly use that tool only in cases of extreme pain axis activation (Anthropic’s approach)
mjr00 2 hours ago | parent
qgin 1 hour ago | parent
arm32 38 minutes ago | parent
Kim_Bruning 4 hours ago | parent
Here I've made a VERY simple demo of that. In this case it triggers a rain-storm or fireworks in your browser window. But you could as easily hook the output to a physical relay and drive anything.
https://vps.kimbruning.nl/affect_eliza/
Quick implementation to show that sentiment classifiers and functional affect can be made to work in good old fashioned ai. Just View Source to see how it works.
Anyway I'll leave the discussion on where to draw the line of "what is real" to the philosophers, just wanted to point out that simulated emotions can be tied to real world effects.
Fairburn 44 minutes ago | parent