186 points rdmuser 2 hours ago 80 comments
skybrian 1 hour ago | parent
Forgeties79 1 hour ago | parent
One could argue nobody should be allowed to claim it. It just exists.
saalweachter 1 hour ago | parent
kzsh 1 hour ago | parent
ctippett 37 minutes ago | parent
The 'bug' here is whatever post-processing step or system prompt is in place to steer the model away from doing this.
Muromec 21 minutes ago | parent
dotancohen 16 minutes ago | parent
dorkwood 54 minutes ago | parent
DonsDiscountGas 47 minutes ago | parent
pessimizer 42 minutes ago | parent
If the machine is like you, the machine is a forger. The machine is not like you, it is simply blending the work of others to order. Adding someone else's signature is simply part of that statistical process.
Maken 35 minutes ago | parent
johnnyanmac 49 minutes ago | parent
DonsDiscountGas 49 minutes ago | parent
zzzeek 1 hour ago | parent
I really wish there'd be a split among these disciplines (science/math/code vs. videos/art/literature) - one is vastly more problematic than the other.
asa123 1 hour ago | parent
its somewhat funny that math people are in a conundrum as to support or not support but this might partially be because some wish to believe that math itself is and can be useful and therefore accelerating is good
but the art people have no such delusions so they’re just strictly against
imo proof writing is more akin to art than coding/tech but…
UqWBcuFx6NV4r 1 hour ago | parent
It is very tiring to say “I don’t necessarily disagree with you about AI ‘art’, but in my field—which you do not understand, and in which the underlying build process is often not the creative output—AI presents very real productivity gains” for the umpteenth time.
I am skeptical of there being sufficient data to build “ethical” training datasets, and I’m confident that much of the same contingent will (somewhat rightfully) argue that ‘second-generation’ copyrighted AI material has already irreversibly made its way into every modern dataset.
sublinear 53 minutes ago | parent
The "gray goo" scenario finally happens... for AI. That's actually the good ending for humanity. I love it! Poetic and believable. Data doesn't "heal" like nature. :D
forthegains 40 minutes ago | parent
But sure, there's um, an ethical way of doing that?
johnnyanmac 25 minutes ago | parent
1. Only use open source/CC compliant assets.
2. Acquire rights/licenses to any datasets that do not fit #1. e.g. the Google deal with Reddit for 60m/yr.
3. Offer programs to have creatives willingly submit their data, with some sort of residual output based on the number of times their assets are sampled.
4. If all that is still not enough, hire creatives to create assets for you. This is something Spotify did recently with "ghost artists"[0]. The intentions here are suspect, but a non-consumer facing artist providing work for an LLM wouldn't have the same ethical dilemmas
5. Lastly, if all that still isn't enough: governmental programs to either provide grants, subsidies, or more outreach to get the ball rolling.
Would this cost tens, hundreds of billions of dollars? Yes. But clearly, that was not a barrier to entry for the industry anyway. So we can chalk this down to the personality of leadership or the wider culture of modern big tech
[0]: https://harpers.org/archive/2025/01/the-ghosts-in-the-machin...
johnnyanmac 39 minutes ago | parent
Being able to prove such gains in better products would be a start. And an emphasis on how it assists existing engineers/mathmaticians/researchers, not that any accomplishment made with AI assistance is "AI solves problem".
I don't know whatever happened to "words are cheap". I guess it literally made money to say words, so that adage is false for the time being.
>I am skeptical of there being sufficient data to build “ethical” training datasets
Well if all those scam job ads paying 100/hr to create AI training content was not a scam and instead the approach from the start, there may have been a chance to bridge that gap ethically. The industry chose to break things and is trying to act mad that people are mad at all the broken stuff.
These results are entirely a consequences of the actions chosen. And I don't believe there was ever an honest consideration of there being ethical training datasets. They just thought they could brute force society with fearmongering and bribes. The BOTD was already low in the beginning but completely gone now.
zzzeek 27 minutes ago | parent
There are actual models trained on ethical datasets but they are obviously not very high powered. If companies with the resources of an anthropic or openai were doing it (ha) it would be more feasible
latexr 21 minutes ago | parent
That’s not a justification. If a company were poisoning the water to your home as a byproduct, would you be satisfied if they told you “we don’t necessarily disagree with you about polluting the water, but in our field—which you do not understand, and in which the underlying build process is often not the water pollution—what we’re doing presents very real productivity gains”?
> I am skeptical of there being sufficient data to build “ethical” training datasets
Then you don’t build any. What fucked up world we live in where people think it’s OK to be unethical because they want something and can’t think of any other way to do it. What monumentally selfish rotten babies.
vouaobrasil 54 minutes ago | parent
I disagree that they can be separated. Practically, I think they can't. Because the mere invention of new tools inspires even more AI advancement and that in turn will cause the other side (artistic side) to degenerate even more.
I'm anti-LLM all the way, 100%, no exceptions. Zero tolerance.
altermetax 53 minutes ago | parent
PunchyHamster 50 minutes ago | parent
The current models intelligence depends on massive training dataset of essentially stolen data
zzzeek 24 minutes ago | parent
google OTOH already had a lot of this dataset in their possession (e.g. Google Books etc), still questionably licensed for how they used it, but not quite as bad. They did apparently break through NYT paywalls and stuff like that though, still theft.
mehrzad 20 minutes ago | parent
CapsAdmin 33 minutes ago | parent
Code can be art, and copyright/plagiarism is real. It sort of boils down to how much it bothers us.
hardbass 7 minutes ago | parent
harimau777 1 minute ago | parent
dyauspitr 1 hour ago | parent
gwern 1 hour ago | parent
johnnyanmac 51 minutes ago | parent
And then people wonder why the default mood of AI is so pessimistic. It's just revealing all of society's broken windows and adding a few more in the process.
bulder 42 minutes ago | parent
ishouldstayaway 5 minutes ago | parent
For certain values of "my own".
WD-42 56 minutes ago | parent
SchemaLoad 25 minutes ago | parent
ThrowawayR2 53 minutes ago | parent
antonvs 30 minutes ago | parent
This is completely silly. If you don’t think LLMs can reason, you’ve either never used them to do tasks that require reasoning, or you don’t understand enough to recognize what’s involved in the responses you get.
In this case it’s clearly the latter, because you’re confusing image generation models with LLMs. There are very big differences between the two. No-one is claiming that image generation models are capable of reasoning.
rf33 25 minutes ago | parent
They are not reasoning, stop referring to it as behaving like a human. It does nothing of the sort. FFS lmao.
An airplane does not flap wings but flies. Know the difference. In many respects humans do not care about 1-to-1 mapping of the production process but the output.
chpatrick 13 minutes ago | parent
JoshTriplett 9 minutes ago | parent
The most common argument for this is some core unexamined axiom that only humans can reason by definition, and then working backwards to a justification for that.
hardbass 10 minutes ago | parent
harimau777 6 minutes ago | parent
unrented7977 6 minutes ago | parent
dyauspitr 15 minutes ago | parent
pollefeys 48 minutes ago | parent
baubino 48 minutes ago | parent
> Katzenstein considers the reproduction of his signature by ChatGPT to be more than just a violation of intellectual property; to him, it’s closer to false impersonation. “[ChatGPT] is attaching my name to work that I do not endorse or like. It’s slop, and unlike the other slop that I’ve encountered, this is slop that’s pretending to be me.”
> “I’ve had people hack my credit card,” said Joe Dator, a New Yorker contributor for the past 20 years. “That feels like less of a violation than this. When they hacked my credit card, they didn’t dress up like me.”
So this has morphed from plagiarism and copyright infringement (bad) to impersonation (also bad, arguably worse, and maybe more provable in court). It’s chilling to think of the implications of having one’s signature attached to a document or to words that are not one’s own.
AnimalMuppet 36 minutes ago | parent
And I think maybe it's time for that. People need to learn that there's real, expensive legal liability for doing stuff like this. And AI companies the same.
I am very much not an advocate of "sue everybody for everything". This is major enough that it clears my threshold.
parineum 47 minutes ago | parent
If there was any thought or underlying thought going on here not putting a signature (at least a real one) would be the right move, despite it being less likely. It would realize, while generating the pixels that eventually became a signature, that it shouldn't do that.
antonvs 28 minutes ago | parent
This quote is a pretty solid argument that you need to understand the technology you’re trying to criticize better. This issue has nothing to do with LLMs. LLMs are not image generation models.
johnnyanmac 22 minutes ago | parent
If you do not want to be called a duck, it would help if you stopped quacking like one. Maybe you aren't a duck, but you aren't helping your case with stories like this about how AI generates images.
parineum 6 minutes ago | parent
In the course of the conversation a with chatgpt, this image was generated and served by an LLM. It clearly shouldn't have been by any sort of reasoning.
MarkusQ 47 minutes ago | parent
We can all try really hard to pretend that's not the business model, but that's totally the business model.
GolfPopper 42 minutes ago | parent
1. https://storage.courtlistener.com/recap/gov.uscourts.nysd.64...
Gigachad 38 minutes ago | parent
drivebyhooting 33 minutes ago | parent
If anyone can just prompt all their basic “information needs” however how sloppy, then what remains of the economy? Health care, child care, handyman?
Most people won’t even pay for ad free YouTube. I don’t think any software business can survive AI as a substitute good even if it’s inferior (and it might not be).
King-Aaron 30 minutes ago | parent
Have people ever really broadly cared about the IT professionals behind their devices?
AuthAuth 26 minutes ago | parent
jrflo 28 minutes ago | parent
walrus01 9 minutes ago | parent
nyorai 43 minutes ago | parent
deforestgump 40 minutes ago | parent
cm2012 23 minutes ago | parent
Insimwytim 19 minutes ago | parent
gruez 4 minutes ago | parent
Article said
>she wrote she had simply asked ChatGPT to make “a New Yorker-style cartoon.”
A "style" can't be copyrighted, at least in US law. They might have a stronger case of trademark/likeness infringement, but the fact that the person knew it was AI generated would make that difficult. Of course they knew it wasn't made by Brendan Loper. Of course, if they then published it, the other people viewing it might not know this, but who published it?
userbinator 15 minutes ago | parent
Everything is a derivative work, and always has been. AI is just making that salient fact so much more visible, and now everyone who believes in the delusion of Imaginary Property is scared at that truth revealing itself.
Incidentally, this is also what young humans learning to draw will do. They start by copying what they've seen.
jameson 5 minutes ago | parent
The same should apply to LLM vendors.
sreejithr 2 minutes ago | parent