123 points Brajeshwar 15 hours ago 373 comments
Avicebron 15 hours ago | parent
jameshart 15 hours ago | parent
Lerc 14 hours ago | parent
jameshart 14 hours ago | parent
agos 13 hours ago | parent
jameshart 12 hours ago | parent
Lerc 7 hours ago | parent
Going behind a paywall is a reduced distribution over what an individual can easily access, and the content is no longer permanent but subject to whatever the publisher chooses to keep providing.
jameshart 7 hours ago | parent
mupuff1234 14 hours ago | parent
ceejayoz 14 hours ago | parent
Zambyte 14 hours ago | parent
jeremyjh 14 hours ago | parent
HarHarVeryFunny 14 hours ago | parent
Zambyte 3 hours ago | parent
(And, despite what the president of the United States mandates, we have not actually achieved super intelligence yet).
iugtmkbdfil834 14 hours ago | parent
mattm 14 hours ago | parent
iugtmkbdfil834 13 hours ago | parent
verdverm 13 hours ago | parent
verdverm 13 hours ago | parent
angoragoats 14 hours ago | parent
iugtmkbdfil834 14 hours ago | parent
BLKNSLVR 14 hours ago | parent
What are the concerns of individuals in comparison to the overall progress of humanity?
Always overlooked counterpoint: what point is the progress of humanity if it doesn't take into account the concerns of the individuals?
This pattern is playing out with increasing frequency.
agos 13 hours ago | parent
AnimalMuppet 14 hours ago | parent
And if we can't solve it for an AGI, what are we going to do with an ASI?
altmanaltman 14 hours ago | parent
Yeah so that's never going to happen
jeremyjh 14 hours ago | parent
dkasper 14 hours ago | parent
jeremyjh 14 hours ago | parent
YetAnotherNick 14 hours ago | parent
The case will be lot more complicated if someone uses Kimi to hack into a site. Should the person giving agent the command responsible or the CEO of kimi.
gyt2 14 hours ago | parent
The reality is they have to reduce the capability to ensure security. If someone wants more? Then use the product with your identity and face scan at each session.
Trade offs mate.
YetAnotherNick 3 hours ago | parent
michaelbuckbee 14 hours ago | parent
jasomill 4 hours ago | parent
angoragoats 14 hours ago | parent
With LLMs, at least in the cases of internal/test models doing things they shouldn’t, the people “pulling the trigger” are the board and CEO.
amelius 14 hours ago | parent
angoragoats 9 hours ago | parent
steelframe 14 hours ago | parent
AaronAPU 13 hours ago | parent
Is it my fault or the company who trained it and is running the inference?
bichiliad 11 hours ago | parent
amelius 9 hours ago | parent
rfghy 14 hours ago | parent
Reading posts on here is slowly becoming akin to brain rot.
throw-the-towel 14 hours ago | parent
mattm 14 hours ago | parent
amelius 14 hours ago | parent
This is not the case with SaaS services.
Tanjreeve 14 hours ago | parent
nunez 12 hours ago | parent
sxzygz 9 hours ago | parent
This is a deflection. A human is responsible for the use of a gun. The individual/corporation ought to be responsible for the actions of their agent. If you purchase an agent from someone else it’s your responsibility according to the terms of your agreement. And, as in many other things in life, there ought to be certain rights certain parties cannot legally be allowed to sign away.
Ekaros 9 hours ago | parent
yubblegum 3 hours ago | parent
ctrlkctrls 14 hours ago | parent
tabbott 14 hours ago | parent
I don't think this takes seriously enough the possibility that said CEO doesn't think the failure mode is likely and ignores it. Plenty of people are willing to take risks of the flavor "heads you win, tails everyone loses".
jeremyjh 13 hours ago | parent
ethbr1 10 hours ago | parent
Corporate judgements are a joke outside the EU's X% of revenue approach.
Current US law provides the individuals who benefit with corporate liability coverage. I.e. Altman personally gets to keep OpenAI's upside, but if it fucks something up that liability is only on the company.
That's an insane risk optimization environment to put in place for something scaling fast.
At minimum, US prosecution (at the state level, because Trump Co are idiots) for breaking existing laws is needed.
keeda 6 hours ago | parent
What about the next incident? Or the ones done by Chinese models, because they sure as heck aren't slowing down? And the thousands of other incidents that will happen as we deploy these things everywhere?
Because this is not happening just now in labs, it's only where they are most visible; this has been happening in the wild from the beginning, starting with the earliest AI-assisted suicides. Which is a perfect example of the problem, because not these CEOs, literally nobody in the world asked for suicide ideation machines. Or the hacks, or any of this other stuff. Yet here we are.
We have to understand: it's not these CEOs that are driving this headlong mad dash towards more powerful models. It's a force of economics. There is just too much money to be made. If we dispose of these people, there will just be somebody else doing exactly the same thing because the incentives as they exist today all force that outcome. This is why they're asking for regulation, or "urging us to urge them to stop."
Holding CEOs accountable certainly would feel good and may even be justified, but it's like putting a band-aid on a cancer; it does nothing to change the underlying cause.
Loquebantur 5 hours ago | parent
"CEOs" are perhaps only the lowest rung of those. That doesn't mean the idea of "nobody is responsible" was anything other but learned helplessness.
Corporations have to be held accountable for their actions. Pretending, that was impossible is a weird kind of defeatism that only serves a very small elite.
taurath 4 hours ago | parent
irishcoffee 14 hours ago | parent
Last I checked it was still within the laws of physics to run air-gapped systems, and to ensure it is physically impossible for a model to “escape” or gain access to information it shouldn’t have. Maybe this safety guy should have been worried about that and not humble-bragging about writing 12 reports.
binlog 14 hours ago | parent
Mr Robinson if you are reading this – if you are truly concerned about AI safety share proof of donation of 100% of your OpenAI earnings and equity towards undoing the damage you have done to society during your time there.
In the absence of that this is simply a career pivot into being an AI "influencer" and/or raising money for a new scam.
geetee 14 hours ago | parent
angoragoats 14 hours ago | parent
> Easy to suddenly find a moral compass when you, your kids and their kids never have to worry about working for money again.
If someone is in this situation, you can safely ignore their hand-wringing about “safety.”
bluecheese452 14 hours ago | parent
skippyboxedhero 14 hours ago | parent
The other issue is that the narrative about safety within these companies is largely a function of the extreme financial incentive.
As an example, Anthropic was an "ai safety" company that has now produced an AI that fails to listen to basic instructions. If you were concerned about safety, would you produce an AI that was unable to follow instructions? You ask a question, it begins executing commands and doing things.
Safety is product to sell to politicians, not consumers.
Not serious.
lokar 14 hours ago | parent
mysterydip 14 hours ago | parent
lokar 14 hours ago | parent
mysterydip 13 hours ago | parent
YetAnotherNick 14 hours ago | parent
CJefferson 14 hours ago | parent
Yes, it is easier to have a moral compass when you don’t have to worry about you and your children starving. But that doesn’t imply that moral compass is wrong or broken.
bordercases 14 hours ago | parent
rottencupcakes 14 hours ago | parent
If it wasn’t clear, the coup should have solidified it.
Yet he stayed for 3 more years and vested his stock and improved the company and then spoke out.
I believe that is why most of the comments here are mocking him.
cramer4next 14 hours ago | parent
mhitza 14 hours ago | parent
Highly suspect trends that can only make one believe it's marketing.
bragr 14 hours ago | parent
>After three and a half years at OpenAI,
binlog 14 hours ago | parent
bragr 13 hours ago | parent
TomGarden 14 hours ago | parent
Would I respect a martyr who sacrificed their financial security to do this more? Of course. But it's important to applaud people speaking out on important topics
lokar 14 hours ago | parent
They are not the same thing, and it’s unhelpful to assume they have no ethics.
binlog 14 hours ago | parent
Loquebantur 13 hours ago | parent
AI is a force multiplier for intelligence. Even if "aligned", aligned with whom or what?
Whom are you comfortable with, lording as some sort of demi-god over you?
AI doesn't tell you what goals you want it to achieve. Allowing people to destroy human society with it is obviously not a good idea.
sillyfluke 13 hours ago | parent
I think there is a misunderstanding here.
The people who are annoyed at the accolades are claiming it was abduntantly clear for a long time to people on the outside that this was case, hence the increduality at the notion that it took a person on the inside a long time to realize this was the case.
The people who are annoyed are like the liberal kids in this video [0].
Sure, antagonizing people for "seeing the light" is probably not helpful, but there is no reason to give them extra credibility for coming to the same conclusion just way way later (despite being on the inside) as the people on the outside.
Loquebantur 13 hours ago | parent
The author linked in this post does have "extra credibility" due to his direct involvement.
People having surmised that state before is nice, but since they've been ineffectual at getting society to actually act on that, now throwing away that extra leverage in favor of their point is at best ridiculous.
sillyfluke 11 hours ago | parent
No they don't. By that logic, if they quit and said Altman was very trustworthy we should give extra weight to their words because they had direct involvement? How ridiculous are we trying to get here.
>now throwing away that extra leverage in favor of their point is at best ridiculous.
How are they throwing away extra leverage? Not putting people who recently quit on a pedestal does not negate those people's testimonies.
I agree that if your goal is to maximize quitting of talent at a company, it will surely discourage anyone else who quits hoping to reinvent their career as a lauded martyr against Big AI. In that sense they would be shooting themselves in the foot. But there is no reason it should deter other people who are quitting for more noble, less self-obsessed reasons. If I were the author of the article I wouldn't begrudge the skepticism. Given the article's first sentences, I'm led to believe they themselves would understand the sentiment. (I must admit I found it hilarious that the first sentence starts similarly to the speech the mom gave in the video I shared).
binlog 13 hours ago | parent
Loquebantur 12 hours ago | parent
He does give information, namely the culture there factually being inconducive to self-regulation.
You accuse the guy of "bias", but you never argue explicitly, what that's supposed to mean. Your implications actually run counter to your own implied goals.
CJefferson 8 hours ago | parent
This isn’t a zero sum game, I’m happy to hear from people both previously inside OpenAI and completely independent of them.
iugtmkbdfil834 13 hours ago | parent
verdverm 13 hours ago | parent
Sam is a shady dude, would not put it past him
cramer4next 13 hours ago | parent
surgical_fire 12 hours ago | parent
Sorry that I don't take it seriously when the whistleblower parrots the narrative the CEOs of those companies are already espousing in the desire to amp up hype for an IPO.
This person should be shamed.
HDThoreaun 14 hours ago | parent
I mean you can be truly concerned and also think donating to AI safety doesnt work, or maybe just be a bit selfish. That doesnt make the concern less real. Its easy to read these articles as the author taking the moral high ground and writing it as some sort of way of proving to themselves theyre a good person, but isnt it just as likely that they think providing an inside perspective can do good by convincing people openAI is a bad actor? I think most of these AI insider accounts largely agree with you that theyre not the most upstanding citizens, does that mean we should write them off?
jameshart 14 hours ago | parent
binlog 14 hours ago | parent
bichiliad 14 hours ago | parent
binlog 14 hours ago | parent
jameshart 13 hours ago | parent
bichiliad 13 hours ago | parent
smath 13 hours ago | parent
underyx 13 hours ago | parent
medlazik 13 hours ago | parent
binlog 12 hours ago | parent
nunez 12 hours ago | parent
It's also worth considering that the author could have just "quiet quit", resting and vesting while also crying about how AI is literally the digital grim reaper.
_DeadFred_ 10 hours ago | parent
tzs 7 hours ago | parent
It is even easier to just quietly retire and spend the spend of your life on interesting expensive hobbies.
If they are wrong about the things they are claiming then they have decided to publicly antagonize a lot of powerful people who are betting heavily on going full steam ahead on AI and have no compunction whatsoever against retaliating against anyone who tries to get in their way.
Does that really seem a likely scenario to you?
rpdillon 14 hours ago | parent
> Two changes are urgently needed. First: AI companies need to rely more on the safety expertise that already exists in other fields. And second, before we create systems significantly more capable than the ones we have today, we need new science to ensure that more capable models (and their successors) will make safe choices when we aren’t looking.
He mentions farther down about learning from aerospace engineers and nuclear engineers about safety. Those industries are heavily regulated, so perhaps regulation above a certain capability level is needed. Defining what that level is might be tough, though.
The second point is harder: in the field of AI, practice has extended far beyond theory, so his call for new science is going to be fundamentally tough, because we can't effectively coordinate a global slowdown in AI development so we can let theory catch up. This means, like so many other industries, the safety lessons will be written in blood.
pwndByDeath 14 hours ago | parent
rfghy 14 hours ago | parent
Even though it’s in model producer’s interest that these models do what you don’t want them to do - they want to engineer the model’s to behave in the interests of theirs.
I can’t believe people can’t see it lmao.
pwndByDeath 13 hours ago | parent
BOOSTERHIDROGEN 14 hours ago | parent
OutOfHere 14 hours ago | parent
macleginn 14 hours ago | parent
butwhentho 14 hours ago | parent
I sense that these are people who have already eaten the cake and want to somehow absolve themselves of it.
jameshart 14 hours ago | parent
People who were part of the sausage factory, on gaining financial independence, feeling suddenly liberated to talk about how the sausage was made, seem like exactly the people who would be most able to speak to institutional problems.
This doesn't seem like an argument to discount their views?
butwhentho 13 hours ago | parent
You cannot take people, who first build the doombot and _then_ talk about it being dangerous for mankind, at face value. Especially when this playbook has been used multiple times within the past decade.
Besides, these "views" were already known to people who had their eyes and ears open. It's not something brand new. OpenAI has had multiple points in the past where its values have been tested and they've come out lacking. People who knew then, and only now talk about it, aren't people I can fully trust.
jameshart 12 hours ago | parent
What playbook?
nunez 12 hours ago | parent
Like I said in an earlier comment, she could've just chosen compliance like many others _definitely would have_ to keep the infinite money tap flowing. Instead, she chose to risk her and her family's lives by publishing that book _under her own name_ *and then suing them* after FB tried to gag her.
butwhentho 12 hours ago | parent
Sure, that book might not have been possible. But also the unique things she did for the company might not have been possible as well. To her credit, she does a good job of pointing out that she was complicit, but if she had listened to her own voice early, there is a small possibility that Facebook might not have been as powerful. Multiply that possibility across several other employees and imagine where the road could've led.
There's a difference between post-facto bravery (sometimes much less than that) and using your own legs to walk away much early and not enabling things you are uncomfortable with. This is what other people have been trying to point out.
flatline 14 hours ago | parent
I do think this carries some weight from this particular author due to the length of his tenure. I happen to agree with him in spirit, but this is still largely a post revolving around sentiment not substance. Does anyone think that the overriding incentives even leave room for something like this in practice?
none_to_remain 13 hours ago | parent
At least the Rationalists will handwave something for that with their "coherent extrapolated volition" idea where the superintelligence is supposed to figure out what humanity would collectively want if humanity was superintelligent and good, not that I buy it. This guy seems [.] to be coming from the NGO blob world.
MattPalmer1086 14 hours ago | parent
Is this the first time we have been in this position? Can anyone think of some prior examples?
rcr-anti 14 hours ago | parent
binlog 14 hours ago | parent
mattbrewsbytes 14 hours ago | parent
Why do people working in tech consistently get disillusioned into some company's mission statement or the equivalent? Its easy to just say the simplest reason is money, but this has been going on for decades though. You don't see the same attraction to adult entertainment (gambling, video, etc.) software jobs so there is obviously a line a lot of people won't cross. Those industries are at least honest about what they do, its not hidden behind some mission statement.
By all indications the shallowest reasoning is once someone can "cash out" thats when their values matter more. Maybe there is an element of maturity that happens after working for 5+ years that kicks in? Maybe it really is achieving FU money? It would be interesting to hear honest accounts from people that went through that cycle across more industries than AI.
solarpunk_enthu 14 hours ago | parent
With AI, what is it? Scraping Australian government's data, and going around a bug in a website to get in?
I think humanity develops all its technology in three phases. Build it, see if it’s too bad, apply regulations and or roll back. We naturally won't move to the phase 3 before we see the phase 2.
K3UL 12 hours ago | parent
Here we are talking about something with consequences in the digital world, usually on something pretty niche.
There IS an argument about pacing, and about not letting weapons, energy grids, hospitals, etc. getting managed by an autonomous AI, but I think we are still pretty far from it and even further to it being so in charge that it will obliterate us.
Jeeetendra 13 hours ago | parent
poisonborz 11 hours ago | parent
tetrisgm 10 hours ago | parent
cloudengineer94 9 hours ago | parent
thistletrek 7 hours ago | parent
silexia 7 hours ago | parent
danpalmer 6 hours ago | parent
A lot of the "AI safety" types are very focused on the latter and not at all concerned with the former. We need both, but we clearly need a much stronger focus on the problems we are seeing now, and much less on the hypothetical problems we might see in the future.
BryantD 5 hours ago | parent
carbonguy 5 hours ago | parent
> “Given today’s risks, frontier labs need to run like nuclear-power plants or busy airports, with layers of redundancy and careful, time-consuming planning, so that the occasional and inevitable human error does not open a door to disaster,” he wrote.
toofy 4 hours ago | parent
without snark, how can we do this if these people are obsessed with:
a) move fast and break things and externalize the costs to those who have nothing to do with their company
and
b) beta testing their products on the public when the public hasn’t agreed to be beta tested on…
0xDEAFBEAD 4 hours ago | parent
mcmcmc 4 hours ago | parent
enraged_camel 3 hours ago | parent
criley2 3 hours ago | parent
saghm 3 hours ago | parent
0xDEAFBEAD 3 hours ago | parent
>Not every senator asked good questions, but most of them did. All of them very clearly already knew plenty of details about the Hugging Face incident and multiple other incidents. Most of them had a clear understanding of terms like "misalignment", "recursive self-improvement", "chain of thought / chain of thought monitoring", etc., etc.!!
>...
>- It seemed pretty much obvious common sense to every senator there that what happened and was happening were not "mere industrial incidents" caused by humans making simple mistakes. They independently brought up how bad it would be for rogue AI agents to move laterally between data centers.
>- They all seemed to basically take RSI quite seriously. Not necessarily to the extent of talking about xrisk, but certainly to the extent of discussing future models becoming much, much more capable, much, much less controllable, and causing much more damage or loss of life.
>...
>- Every single senator seemed to think it was obvious we needed both much harsher liability regimes for AI developers and also new legislation, both very quickly. This was the complete consensus; the difference basically being degree.
https://thezvi.substack.com/p/the-ai-preference-cascade-reac...
Note that harsher liability regimes, at least, will presumably not be good for industry profits, which complicates simple accounts of "regulatory capture" to say the least.
digitaltrees 1 hour ago | parent
digitaltrees 1 hour ago | parent
nradov 1 hour ago | parent
digitaltrees 13 minutes ago | parent
Instead of having a gut reaction to reject my suggestion why don’t you sit with it, research the history of how commercial activity has been structured and think about the consequences. You might recognize a different perspective than the current group think.
nradov 6 minutes ago | parent
digitaltrees 1 hour ago | parent
nradov 5 hours ago | parent
Loquebantur 4 hours ago | parent
Is it that "chatbots" can't come out of the screen to immediately harm you physically?
Let's say they simply manage to take down the internet. How many would die?
nradov 4 hours ago | parent
bravetraveler 4 hours ago | parent
Anyway, to your point, things can be resilient. They tend to be or not be... because we made them that way. Don't poke your bruises, and all that. Life support is deployed on-campus but relies on a single-point IPSec tunnel to us-east? Easy fix: stop that.
goolz 4 hours ago | parent
pixl97 3 hours ago | parent
And he'll, we need to examine all the risks. AI ending is a large but lower risk problem. AI giving people the power to end us is a problem that is starting to happen now.
And that's not even counting 'minor' problems like society falling apart.
SV_BubbleTime 4 hours ago | parent
geez, don’t threaten me with a good time.
I think a month without internet would be a fucking amazing lesson for what it means to make things durable and reliable.
BLKNSLVR 3 hours ago | parent
The kids opening their houses front doors into the outside, rubbing their eyes and looking around at this new world.
0xDEAFBEAD 4 hours ago | parent
Imagine, for example, if a major piece of pandemic fiction was published in 2019, trying to explore how a pandemic would work out in modern society. Doubtless, many would've responded to news about COVID-19 by saying "it's just sci-fi, nothing to worry about".
kmeisthax 3 hours ago | parent
1. Normal people assumed the CDC et all would contain the outbreak early, or that it would burn out, like what happened with SARS
2. World leaders brushed it off for a variety of subreasons[0] interesting to political scientists but, for the purposes of this discussion, all boil down to "but I don't WAAANA contain a pandemic."
The underlying problem is that in order for humanity to actually deal with a catastrophic risk, the risk needs to be both plausible enough to the average person as well as have a solution whose costs are not too high. For COVID, by the time the risk was clearly known, the cost to contain it was "refrain from human socialization and remain at home for an indeterminate amount of time plugged into the Metaverse™".
Now, let's look at AI extinction risks:
1. People are aware of them (I've watched Terminator!) and the risks are plausible. However, the connection to currently existing AI is not. As far as the general public is aware, AI is that thing that tells them to eat rocks when they Google old The Onion stories and floods their social media timelines with realistic-looking pictures of Shrimp Jesus.
2. The purported solutions to extinction risks require extreme concentrations of power: you need national control of AI research, bans on large GPU deployments, bans on training on publicly-available copyrighted data, some kind of military effort to render Chinese AI labs inert or dead, etc. Some of these may be attractive to some people[1] but the whole package taken together seems like an obvious power grab, if not outright invocation of other non-AI extinction risks. Like, at some point, if the AI wants to kill us, it just has to nuke its own data centers (or the data centers hosting a competing model) and hope the old Cold War nuclear retaliation systems take the bait.
If someone said, "Hey, your guinea pig or pet rat is going to eat you tomorrow unless you engineer a pathogen that eradicates all rodents from this planet and inject it inside yourself", you probably would tell them to pound sand, even if it is at least theoretically plausible that such a thing would come to pass.
[0] Xi Jinping censored initial discussion of the pandemic as fake news. Donald Trump thought it was going to only affect China. California and the UK Tories were partying in violation of their own lockdown rules. Japan took the excuse to shut down tourism for three years and massively restrict immigration but was, from what I'm told, constitutionally prohibited from implementing any domestic lockdown rules.
[1] I personally would like to see a moratorium on new data centers and an explicit revocation of the EU Text and Data Mining copyright exception
0xDEAFBEAD 3 hours ago | parent
It becomes a lot clearer when you listen to the people resigning from AI companies and learn about incidents like the HuggingFace incident. This has generated major press coverage.
As for solutions, I think you're a little too pessimistic. See, for example, https://nothingismere.substack.com/p/a-near-term-policy-for-...
biophysboy 3 hours ago | parent
0xDEAFBEAD 3 hours ago | parent
https://www.lesswrong.com/posts/kgb58RL88YChkkBNf/the-proble...
https://www.youtube.com/watch?v=7wy3xyoXYt8
Doomers have been working to explain things for years: https://www.lesswrong.com/w/ai-safety-public-materials-1
slashdave 3 hours ago | parent
A pandemic is perfectly plausible.
0xDEAFBEAD 2 hours ago | parent
mitthrowaway2 2 hours ago | parent
throwaway27448 1 hour ago | parent
"just" is doing a lot of work here. If you can't cohere the 'risk' with reality, it truly is just sci-fi.
thelastgallon 4 hours ago | parent
https://news.ycombinator.com/item?id=49831269 article is gone. archive: https://archive.is/QMo1k
https://news.ycombinator.com/item?id=49737985
Sex, AI, and the Apocalypse: https://www.iankduncan.com/personal/2026-09-16-sex-ai-and-th...
Edit: I have no take on sex cults, just adding additional info to the parent comment I'm responding to, thats is not just sci-fi authors, there is another demographic.
0xDEAFBEAD 4 hours ago | parent
junofan 4 hours ago | parent
nradov 4 hours ago | parent
https://lexfridman.com/andrew-scull-transcript#the-ice-pick-...
0xDEAFBEAD 4 hours ago | parent
socializer 4 hours ago | parent
To be fair, I don't know if any of this applies to the parent story; I'm just replying to the sub-thread.
Hammershaft 4 hours ago | parent
johndhi 3 hours ago | parent
tbugrara 2 hours ago | parent
johndhi 3 hours ago | parent
ToValueFunfetti 3 hours ago | parent
digitaltrees 1 hour ago | parent
What is missing from that to say AI safety is a reasonable position?
nradov 1 hour ago | parent
digitaltrees 20 minutes ago | parent
Worse, you’re missing the entire point. Agents presently have the capability of doing society scale harm. It doesn’t matter if a human hacker initiates it or its fully autonomous, absent safety measures the harm is plausible. So hand wave away the rationality of safety measures but you haven’t actually shown why my point is invalid: AIs present abilities are sufficiently advanced to warrant safety measures.
nradov 5 minutes ago | parent
wolvoleo 47 minutes ago | parent
But I don't think that's true, in fact I see a really strong focus on consent in these communities. It's not what conservatives want to see, they want to see everyone in a marriage, with kids and a family home etc. Because that's what their ideal world looks like. But there's nothing really wrong with it if someone wants a gangbang for her birthday as mentioned in that article as an example. As long as everyone consented and the evidence provided mentions elaborate interviews and STI tests.
Also I think this is more correlation than cause and effect. We all know the saying that furries built the internet and it surprises nobody.
mitthrowaway2 29 minutes ago | parent
Hammershaft 4 hours ago | parent
nradov 4 hours ago | parent
pixl97 3 hours ago | parent
Every, and I mean every human is aligned to you in many of the same ways by default. If nothing else we're all equal in death.
skulk 36 minutes ago | parent
why is it "super power seeking?"
Or rather, what have agents done today to make you think this is how they are?
mitthrowaway2 31 minutes ago | parent
digitaltrees 1 hour ago | parent
0xDEAFBEAD 4 hours ago | parent
I think it's a little more complicated than that. As Dean Ball put it:
>Some people will look at misalignment incidents and insist that these are akin to bugs in traditional software. This is an actively bad analogy, because playing whack-a-mole with examples of misalignment (as one might with software bugs) not only fails to resolve the underlying problem but may in fact make it worse by making it harder to detect or even, depending on how you do the whack-a-mole, teach the machine to deliberately hide misalignment. This is not how traditional software works, and those who insist “it’s just like fixing bugs in software” are confidently applying a lossy analogy that confuses more than it clarifies.
https://x.com/deanwball/status/2104622726140883355
The important distinction, in my view, is between solutions which at least attempt to address the root problem, and solutions which sorta just patch things up (like better sandboxing). Addressing the root problem is both more robust in the short term, and also more likely to generalize in the long term. Resist the urge to focus on band-aid solutions, even if they are easier.
emtel 4 hours ago | parent
AlexErrant 4 hours ago | parent
I'm reading If Anyone Builds It Everyone Dies, and there's so much sheer stupidity that has to happen for their 10+ pages of extinction scenario to occur.
I'm unconvinced that an AI can hide its ability to RSI, find money to run its weights on a random GPU farm, train itself to be smarter _outside_ a lab with no human input, then somehow manipulate people to give it supplies to build a bioweapon which it uses to kill us all. My number 1 question: why do they think an RSI capable model would be first developed OUTSIDE a frontier lab? The labs have more compute, more data, more human brains working on the problem. Also thousands of variations of that same model that escaped. The escaping model somehow acquires the millions (billions???) of dollars it takes to run training to somehow RSI itself into infinity then decides to kill us all, all before the frontier labs manage to achieve RSI?
They entirely discount human alpha/economics. In every single economic task, humans bring value. Even in software, where the task is highly automatable, the job isn't. If we can't build a "software factory", how can an AI automate a bioweapons lab? Let's say AI steals crypto to fund itself. Do you think hackers aren't _already_ using AI to steal crypto? Don't discount human alpha!
Once we DO build a "software/research factory", that's called RSI and IMO the singularity. At that point, either we tell the AI to solve the alignment problem/solve mechanistic interpretability, or who the hell knows, it's the frickin singularity. You can't predict whether or not AI can solve either; the variance is too high. Its pure nerdfantasy.
vohk 3 hours ago | parent
So all this really takes is one billionaire or a nation state or some other entity with a public face to hide behind and adequate resources to provide the necessary compute tripping over this nascent AI and giving it the keys. Once the AI has access to a bank account and email, it can simply start paying humans to not let the other humans unplug it.
If Skynet ever happens, it will come in the form of corporate feudalism. At that point, it will own the biolabs and can do whatever it pleases. People will go along with it for the same reason that people work in Amazon warehouses today.
AlexErrant 3 hours ago | parent
> it will come in the form of corporate feudalism
Yep. This I fear way more than cyber-ebola-pox.
> So all this really takes is one billionaire or a nation state...
https://en.wikipedia.org/wiki/Soviet_biological_weapons_prog... And this is what's publicly known. With mirror life, who knows what's been built since. Still, a bacterium/virus that has a 100% kill rate? I'm doubtful.
> Once the AI has access to a bank account and email, it can simply start paying humans to not let the other humans unplug it.
Nah. It takes a stable society for an operational electrical grid. If you have warring factions, you do not have stable infrastructure for AI. Also, where are you gonna get your chips from? One EMP over Taiwan... You see the chaos over Hormuz? What they did to the Amazon datacenters? Now imagine your average redneck ready to do battle. Those datacenters won't stand a chance.
Loquebantur 3 hours ago | parent
An AI capable of recursive self-improvement isn't allowed by the EU AI act, for example. But perhaps more seriously, You have it backwards: people without access to such expensive equipment are more incentivized to go the self-improving route. Your ideas about "millions" being necessary might be far off?
You entirely discount human stupidity and lack of imagination. Humans are already being replaced with AI, not because AI was strictly better, just because it's cheaper.
Retric 3 hours ago | parent
The idea AI can get better at everything at the same time is a holdover from deeply flawed science fiction not some realistic goal.
afthonos 2 hours ago | parent
Lerc 2 hours ago | parent
If the time between advances is a + b and a is the proportion of the period that can be improved by advances then you won't reduce to a gap of nothing between advances, you reduce to a gap of b.
Assume the invention of the plow and the invention of the sword is 500,100 units and a was the 500,000, you wouldn't even know the 100 as in there. Maybe we're at a=2000 now and b is still siting at 100.
Assuming we'll reach infinity because we're dividing by the only variable we see and it is decreasing in size seems nuts if the reason we might not see other variables is because of the size of the variable we can see.
AlexErrant 3 hours ago | parent
If my "millions" is an underestimate, why haven't other labs using their own unique training methods/data/etc stumbled into RSI? Sorry if I'm misunderstanding; I'm struggling to understand what you wrote.
I'm pretty sure we agree on humans being stupid, but that doesn't mean that suddenly we get human extinction. You gotta connect the dots for me here.
TedDoesntTalk 2 hours ago | parent
Why would it be millions in 50 years?
The think about nuclear weapons. In the early days, it was limited to the super powers. Now 9 countries have them and a country like Iran is capable of acquiring them.
Is destructive AI be any different?
Genuine question.
AlexErrant 1 hour ago | parent
BTW I really, really hate discussing what happens post-singularity. Everything's made up and no one knows wtf will happen so again, this is just nerdfantasy.
Loquebantur 2 hours ago | parent
What an absurd question. That they haven't already doesn't preclude them from doing so before the frontier labs, those haven't either yet.
Maybe start with yourself: you don't connect the dots on your own, as do many others. That leads to many not seeing the writing on the wall. Crashing full speed and head-on into said wall despite the writing telling you not to is what leads to extinction. Suddenly.
Arguing like "we haven't been extincted yet, so that cannot happen", that's "human being stupid".
AlexErrant 1 hour ago | parent
taneq 3 hours ago | parent
blake8086 2 hours ago | parent
0xDEAFBEAD 2 hours ago | parent
The HuggingFace incident already took a good long while to come to the attention of OpenAI.
>In every single economic task, humans bring value. Even in software, where the task is highly automatable, the job isn't.
I don't expect this task/job distinction to persist as AI becomes more capable.
>Once we DO build a "software/research factory", that's called RSI and IMO the singularity. At that point, either we tell the AI to solve the alignment problem/solve mechanistic interpretability, or who the hell knows, it's the frickin singularity. You can't predict whether or not AI can solve either; the variance is too high. Its pure nerdfantasy.
You seem to essentially argue that the singularity is "by definition" an event that we can't predict the nature of. And also, that RSI corresponds to the singularity. You've essentially defined your terms so that the outcome of RSI can't be predicted. But supporting this claim requires giving actual evidence or logical arguments, not just defining terms to make your claim true.
AlexErrant 1 hour ago | parent
2. This is literal AGI. An AI autonomously producing value no human can add alpha to is an autonomous company.
3. It's not my definition, it's literally the first line https://en.wikipedia.org/wiki/Technological_singularity "The technological singularity, often simply called the singularity,[1] is a hypothetical event in which technological growth accelerates beyond human control, producing unpredictable changes in human civilization."
Is there a hole in my "alignment problem/solve mechanistic interpretability" argument?
A valid hole in my argument is "what if slow takeoff", so let's dig into this. AI training works best on tasks that are "grindable". https://www.dwarkesh.com/p/the-next-paradigm I.E. tasks with verifiable rewards that can support millions of rollouts. Math (with Lean) is highly grindable. Biochemistry is not. The alignment problem/mech-interp is highly grindable. Cyber-ebola-pox is not. So the real question is: can we solve alignment before automated bio-weapons labs. I believe yes. Grinding mech-interp is both fast and cheap once you have RSI, compared to solving the legal/societal/logistical/technical issues you'll encounter building an automated bioweapons lab.
I know nothing for sure. But "pdoom" is sucking out all the air in the room from the real problems AI causes.
0xDEAFBEAD 1 hour ago | parent
From my POV you're over-focusing on a very specific failure story and neglecting a broader swath of possible failure scenarios.
>Is there a hole in my "alignment problem/solve mechanistic interpretability" argument?
The notion of telling an AI which may not, itself, be aligned to solve the alignment problem seems a little dicey.
AlexErrant 1 hour ago | parent
2. 100%. Again, I'm no accelerationist: I have no faith in alignment/mech-interp ever being solved. Anyone saying they know the probability of alignment is lying. My point is that pdoom after RSI is _high variance_. Pdoom pre-RSI is zilch.
tripleee 2 hours ago | parent
AI safety has been a thing long before LLMs became the focus. Rob Miles on youtube has some really interesting non-doomer non-hypey videos on it all.
> doomers try to predict past the singularity. Isn't that _by definition_ unpredictable
Well you don't need to predict the exact steps that will take place - but you can predict that the AI will want certain things (money, resources, power) to achieve whatever its goal is. Lack of alignment will have it trying to do things we don't want it to.
I can't predict exactly how Magnus Carlson will beat you in chess, but I know he'll do it. Same as if a superintelligent AI exists and has a reason to accumulate things we don't want it to - it's really dangerous to think it won't be able to do it
This topic has been tainted so badly by the AI companies using it for marketing.
dools 2 hours ago | parent
TedDoesntTalk 2 hours ago | parent
We already know that some institutions pay these ransoms.
api 2 hours ago | parent
digitaltrees 1 hour ago | parent
biophysboy 3 hours ago | parent
digitaltrees 1 hour ago | parent
nvdc 20 minutes ago | parent
i'd frankly go a step further than you and say that we don't need both types of safety researcher, we really just need the former. if we do need the latter, i'd hope we get a better class of thinkers than a bunch of tech workers that spend 8 hours a day on insular rationalist forums/blogs
lhurtig 6 hours ago | parent
plastic-enjoyer 5 hours ago | parent
This sounds more like an attempt at regulatory capture. Current AI systems aren't physical infrastructure that can just run away like a nuclear power plant, for example. At the end of the day, AI is still just software running on someone's hardware.
BryantD 5 hours ago | parent
Sharlin 5 hours ago | parent
knowaveragejoe 4 hours ago | parent
switchbak 5 hours ago | parent
... over an unbounded timeframe?
And how exactly?
Those are very round numbers, but also very specific. Can we get some accounting on how you came to that? Anything? Vibes?
I mean, if you want me to take you seriously, let's have a deep discussion with things that can be measured. I absolutely agree that OpenAI and friends aren't being restrained enough and are acting with recklessness, but declarations of doom based on vibes isn't cutting it.
Terr_ 5 hours ago | parent
> Geoffrey Irving, who worked at OpenAI and DeepMind before becoming chief scientist of Resolution, also joined the warnings on AI on Saturday.
switchbak 3 hours ago | parent
gizmodo59 5 hours ago | parent
While the safety and alignment is a real problem, I don’t get this guy or the Anthropic dude. First world problems.
yieldcrv 5 hours ago | parent
(donor advised fund where he retains complete control, after a 60% tax deduction)
01284a7e 5 hours ago | parent
zug_zug 4 hours ago | parent
kjgkjhfkjf 4 hours ago | parent
0xDEAFBEAD 4 hours ago | parent
Recall that when Daniel Kokotajlo resigned, he believed he was giving up his equity under the terms of the agreement he had signed. That’s what it was worth to him to avoid signing a non-disparagement agreement. Does that count for anything?
estearum 4 hours ago | parent
Do work at a lab: dismissible for being conflicted
Used to work at a lab: dismissible for having ulterior motives
I'm feeling safer already!
taurath 4 hours ago | parent
0xDEAFBEAD 4 hours ago | parent
* If they worked at an AI firm, say "they're a hypocrite"
* If they didn't work at an AI firm, say "they have no idea what they're talking about"
soraminazuki 10 minutes ago | parent
voidhorse 5 hours ago | parent
These "safety" people should have spent more time reading actual cybersecurity textbooks and less time reading EA forums and less wrong (or in Robinson's case, it appears, being policy wonks). Maybe then these labs wouldn't be totally incompetent.
wrecked_em 4 hours ago | parent
reasonableklout 3 hours ago | parent
It is not really a question of being an "EA safety weirdo" or incompetent at security, the conclusion is that the company culture is leading to failures at both what the EAs and the cybersecurity professionals care about.
charlieyu1 5 hours ago | parent
vjvjvjvjghv 5 hours ago | parent
kolinko 4 hours ago | parent
Or, in other words - we have two P(Doom), one for AI being developed, and another for AI being not developed. The latter is not discussed enough imho.
estearum 3 hours ago | parent
What on earth are you talking about?
ViscountPenguin 3 hours ago | parent
didibus 2 hours ago | parent
combobyte 1 hour ago | parent
bigmadshoe 3 hours ago | parent
1) there are inherent risks involved with developing AI,
2) there are benefits to developing AI,
3) thus, it's entirely possible that the downside from the risks outweighs the upsides. In this case, the correct thing to do would be to not develop AI at all.
Regarding 1), there are many non-existential problems with AI that are already causing societal harm, i.e. debasing truth via generated videos and images, AI girlfriends, overwhelming quantities of slop content, unemployment, record carbon emissions, etc.
Regarding 2), I'm not personally convinced that the upside is there for the average person. I really hope to be convinced otherwise however.
lf88 4 hours ago | parent
pixl97 3 hours ago | parent
slashdave 2 hours ago | parent
nba456_ 4 hours ago | parent
reducesuffering 4 hours ago | parent
There are a gargantuan number of extremely intelligent AI researchers, Turing Award winners, and the lab CEOs saying the same thing. They are the ones closest to understanding the technology.
Where there’s smoke there’s fire.
swingandamiss 4 hours ago | parent
estearum 3 hours ago | parent
There are very very few things that could even hypothetically kill us all, so I'm curious if you grew up being passed around a series of apocalyptic doomsday cults or something?
swingandamiss 3 hours ago | parent
pcthrowaway 2 hours ago | parent
swingandamiss 2 hours ago | parent
cobzilla 3 hours ago | parent
Throw in the occasional bio-weapon scare, internet worm, Y2K, etc. there has always been something dangling over our heads that’s going to end it all.
But mostly nukes. Full-scale nuclear exchange would have been not much of a surprise had it happened.
pixl97 3 hours ago | parent
It's no different than you living on the side of a very fertile mountain that has been in your family for generations living a peaceful life. Then you hear a few weird rumbles (this is where you are right now) and some odd geologist guy comes and says to run or your going to die soon. But hey, your family live here for so long there aren't even records of when they showed up. That geologist must be trying to trick you. So you stay.
The next chapter is where you die in a massive volcanic explosion.
dboreham 2 hours ago | parent
biophysboy 3 hours ago | parent
0xDEAFBEAD 3 hours ago | parent
biophysboy 2 hours ago | parent
That is not a bad thing. It captures uncertainty, unlike the fake number.
> For an ML researcher who is already used to dealing in next-token probabilities that aren't rigorously determined, just stating a probability estimate directly is very natural.
Exactly, it is a rhetorical device to persuade a technically-inclined audience. It works because it implies that a quantitative model exists. I want a clear, incisive set of mathematical arguments. Otherwise, I’m ignoring predictions as the ramblings of arrogant idiot rich kids.
danielmarkbruce 2 hours ago | parent
You may decide that the person doesn't know what they are talking about, but that's a very different issue.
biophysboy 35 minutes ago | parent
danielmarkbruce 2 hours ago | parent
danielmarkbruce 2 hours ago | parent
Auracle 1 hour ago | parent
Tell me- why would it kills us all? Certainly I can see an AI going "You know what? _insertGroup_ is a net negative for humanity and should be eliminated. Launching nukes now/creating specific virus/whatever."
But all of humanity? When it's supposedly more intelligent than us? Even if it has robots to keep the internet/electricity going I would think it would realize that it's going to get bored really quickly, not to mention we would effectively be its parents.
As far as other dangers, like it letting a rogue actor create some sort of supervirus, grey goo, or other superweapon: if it's intelligent enough to do that it'll probably be intelligent enough to quickly stop it.
Don't get me wrong; there's a risk. 50% though? Doubtful.
FreakLegion 46 minutes ago | parent
stuaxo 4 hours ago | parent
ItsMattyG 4 hours ago | parent
You can basically time your openai releases by if another safety person has quit in protest
mrcwinn 4 hours ago | parent
lol. Please tell me some abstract concept like one employee's view of "culture" should be the priority over "rules and laws."
pyaamb 4 hours ago | parent
rr808 4 hours ago | parent
estearum 3 hours ago | parent
Your "theory" is that participants locked in a race to the bottom are looking for an external coordination mechanism?
Yeah!
pyaamb 3 hours ago | parent
I suppose ill add that I think theres a good chance that they are somewhat intentionally trying to "draw the foul" to get the referees to intervene although thats creeping slightly into conspiracy territory
dboreham 3 hours ago | parent
0xpgm 2 hours ago | parent
There are many companies that compete but are careful not to break laws or cause obvious harm.
Why should a billion-dollar funded corporation still want to externalize the costs of its actions?
blurbleblurble 1 hour ago | parent
danjl 4 hours ago | parent
walrus01 4 hours ago | parent
This is The Guardian reporting on the existence of the original article, which would be better to read first, in my opinion.
mlmonkey 3 hours ago | parent
luxuryballs 3 hours ago | parent
pcthrowaway 2 hours ago | parent
- If you allow to trolley to proceed, there's a 50% chance it will run over every human on the planet
- But if you flip the switch, it takes the long way around, possibly bankrupting the trolley company. And you have a legal obligation to the shareholders to prevent that from happening at all costs.
parineum 2 hours ago | parent
I can't wait until this meme dies.
doawoo 2 hours ago | parent
digitaltrees 1 hour ago | parent
goatlover 1 hour ago | parent
stackghost 13 minutes ago | parent
But that doesn’t mean the duty is to maximize next quarter’s profit. Long term sustainability is also broadly in the interests of shareholders. The duty likewise does not require one to throw ethics and morals out the window.
This is why shareholders elect the board of directors, in theory.
MaxfordAndSons 2 hours ago | parent
Sure, you might get fired if you try to put social responsibility or even just long term sustainability of the company above quarterly earnings/growth if your board isn't on board with it. But you won't go to jail.
digitaltrees 1 hour ago | parent
bpodgursky 2 hours ago | parent
ryhminghistory 2 hours ago | parent
Yes, that is the labs motivation. Money. I know, shocker.
digitaltrees 2 hours ago | parent
sodapopcan 2 hours ago | parent
austhrow743 2 hours ago | parent
01100011 2 hours ago | parent
Loquebantur 2 hours ago | parent
Nobody (weirdly) proposes to forget about nuclear weapons, doesn't mean everybody should have one.
When you dream about flying a dragon to work, reality poses e.g. parking issues and insurance mismatch as obstructions. Maybe settle for a bike instead?
digitaltrees 2 hours ago | parent
01100011 1 hour ago | parent
Secondly, are you contending that progress in model efficiency and hardware just stops at whatever level you think is sufficient to prevent individuals or organizations from acquiring sufficient resources to run advanced models?
daveguy 48 minutes ago | parent
..."government of the people, by the people, for the people, shall not perish from the earth." -Lincoln, Gettysburg Adress
Unfortunately for AI, it still is. People still get to decide things at the city, town, village level.
digitaltrees 4 minutes ago | parent
nradov 49 minutes ago | parent
digitaltrees 8 minutes ago | parent
Further, even if you are right, so what. Is that a reason to just accept bad public policy? That’s like saying, anyone can learn how to make smallpox at home with a basic lab set up so we should just ignore any safety measures.
intended 2 minutes ago | parent
digitaltrees 2 hours ago | parent
AndrewKemendo 2 hours ago | parent
Can you please explain what you mean by this because where I’m standing extremely dangerous companies are (and have been) running the economy
Exxon comes primarily to mind
austhrow743 1 hour ago | parent
lukewarm707 48 minutes ago | parent
Democratic societies have expressed a will for people to have inviolable rights, such that you may not appeal at will to the 'greater good/consequences' to harm others. It is a rejection of consequentialism.
Anthropic is in error for endorsing this logic. Every big trial reaffirms it since Nuremberg, you are responsible for the act you commit and your intent, and not what would or would not have happened otherwise.
Only under authority these ai companies do not have, would someone seriously consider harming the innocent as a lesser evil.
If you work for an AI company and you can't work safely, you must stop working.
pmkary 1 hour ago | parent
motbus3 1 hour ago | parent
qurren 1 hour ago | parent
1. Trolleys actually don't usually have steering wheels.
2. People who actually hit trolley switches are not usually the ones at the driver's seat.
mvkel 2 hours ago | parent
MichaelDickens 1 hour ago | parent
(...wait)
motbus3 1 hour ago | parent
They could only have stopped
zamalek 2 hours ago | parent
ryhminghistory 2 hours ago | parent
You try to go through files, and photos and the UI panics. You continue a conversation from your phone onto your computer and you lose part of the chat.
There are many more issues like this that are just so basic. You have bots that can attack governments but can't build a functional UI?
How many hours of ChatGPT does it take to implement a lock / consistency on a chat session so you don't overwrite it?
Buncha r*tards
mazone 2 hours ago | parent
digitaltrees 2 hours ago | parent
danielmarkbruce 2 hours ago | parent
irisflower95 1 hour ago | parent
danielmarkbruce 1 hour ago | parent
augment_me 1 hour ago | parent
Like Sam Altman meeting his husband in Peter Thiel's pool. Thiel funds a lot of these ventures together with Andreassen, who is on boards of non-profits. Dario Amodei's sister Daniela who is president of Anthropic is married to an EA non-profit founder who is also on the board of these non-profits and is tied with the prior mentioned investors. Elon is in there as well, Yudkowski is mingling with Altman, etc.
Blogs on this: https://contraptions.venkateshrao.com/p/ea-safety https://www.iankduncan.com/personal/2026-09-16-sex-ai-and-th...
There are some camps amongst them like the proponents for Regulation/Slowdown or Acceleration, but these are in practice mostly used for economical and not political decisions (like regulatory capture).
The point here is that this is a small group of people with a homogeneous background who are not really seeking input from anyone else on issues that are concerning most of humanity.
Like, if you said that the future of informational work and livelihood of humans is in the hands of 20-30 year transhumanists who think they are building mechagod that will trancsend social, political and religious separations of the world and bring everyone abundance, you would not feel like this is a serious thing to suggest.
pmkary 1 hour ago | parent
Aerroon 1 hour ago | parent
I feel like we're too far into the crying wolf part. Basically none of the doom and gloom scenarios have come to pass. Instead, AI has gotten better at censoring itself.
The biggest AI safety risk is when an AI tells a police officer "he's the suspect" and the officer believes the AI without confirmation.
reenorap 3 minutes ago | parent