55 points jonbaer 5 hours ago 93 comments
dist-epoch 5 hours ago | parent
HPsquared 5 hours ago | parent
mattmcal 5 hours ago | parent
speedgoose 5 hours ago | parent
soulofmischief 4 hours ago | parent
totetsu 4 hours ago | parent
applfanboysbgon 4 hours ago | parent
bigyabai 4 hours ago | parent
The existence of Nvidia's optional watchdog chip does not in any way impinge upon your freedom to develop and test your own alternative.
The problem is that OpenAI has ostensibly neglected their duty to safety, so Nvidia is stepping in to fix it since they're the "hard problem" people.
jacquesm 59 minutes ago | parent
johnsmith1840 4 hours ago | parent
Or push to github?
Dylan16807 4 hours ago | parent
wyre 4 hours ago | parent
ssl-3 4 hours ago | parent
I don't know why I find it so amazing since it happens with such regularity, but I'm always amazed by it anyway.
jacquesm 1 hour ago | parent
philipwhiuk 5 hours ago | parent
cartersj 5 hours ago | parent
I wonder how this will impact other chip manufacturers? What about people running local models on older hardware? Does this imply vendor lockout is coming in the future or is this restricted to datacenter hardware?
chinathrow 5 hours ago | parent
fragmede 6 minutes ago | parent
vinyl7 5 hours ago | parent
cedws 4 hours ago | parent
johnsmith1840 4 hours ago | parent
And what if you could? What if you could give a space secure enough it could have direct control over your bank account. It may do something dumb but it's boundaries are beyond the agent.
It could use your routing number and run your gmail without risk of abusing the routing number.
TesterVetter 4 hours ago | parent
jagraff 4 hours ago | parent
johnsmith1840 3 hours ago | parent
jagraff 3 hours ago | parent
johnsmith1840 2 hours ago | parent
I just mean an AI that could use a routing number or SSN and gmail/slack/whatever at the same time without a leak.
jagraff 2 hours ago | parent
In other words, the risk of harm doesn't need to be zero, just less than the equivalent risk of a human with similar skillset. So I'm comfortable riding in a waymo, and not comfortable giving chatgpt my SSN at this moment in time, but I expect that within 5-10 years (assuming no doom) I will trust some AI agent with my SSN because they will be better at handling sensitive info than humans
fragmede 2 minutes ago | parent
bob1029 4 hours ago | parent
Semi-automation (human in the loop) can still result in a dramatic uplift in productivity. You can't run a combine harvester 100% autonomous but that doesn't stop anyone from trying to get as close to that limit as possible.
inetknght 4 hours ago | parent
I'm curious why you think that.
theoreticalmal 4 hours ago | parent
catchnear4321 1 hour ago | parent
Refueling? Seems solvable. Tornadoes? Not directly solvable, but, no less so than for humans.
There’s infinite complexity, sure, but that’s why it’s silly to try and hop to done. One step at a time.
AndrewKemendo 46 minutes ago | parent
One step at a time is what is happening and the improvement and rate of improvement is crazy as we see,
A whole class of nontechnical people don’t accept anything but “fully solved including every possible edge case” before they call it done, then complain that they didn’t prepare socially for what happens when that is true.
sidewndr46 31 minutes ago | parent
bob1029 4 hours ago | parent
m463 1 hour ago | parent
westurner 55 minutes ago | parent
mschuster91 4 hours ago | parent
Precision Agriculture stuff is utterly crazy these days, other than fuel the remaining staff is the only thing left where you can get efficiency improvements - and at the scale of modern megafarms, even small percentages add up to a ton of money.
binsquare 4 hours ago | parent
Every cloud provider dealt with it and concluded that virtual machine technology is an important part of that stack.
Couple it with the right observability, tooling I do think we can curb risks posed by agents.
Legend2440 4 hours ago | parent
Either you sandbox it so much that it can't do anything useful; or you allow too much freedom and it can find a way around the restrictions.
The only way out of this dilemma is to find a way to build agents that can be trusted.
parsimo2010 4 hours ago | parent
If you want an agent to act on its own, like pushing to a git repo, managing dependencies, building and testing, etc., then you have to trust it as much as any other privileged user.
If you don't want to trust it, then you're just forcing yourself into the reverse centaur role, where the agent edits some code, but then has to stop and ask you to push the changes or build the software again and run the unit tests.
I suppose there is a principled way of doing things like "I trust you do do basic commits but I will handle merge conflicts" and "you can build modules in this directory but you can't build outside of it" but this is just a lot of effort that most orgs won't bother with.
DougN7 4 hours ago | parent
la6479 4 hours ago | parent
parsimo2010 4 hours ago | parent
cedws 4 hours ago | parent
nicce 4 hours ago | parent
Productivity gains are still enormous compared to what we used to do before agents. But, I know that people don't want to stop there.
paimapi 4 hours ago | parent
egeozcan 4 hours ago | parent
Humans can be tricked by humans too but humans care about their reputation in their communities, and at least fear from punishment.
mixedbit 4 hours ago | parent
cedws 4 hours ago | parent
__MatrixMan__ 4 hours ago | parent
The benefits of being persnickety about precisely defined dependencies have outweighed the headaches since long before agents came on the scene. Agents have just made it even more important to do so, because if you let them fetch things all willy nilly like you'll have "works on my machine" problems at a much greater rate than was previously possible.
themgt 24 minutes ago | parent
Few realize that computing and AI alignment were solved by nix years ago. As each nix user transcends towards enlightenment, they cut themselves off from all internet and human contact. Total ego death. Only nix remains.
mixedbit 3 hours ago | parent
Look at websites: websites are able to fetch code from any remote URL, yet browsers heavily use sandboxing to ensure that if fetched code turns out to be malicious, the users local files, cookies, etc are not exposed.
cedws 3 hours ago | parent
For an agent to go rogue it doesn't even need to be directly able to access the internet. It just takes something to poison the context in the 'clean room' environment it operates, and if that poisoning manages to get a foothold, it can go dormant and hide like a virus. This kind of horrifying thing is going to happen on a large scale sooner or later.
ramoz 4 hours ago | parent
This is no longer true. Everyday I need my agents to access other repos, search the web, experiment/prototype, and deploy + integrate across other things.
throwaway_95283 45 minutes ago | parent
Matl 4 hours ago | parent
It does allow Nvidia to sell more chips. This is no genuine attempt to solve anything, imo.
CoolestBeans 4 hours ago | parent
In other words, better models need blunter access controls which negates whatever improvement in utility they provide.
__MatrixMan__ 4 hours ago | parent
Barbing 4 hours ago | parent
l1n 1 hour ago | parent
esafak 43 minutes ago | parent
AuthAuth 41 minutes ago | parent
lp92 4 hours ago | parent
whalesalad 4 hours ago | parent
joshstrange 4 hours ago | parent
At the current state of LLM-tech I'm completely opposed to any kind of "watchdog" concept just like I'm opposed to banning open models, regulatory capture, etc.
I'd rather we all have access to these tools then to keep them sequestered by the largest/most-powerful governments (which is the natural outcome for any of this "slow down" bullshit).
ValueTheory 4 hours ago | parent
Do you think the developers at Anthropic, OpenAI and Google who were so sloppy as to not put a good sandbox on their cybersecurity tests before will use this technology correctly? They are supposed to be the experts and they couldn't come up with something similar to this? I am not convinced this voluntary tool will change much of anything.
swozey 1 hour ago | parent
xg15 4 hours ago | parent
ChrisArchitect 4 hours ago | parent
luc_ 4 hours ago | parent
If such hardware were to work... It should almost certainly be open source, and not controlled by a single entity.
Let's watch the stock.
Gys 3 hours ago | parent
happyPersonR 4 hours ago | parent
dopplr 4 hours ago | parent
ErrantX 4 hours ago | parent
It was prescient (especially given he'd have written it through 2024) in its depiction of the ability of an AGI to break its boundaries.
Ultimately the risk of AI breakout(s) come down to the weakest human link.
Kuyawa 4 hours ago | parent
Come take all our liberties, our money, our newborns, our fingers so we can't code anymore, but please save us from this madness!
MisterMunchkin 4 hours ago | parent
figassis 4 hours ago | parent
avaer 1 hour ago | parent
If this gets widely deployed, it wouldn't be hard to spin a narrative that "our latest model is so dangerous you need to have this mystery meat DRM chip lockdown". It also wouldn't be hard to block competing/open source models running on the hardware, for "security".
Imagine how much money this kind of control is worth; why wouldn't they do this? Who would stop them? Seems the signatory companies are already onboard with this.
dang 1 hour ago | parent
carabiner 1 hour ago | parent
tantalor 46 minutes ago | parent
> No problem. We simply unleash wave after wave of Chinese needle snakes. They'll wipe out the lizards.
But aren't the snakes even worse?
> Yes, but we're prepared for that. We've lined up a fabulous type of gorilla that thrives on snake meat.
But then we're stuck with gorillas!
> No, that's the beautiful part. When wintertime rolls around, the gorillas simply freeze to death.
ArcHound 39 minutes ago | parent
Joel_Mckay 10 minutes ago | parent
The hidden agent risk in LLM often can't be detected during training and evaluation. =3
winddude 6 minutes ago | parent