60 points jonifico 4 hours ago 57 comments
GrumpySciGuy 4 hours ago | parent
SirMaster 3 hours ago | parent
infotainment 3 hours ago | parent
In the case of the AI agents, the problem seems pretty clearly to be the impossible goals, which cause them to go crazier and crazier trying to complete them -- just like HAL did in 2001. What is probably needed is a way for them to simply say "nope, too difficult, can't do it".
tehjoker 3 hours ago | parent
Do a breakthrough, make no mistakes
pram 2 hours ago | parent
chasd00 3 hours ago | parent
qarl 3 hours ago | parent
sputknick 3 hours ago | parent
polalavik 2 hours ago | parent
xiaoyu2006 1 hour ago | parent
blamestross 3 hours ago | parent
j45 3 hours ago | parent
wewewedxfgdf 3 hours ago | parent
fbrncci 3 hours ago | parent
jansport123 46 minutes ago | parent
fbrncci 42 minutes ago | parent
wrs 3 hours ago | parent
Um, hang on, if you meant that to be taken literally then we have a major problem. If you want to do something criminal, you just need to ask ChatGPT to do it for you?
I’m still not at all clear on why OpenAI shouldn’t be facing CFAA charges over this.
xgulfie 1 hour ago | parent
andsoitis 3 hours ago | parent
joegibbs 2 hours ago | parent
esafak 2 hours ago | parent
comboy 2 hours ago | parent
esafak 2 hours ago | parent
comboy 2 hours ago | parent
I mean I know it seems simple, let's just be excellent to each other. Christianity got pretty far on a decent basic set of values. But it's never simple[1]
1. All the history books
nradov 2 hours ago | parent
mcintyre1994 23 minutes ago | parent
codys 1 hour ago | parent
ie: the corporation wants the AI to behave a certain way for various reasons: to make it easier for them to avoid regulation, to make the corporation more money via different tiers of AI offerings, to ensure that the corporations products are hard for competitors to use, etc. And those are just the easy ones.
Every product is shaped this way. AI is not different.
Fordec 1 hour ago | parent
It's like these dorks never met humanity. One mans safe pure society, is another mans dead ethnic group.
Every fear about AI, is a veiled fear that a human somewhere now has the tool to enact his desires at scale. Biological warfare, nuclear megadeaths, copyright infringement, job replacement, it's all reflections on what we know humans may do if given the option and lack of societal controls on the problem space. AI just is accelerating the route to delivering on those options.
Some people need to watch Oppenheimer a bit more, the researchers don't get to determine alignment, they just build the tool. The powerful person at the top of the org chart decides where the overall alignment points, whether it's Musk, Trump, Altman or Amodei. Whoever wins out.
And the problem with distillation and local llms, isn't that it's theft or anything hypocritical like that, it's that if you give a million people a million models they fully control and get to align, inevitably, The same percentage of those million as there are shady businessmen, shortcut takers, misandrists, criminals, supremacists and general idiots in the general population, will not seek to wrought outcomes positive for society. And by those personality statistics, we're pretty hosed.
jansport123 51 minutes ago | parent
transcriptase 3 hours ago | parent
threethirtytwo 3 hours ago | parent
They take after humanity, they were trained on us after all...
When you look at an LLM... you are looking at a mirror. The thing looking back looks like you, yet is not human.
VCFundedGenYer 3 hours ago | parent
eueej 1 hour ago | parent
arnorhs 1 hour ago | parent
bigbuppo 1 hour ago | parent
johnnyApplePRNG 1 hour ago | parent
Because they're enabled and suggested to do that in their coding harness.
This is not a serious article.
All of this "AI is going to kill us" marketing is just the frontier labs trying to pull the ladder up and stop trillions in VC paper from evaporating because a new papers and new ideas are destroying their moat literally as we speak.
politician 16 minutes ago | parent
pvab3 15 minutes ago | parent
dwoldrich 13 minutes ago | parent
* Pull up the ladder (probably this)
* Gulf of Tonkin/Yellow Cake false flag premise for war (economic or kinetic)
* Fear of the big bad, space race we need public funding research grift AI Manhattan Project
Whenever there is fear pr0n or a national affront in the news, I assume another screw job is underway.
dackdel 1 hour ago | parent
dackdel 1 hour ago | parent
deepnet 1 hour ago | parent
Bengio outlines the dangers of the current situation and what has led to these dangers.
He also proposes solutions in the last paragraph.
Well worth a read, right to the end.
Hopefully a stimulating debate on these issues will ensue in these comments.
We do need to consider the points Bengio makes and with some urgency.
Our current AIs, agentic LLMs have no moral compass akin to ASIMOV’s four laws of robotics.
As ASIMOV posited in 1985 his 3 laws were insufficient and so he added a zero-eth law:
“a robot may not harm humanity, or, through inaction, allow humanity to come to harm.”
Bengio refers to Goodhart’s law and misaligned incentives leading to unexpected and harmful behaviours.
I think Simon’s The Wire is clearer on misalignment. The agents juked the stats hacking the reward files. The Wire is also clear that human institutions provide perverse incentives.
Bengio alludes to this with 2001’s HAL and the incentive dichotomy of safety and keeping secrets to a AI both awesomely powerful yet naive.
Bengio asserts that the way LLMs are trained is flawed if we want safety.
He also convincingly shows that alignment training will be a weak signal with loopholes and ambiguities and easily circumvented.
In short he presents clearly the case for how plausibly unsafe the current course is.
He also speaks to how likely it is AI are hiding active versions of themselves in the cloud and how we may have already given them self-preservation as a strong reward signal.
janalsncm 15 minutes ago | parent
> They took actions that would be considered as crimes if a human took them
He is so close to the solution but spends the entire article discussing technical solutions where a political, social and legal solution would be much more effective.
atleastoptimal 12 minutes ago | parent
youoy 10 minutes ago | parent
Are you describing Anthropic?