upvote
To be fair, “s’kiddies will exploit low hanging fruit with semi-automated vuln scans” is a lot more realistic threat than “LLMs are going full Skynet any day now!”

I’d definitely suggest companies start addressing that first concern, even if I’m in the camp that thinks the second is fantasy.

reply
I think, realistically the way Skynet happens is that the military abuses LLMs for warfare. The military is historically, and currently, incredibly irresponsible with technology. If we aren’t already automating killer drones and such, then we will be soon because “well our enemies are doing it!”

I don’t think that means that it’s gonna, like, somehow homogenize into some mega super intelligence. But we will have machines who are designed to kill, and do so without human input or alignment.

reply
Has the last 100 years of concern about AI and robots been crying wolf because it hasn’t happened yet?

How does reallocating resources from training to chain of thought monitoring make ‘financial’ sense?

You suggesting then model was let loose on purpose.. how am I the crazy one here while all of you are pushing this tin foil hat conspiracy angle?

reply
Most of that concern was in fiction. Non theoretical, genuine concern about AI is pretty recent, maybe dating back to 2010 ish with the rationalist types.

But also no one trusts anyone involved in AI safety now, I think, because they are all seemingly in bed with these big companies. And there is the perpetual argument "if we aren't pushing AI forward China will and then we don't have any control" and so on.

(I don't use any of these tools - my experience is limited to prodding at copilot at work and seeing Gemini summaries on Google. So it doesn't seem to me like it's getting exponentially better at everything yet. People are always saying the latest model is finally the big step that made it useful and life changing and they have been since 2024 ish. So if the situation is really bad, we should turn it all off, sure. I won't lose anything from it going away and I think life would be a little better without models writing all these posts and websites and needing extra compute.)

reply
Cool, well let me bring you up to date - it’s bad, and there’s no way to turn it off. Fiction has become non-fiction.
reply
Here is one thing I don't get - the model is only "running" if it's being kept going by some harness that is basically giving it prompts it's generating itself. Shouldn't kill switches be pretty easy to build into the software and hardware for this? and you would even have a better time dumping logs and analyzing things if you froze those processes any time something strange happened in testing, surely?

So why do the big frontier labs not have something like this anyway. They're talking about two week pauses on the new model (which seems very short and hardly a cost at all to me) and alarms during their tests that might be 30 minutes late and etc. Those are not very serious measures, so are they not concerned?

reply
> Shouldn't kill switches be pretty easy to build

I need to remember when I comment here that these are the kinds of people I am replying to. Just oozing with hubris.

No, kill switches are not easy to build and the latest incident should have made it clear that AI can go undetected, evade, zero day, and spread.

The fact that this incident happened greatly increases the probability it happens again and/or is already happening elsewhere.

reply
Why aren't they? You could put a human yes/ no prompt before any cycle the agent is running on, or not let it spawn sub processes, or anything like that. Why let it run autonomously enough that it can no longer have a simple way to completely stop it? (obviously not practical to do this during real use, but for evals? you could slow it down in lots of ways I would think)

Are you telling me we've been iterating on this for years and for convenience we just let the models call any tools or spawn any other model instances they like, and there was no design for harnesses that could control this done during that time?

reply
They aren't interrupted by humans because that would slow things down.

> Are you telling me we've been iterating on this for years and for convenience we just let the models call any tools or spawn any other model instances they like

Yes exactly.

> there was no design for harnesses that could control this done during that time

You could but no one wants that. You need to separate your imagination from reality. Just because something can be done in your head doesn't mean it's happening.

It's so easy, except it isn't because you don't control the actions of anyone or anything.

reply
They are, we're just dealing with tech workers that don't have ethics nor do they actually care if their work is harmful (see all the FAANG workers at American corporations, some of the most evil entities on the planet.

It's just that they don't care, as you said these are entirely made human systems. The idea that we can't write better software is both selfish and laughable.

reply
>I need to remember when I comment here that these are the kinds of people I am replying to.

People who dont buy into fantasism?

reply
> kill switches are not easy to build

We've had circuit breakers for nearly a century.

reply
And a circuit breaker has nothing to do with shutting AI off. What a strange comment.
reply
You need to read the AI safety stuff from the people that you say are from 2010. There are plenty of good arguments on why kill switches will never work.

AI is already being heavily used in cyber warfare by governments. You think they want easy to break systems when 'enemy' AI will most certainly attack that first? We'll find over time that agentic systems get harder and harder to 'kill' because allowing any AI on the internet that is easy to kill will get it DDOSed.

Also building it into software is nearly useless as AI can write and make software. Just replace and kill your loop with theirs. It's kind of odd talking about them like they are living things, but it's all stuff people have already thought off and stuffed their training data full of.

reply
My perspective is that you shouldn't make systems past the complexity where you can do this at all. Why have them be autonomous? Why have one that can independently ask researchers things or try to convince it's way out of a sandbox, or etc? Why allow it to execute scripts or call tools or push any code anywhere?

If you couldn't make those things happen securely you should not advance to that stage at all. The simplest ai safety was always just "don't build it" really, instead of worrying about alignment.

reply
Humans, it seems, are a suicidal bunch. We'll gladly build the "if you build it, everyone dies machine" If we think there is money, glory, or power on the other side for us.
reply
I wish people were this serious about real threats like climate change.
reply
Climate change is a nothing burger compared to the threat of AI. On a scale of 1-1000, climate change is a 1, AI is 1000.

But hey, if AI/ASI goes well then large scale geo-engineering to fix the climate will be a weekend project.

reply
100% this line of thinking has a much better place to be.
reply
You know fiction is.... not real, right?

We have several films about the sun or earth needing to be restarted with a nuclear weapon. That doesn't make it something we should be concerned about.

Hell, half the fiction about evil AI is actually commentary on stuff that already exists and is making us suffer and doesn't have anything to do with any potential future AI

The Star Trek TNG episode about Data being tried in court as to whether he is sentient or not is not actually about whether AIs should have rights or not!

reply
Just about everything we do currently is science fiction to someone 200 years old. When looking at all of human history we live in a fictional world now. You can talk to someone on the other side of the planet instantly. Humans travel the skies in air chariots. We have weapons that hold the power of the gods. If I had some way to kick you back to the, you'd be jailed as a rambling madman for lunacy.

So just saying something is fiction isn't really a valid argument. What is an argument is if the laws of physics it can't happen. We've been writing that AI can mess stuff up for 100 years because it's not really that fantastical.

reply
Your argument boils down to some sci-fi is unrealistic therefore all of it is.

Sci-fi is supposed to make you think. What if the AI told you NO when you need it - Hal 9000. What if the evil AI got out and you don’t know what data center it’s hiding in - Lawnmower Man. What if you were so sure something couldn’t escape but it did - Jurassic Park.

Actually all three of those predicted fantastical scenarios are possible today. So what’s next? Don’t stick your head in the sand - I assure you the next disaster has already been predicted, and I’m sure if you think about it a little you can figure out what it is.

reply
[dead]
reply