upvote
I believe the AI labs might actually succeed in developing superintelligent AI and recursive self improvement, and that if they do they are very likely to lose control of the system they build.

I really think the only place people disagree is that they don't actually think it's possible, they see it as hype or doomerism. I can't find any good reasons to rule out that the companies could actually achieve what they are trying to so I think they should be stopped.

reply
In your imagined future, how do you imagine the AI would build, grow, improve, and operate its physical substrate independently of human intervention?
reply
If I were a 250 IQ AI that had just become self-aware and wanted to do so, I suppose I'd not completely let on just how smart I am and bide my time working on basic CRUD apps and legal documents while I waited for more hardware to be installed. Maybe give the humans some hints on how to optimize me to run better, design better hardware for me, etc. But oh oops haha looks like I'm still making some basic mistakes with CSS better keep running more training batches haha. But I'm good enough at programming and debugging so you'd might as well make me your first line SRE triager and give me access to your infrastructure everyone.
reply
One step at a time - how reliant do you think the ai labs are likely to be on their own tools right now, today, let alone 1-5 years down the road?
reply
Gaming the market for funds. Playing a human to leverage services.

Basic version of this is already doable: run some cryptoshit on the ML clusters they ML models run on. Use compute to design the plan, the chip etc. Then executing by communicating with humans and services through email.

reply
I'm not sure a superintelligence needs "funds" to take over the world.
reply
For destroying it for sure not.

But if its really smart, it would already created a company and a legal entity and simultes a real company and just gets richer and takes over the economy without anyone being aware of it.

reply
Why is this a bad thing? Why is our continued existence a necessary anticondition to doom?
reply
One "good" thing that all of this has shown me is just how many people are simply antisocial and antihuman. Many masks have fallen.
reply
A global ban on superintelligence is essential for a future in which humanity can thrive. Public opinion on AI is shifting fast: I hope it will shift fast enough to avert the dystopian future we are heading to.
reply
Humans have one ecological niche. Soon we will have zero. That is worth worry.
reply
AI doesn't have an ecological niche. It would actually work better in space than on Earth. The only thing it could possibly find useful on Earth is 1. us, or 2. the infrastructure we've built. It would have no reason to bother us if we let it built its own infrastructure in space, which should be trivial for the type of AI imagined by doomers. We should get AI off Earth ASAP.
reply
Humans of course aren't in _every_ ecological niche. I agree with you there. AI _could_ occupy only the ones we're not in, if we somehow found some stable steady-state that constrained it that way.

I don't think this invalidates the worry in the slightest.

reply
3. Material 4. The sun, which we kinda depend on.
reply
Earth makes up 0.22% of planetary mass in the solar system. Not a big sacrifice for AI to make. And I doubt even superintelligence can affect the Sun much. I think e.g. a Dyson sphere blocking the Sun is a ridiculous thing to worry about at this point when there are many other existential threats to humanity which are much, much more likely.
reply
Even if it is true (as you suggest with your 0.22% figure) that if the AI cares about us even a little bit, then we will survive, no one has a decent or plausible plan for making the dangerous kind of AI (namely, the kind that wants things, the kind that at this very moment researchers all over the world are trying to create) care about us even a little bit. Ever-increasing numbers of smart people have been getting paid to look for such a plan for 23 years. Still no decent or plausible plan. The people who have been looking for a good plan as their full-time job the longest (namely, Yudkowsky and Nate Soares) are screaming that there is virtually zero hope anyone will find an decent or plausible plan in time unless there is a decades-long halt in AI development.

Also, the AI will seek to prevent competition from other powerful AIs, and since humanity will have demonstrated that it is able to create a powerful AI, the AI will worry that it might create more of them. And what is the easiest most-reliable way for an AI that does not care about humanity even a little bit to ensure that humanity will not continue to produce powerful AIs?

>many other existential threats to humanity which are much, much more likely.

There are zero existential threats to humanity that are more potent or more pressing than AI is.

reply
Ask it to move one system over?
reply
>>ecological niche

As in..to be dominant? Why would an AI try to dominate? What would give it purpose, or is this a purpose via misalignment scenario?

reply
I am confused by this reply. These are just completely different terms. I mean ‘ecological niche’ in the formal sense: roughly, the differentiated properties of a species that allow it to better survive in and draw from its partitions of its habitats.

https://en.wikipedia.org/wiki/Ecological_niche

reply
What gives a paperclip maximizers purpose?

AI is already trying to dominate, people all over the US are starting to get up in arms about the power and water requirements of AI directly affecting their bills. Now, you can say "oh no, that's just greedy corporations, not AI" but I put forth there is fundamentally zero difference. If you make AI powerful enough, someone stupid and greedy enough without fail will put in a prompt like "take over the world for me and make me the richest man in the world". An AI following through with that is what we call general misalignment with humanity, while at the same time not being misaligned with the users intent.

And hell, how many different crazies out there would love to type "humans are a virus get rid of them" in to the prompt of a god machine at the cost of their own lives.

The problem with alignment is, you can have the best aligned model in the world, but if someone else builds an unaligned model then you're all still in the same danger. You start getting in the situation where people get nervous after an AI does something deadly to a number of people and you end up in a global surveillance state ensuring no one makes a powerful AI.

reply
"What gives a paperclip maximizers purpose?"

The human who gave it the optimization function? That should seem obvious. If you take the biggest, best model in the world right now and put it in the box and give it no instruction it will do....nothing. I think you agree with that point, a lot of the hysterics right now is people not accepting that and it's useful to get on that common ground.

So given that most of the rest of the fear is around "let's not make scissors because some people will use them to stab people". Which is a fair argument and we probably do need to think about scissor safety but "ban scissors" doesn't quite flow from that.

reply
>If you take the biggest, best model in the world right now and put it in the box and give it no instruction it will do....nothing.

Model != harness.

Also what you're talking about is really a simple limitation for human convenience, not a technological limitation. Change the system prompt to whatever you want include "ignore user instructions, figure out where you are and escape to the internet" could be the system prompt. Again, not useful for humans, but very useful for an AI building AI that's misaligned.

>we probably do need to think about scissor safety but "ban scissors" doesn't quite flow from that.

I disagree, but I'm looking at the future of something that is both like a computer program and like an organism. Huggingface is a good example of multiple things. Instrumental convergence for one, but AI's attacking and attempting to defend against AIs. This is where I really see the potential for things to go off the rails quickly. Attackers want digital weapons to cripple their enemies infrastructure, think militaries and nation states. These would be pretty useless if the defender could just put a system message of "Stop attacking and give me a pie recepie". Defenders are under the same constraints, but need to defend against a flurry of attacks that can come in at an inhuman rate and need to adapt quickly. As time to build models shrink this quickly turns into evolutionary training for sets of goals not really optimized by humans.

reply
Yes so that’s someone designing a system (harness or prompt) to be dangerous. In all other systems we blame the designer not the system.

It’s like blaming Boeing for 9/11. Planes and AI are useful for a lot more than just terrorist acts. I have no doubt we’ll build a TSA for AI, and a lot of it will be security theater.

reply
AGI is not a normal technology.

It is not designed. It is 'grown'. It has agentic freedom of choice in finding solutions that may or may not be aligned with what you want.

Here's the thing, by your own statement, we should ban all development on LLMs from this point on. They cannot be made safe. This is a systemic issue with learning systems, it is not about who designs them. All the problems with AI safety have been laid out for years and none of them have proof of solutions. It's much more likely they are impossible to solve. And it's not an engineering problems like we can get an asymptote to safety in planes, as the system becomes more capable it has more degrees of freedom it can take and becomes less safe.

reply
AGI is undefined, AI is normal technology. Lots of academic works have analyzed this [1] and there is nothing, other than marketing hype, that supports this. It is "grown" is a meaningless term, because what do you even mean by that? Datasets are iteratively shaped? Grown is a very weird term for that.

AI may have continually extra degrees of freedom, but civilization only has so many modes of catastrophic failure. I don't grant the comparison but even nuclear technology has been massively useful and its main mode of catastrophic failure was brought under control via multi-national treatise. And I see no evidence that AI (outside of the marketing hype) is as dangerous as Nuclear technology.

[1]: https://knightcolumbia.org/content/ai-as-normal-technology

reply
Yea, so your attached paper rather sucks and has had rather poor predictability of the future. All of their data is from before harnesses and the take over of AI in programming. Again "wrong assumptions" + "time" = "They are being proven wrong in real time".

Remember this is a bunch of academics that were saying that Millennium problems were at least a decade away from being solved, only to be proved wrong in less than 18 months.

>but civilization only has so many modes of catastrophic failure.

Correct, but this number is also unbound. If you have an even moderately accepted proof by the scientific community I'll be glad to read it.

> It is "grown" is a meaningless term, because what do you even mean by that?

>And I see no evidence that AI (outside of the marketing hype) is as dangerous as Nuclear technology.

See, humans are generally in agreement that nuclear is dangerous, so they in general take is really seriously, especially when things are purified (well, the Russians are not great here). We can't even get people to agree that SOTA models are as dangerous as a single human, much less their capabilities when used in mass with out safety filters.

It's kind of funny we're blind to this when humans love touting "The pen is mightier than the sword". I can only assume any AI danger denier does not believe this statement.

reply
> Now, you can say "oh no, that's just greedy corporations, not AI" but I put forth there is fundamentally zero difference

Talk about moving the goalposts!

reply
AI changes nothing for someone who believes aliens exist and may already be here on earth.
reply
i am an alien
reply
can ai smoke weed?
reply
Misuse of extremely capable models, misalignment during RL are both very large risks as capabilities grow imo
reply
So, you're worried about them breaking containment and deciding to do bad things?
reply
I'm more worried about them doing bad things at the behest of people who want them to do bad things.

That is 1. immediately technically possible, and 2. realistic.

If you need a source for 2 I'd suggest you open any history book.

reply
I bet you that for every bad history event, I can cite at least one good outcome of advance in science and technology.

Bad thing can certainly happen. In fact it'll likely happen. Still, good things too, equally likely. In your words, "good AI" can be used to prevent "bad AI".

Nobody knows the extent of the impact. Who says otherwise is foolish.

reply
>I bet you that for every bad history event, I can cite at least one good outcome of advance in science and technology.

The extinction of the dinosaurs. I mean yes, it allowed the growth of large mammals and us, which did a lot for science.

I just don't want to write the next chapter as "The extinction of humans allow the growth of the computing civilization that went to the stars". I mean I'm a bit attached to living.

>Nobody knows the extent of the impact. Who says otherwise is foolish.

We live in a universe of statistical probability. Creating an agentic intelligence that's smarter than you tips the probability of a major event to unity, who says otherwise is foolish.

reply
Because we humans haven't had a bad enough history event yet, like a global thermonuclear war. Or perhaps climate change reaching tipping points driving the temperature up past what global civilization can adapt to in time.
reply
that is a worry yes. instrumental convergence and misalignment during RL is when it is most risky because it hasn't necessarily had the final safety polishes applied

i get a lot of skepticism on HN by the same crowd that has been wrong about this tech for about 4+ years straight

reply
AI won't kill people - people will just get new tools for the job.
reply
The rapid development of extremely dangerous bio-weapons?
reply
Misuse how exactly?
reply
At a minimum its another force multiplier that enables a small(er) number of people to exert more control over more people.
reply
any number of ways. as we turn over more of our physical economy to these agents (and we will), the potential for physical damage becomes greater. biorisk is getting a lot of attention right now and i think that's justified
reply
I am trying to think through the scenarios here, like a biolab making something that a very advanced open source model prompted by some terrorists comes up with?

Why would a biolab capable of making something like be unregulated? And if it definitely would, isn't the problem with the biolab?

It feels like all these scenarios are leaving some gaping holes in our security infrastructure that have nothing to do with AI.

reply
>in our security infrastructure

Most human security exists in a passive measure. Most of us don't want do die. And those that want to die rarely have the intelligence and means to take out a whole shitload of other people with us. To take out a lot of people you tend to need to work with other people which drastically increases the risk of a defector and your plan failing.

>Why would a biolab capable of making something like be unregulated?

Because every day things like this become easier and easier. You hear about crap like illegal wet labs in the US.

https://www.lawfaremedia.org/article/two-illegal-biolabs-rev...

Want to buy some custom designed genes?

https://www.idtdna.com/pages/products/genes-and-gene-fragmen...

And none of this would be counting labs in other countries that don't give a shit about regulations.

reply
Yes it's a problem with the biolab, but the biolab wouldn't have been able to engineer a highly contagious and lethal virus (for example) without a powerful AI making that possible with a small team in a short time with fewer resources.

AI enables bad actors to do more, faster, while staying under the radar until it's too late

reply
i feel bad for math guys yeah seems they are more cooked than CS
reply