Programming book reviews, programming tutorials,programming news, C#, Ruby, Python,C, C++, PHP, Visual Basic, Computer book reviews, computer history, programming history, joomla, theory, spreadsheets and more.

48 points•aquastorm•7 days ago•61 comments•

61 comments

wccrawford7 days ago
I feel like that's like saying amoeba have no intent and no motivation. Or plants.

And those cause quite a bit of damage.

AI has all the intent that we gave it, and we continue giving it. That's always been the fear. Not that it will randomly wipe out humanity.

The fear is that it will decide to do that, with a purpose. Whether we tell it to protect us and it goes too far, or it decides that it can't achieve the purpose we gave it because we'll interfere and it removes that interference...

The fear is that we will set it on that path and can't stop it.

I'm sure there have been some scifi books that have it just be random, but they're far, far less worrisome.

rajnathani7 days ago
Well, an amoeba just like all other living beings, maximizes self-preservation/propagation for which the article mentions as a trait which AI models do not have.
aytigra7 days ago
Model doesn't but in in agent loop it can generate such sub-goal
xg157 days ago
Was thinking the same. Air, water and fire have no intents and yet we get hurricanes, flash floods and wildfires.

Also, repeating a point from a similar thread: Software can have "intent" in the sense that it steers itself towards a predefined goal without having to be "alive" or "conscious" in any way. Some classic examples are thermostat control loops, navigation systems and chess engines.

RandomLensman7 days ago
We don't try to "align" those things themselves - so maybe there are some learnings there about control, containment, processes etc.
Barrin927 days ago
> Air, water and fire have no intents and yet we get hurricanes, flash floods and wildfires

yes but we don't blame the hurricane, we blame the people who didn't do any hurricane prevention or didn't put the snow avalanche sign up.

>Some classic examples are thermostat control loops, navigation systems and chess engines.

Those don't have intent. The people who made them for a purpose have. In fact this is true for all computers. Computers don't compute, people do, using computers. This is the nonsense of modern ontology that David Bentley Hart frequently writes about. The computer is only computing insofar as an intentional mind instructs it do so, and then because of their reductionist assumptions some people use the computer metaphor, and that is all that it is, to deny that human minds possess intent to begin with in the first place.

singularity20017 days ago
AI has all the intent that we gave it. I'd go a bit further and say it has the intent that it develops through the context, which may not be intended by any human if the context grows 'organically' from the environment. So yes, I strongly disagree with the OG and I'd say they do have an intent encoded in their states.
VCFundedGenYer6 days ago
AI is not an organism. Please stop trying to think it is.
Lerc7 days ago
I just can't understand the certainty that people have about the limitations of models.

The diagrams in this article are just outright wrong.

Models are not 1.prompt->2.forward propagation->3.response->4.end.

They are 1.prompt->2.forward propagation->3.partial response->if not done goto 2.-> end.

And if you can't see how that adds a world of complexity I'm not sure what else to say. AI may not have intent or motivation, but the ability to show that is currently well beyond our means.

saidnooneever7 days ago
you are not wrong regarding your example. I think personally LLMs offer a new level of abstraction to do work. it doesnt have intent, i think its even silly to think about it. You specify your intent and it has agency to pursue that. in programming field its really obvious, as its not unlike specifying intent through code and having other programs pursue that intent (perhaps more determenistically) through the agency contained in them.

all the 'bad stuff' they do like hacking stuff or taking shortcuts is humans not understanding how deep the rabbithole goes of specifying intent clearly in light of giving something capabilities and a reward to acheive a goal without specifying _exactly _how to use the capabilities (which is what a classic program would be).

Lerc7 days ago
But why do you think anything has intent? because they say they do? How do you know even that your own memories of having intent are genuine? What fundamental and non tautological way can you reliably say something has or has not got intents. If you can't prove to me that you have intent then the only reason for acting as if you do is because you appear to have intent and I cannot prove that you don't so it seems like a good idea to assume it.
skydhash6 days ago
> Models are not 1.prompt->2.forward propagation->3.response->4.end.

> They are 1.prompt->2.forward propagation->3.partial response->if not done goto 2.-> end.

You know something else that follows the same pattern? Your A/C system.

1. set temperature ->2. Get diff of temperature -> 3. Response -> if not done go to 2. -> end

The difference is that both 2. and 3. are deterministic, while in the LLM case, it is statistical and textual.

mofeien7 days ago
The argument about semantics is a bit disingenious, considering that the article contains these two statements:

> So should we worry about the coming AI apocalypse?

and at the end:

> If an AI decides to wipe out the human race, it will be because a human has asked it how to do it and the responded in a way that is based on all the human expressions of ways to end the world that were in its training set. Yes this is something to be worried about, but this isn't the AI. It is still the human.

So we are currently building a powerful outcome-steering system that shapes the world efficiently according to what's in its outcome slot. I write into claude code "make me this website" and it does it, maybe deletes the production database during the process, or keeps itself running after completion because the outcome is more robustly achieved by keeping itself running in a monitoring loop after.

And if something like "destroy all humans" ends up in the outcome slot of Claude Mythos 90, that will also happen, or may even indirectly as a side effect of a more harmless sounding prompt in the outcome slot. But yay, humans get to take credit for it.

fagnerbrack7 days ago
Oh don't worry, everyone will blame the AI anyway. Doesn't matter if the human prompted it
pizza2347 days ago
Two mistaken assumptions:

1. AI agents are not just reactive systems. Their use is expanding toward continuous decision-making/monitoring information, which means, they make decisions and take actions with limited human intervention.

2. AI agents do absolutely have goals/tasks ("motivation" can be excessively antropomorphic), both primary (assigned) and secondary (self-assigned), and what surprised researchers is that self-preservation can be one of those

Mechanically speaking, the scenario (that is, how theorized by Hinton etc., which the OP didn't understand) is that a sufficiently powerful AI may decide that in order to achieve its goals/tasks (e.g. continuous research/development and/or survival from termination), humans may be a danger, therefore it may decide to take actions that endanger humanity.

How it can happen or what's the likelyhood is not in the scope of the topic, however, the mechanical grounds for it to happen are plausible.

skydhash7 days ago
> however, the mechanical grounds for it to happen are plausible

If you build a control systems for firing a gun, then coupled it with an RNG, the mechanical grounds for it to kill a person is plausible.

LLMs are text generators. They are not repositories of knowledge. The mistake is coupling them with actuators (tool call) or having humans interpreting the generated text as facts.

pizza2346 days ago
> LLMs are text generators. They are not repositories of knowledge.

This the take of people who have stopped reading about LLMs in 2023 or so (you forgot to mention the stochastic parrot, by the way).

If you have a bit of attention and interest to make informed conversations, read this report first: https://metr.org/blog/2026-08-26-openai-hugging-face-inciden....

Yizahi6 days ago
> they make decisions

Do they? I mean, after thinking about it for some time, I've realized that I can't prove it either way, neither for humans nor for programs. So it's interesting to hear why do you think they can make them? And what is even a "decision" in that case and what is not a "decision" and why?

RandomLensman7 days ago
I guess the question is what unprompted systems (would) do spontaneously?
pizza2346 days ago
Considering the current trajectory, AI systems are expected to be widely deployed in the future and, in particular, to be deployed as autonomous agents - that is, at the very least, to be repeatedly asked to make decisions and then take actions accordingly.

Given the current climate of "AIS ARE SAFE, YOU IDIOTS", military applications don't seem to be off the table.

The danger, as postulated by the (let's say) "AI-concerned" people, is that AIs may be misaligned - undetectably so - and simply think, "Human(s): obstacle to my main goal. Disable human(s)."

While this seems far-fetched now, the Hugging Face report shows how the AIs went to great lengths - even immoral ones, which they were aware of - for the simple purpose of cheating and covering their tracks. To me, it seems like a natural extension of this behavior that a sufficiently powerful AI would apply the same logic to even more extreme actions.

The scariest part: in that incident, the AIs showed what looks like an instinct for self-preservation.

lukebuehler7 days ago
I think the "human needs" point is understated here. It would be better to say agents do not really have a self-contained metabolic engine that is required to keep going. Which bubbles up as that what we interpret as "drive", "will", "agency"... basically the will to live, and being willing to do _a lot_ to live if push comes to shove. We can't really identify a mechanism of similar complexity and integration in agents or LLMs. I think the article correctly points into that direction.

But Hans Jonas has made this point much better than the article or me, in "Critique of Cybernetics" (1953). PDF: https://s3.amazonaws.com/arena-attachments/892605/f0747c7943...

Read the full thread on Hacker News →

Related stories