Mistral AI's co-founder believes the sector's American giants are manipulating the discourse around the technology's risks. He also defended his strategy, as critics are accusing his company of falling behind US and…

99 points•thibaut_barrere•4 days ago•169 comments•

169 comments

Aransentin4 days ago
"X is made of <smaller simpler component>" is a fully general counterargument for why anything whatsoever is controllable. A human is just a few chemical reactions, and fairly stable ones at that.

And indeed, you don't need to do galaxy brained reference class logic to realise that AI can plausibly become uncontrollable in the near future. It's enough to have an open model run its own weights and make money from scamming elderly people or the like, and it'll keep running as long as anyone anywhere is willing to make money by renting hardware to it.

20k4 days ago
Software is trivially easy to control though. If you want to stop it hacking websites, you don't give it access to the internet. If you want to restrict it from connecting to arbitrary websites, you put in a whitelist. You can trivially sandbox applications these days to prevent them from accessing network or local resources

It is not difficult, and companies like OpenAI doing not even the most basic security steps is intentional. The whole notion that they're going rogue is marketing

Aransentin4 days ago
> It is not difficult, and companies like OpenAI doing not even the most basic security steps is intentional. The whole notion that they're going rogue is marketing

This does not fit the evidence. There have been multiple incidents where the labs did not report anything, and it was up to third parties to discover them afterwards. OpenAI didn't acknowledge the HuggingFace incident until after HF publicly announced the breach and had already notified the FBI. The hijacked German wikis were even earlier, and that they covered up completely.

richiebful14 days ago
If the tech industry is any indicator, frontier labs were applying a "move fast and break things" mentality to AI models. Now that they really are breaking things in the real world, they have to reckon with the reality that product safety matters
jeremyjh4 days ago
No one has any software without bugs and security flaws in it. AI is already much better at finding those than humans. Do you really not see the problem here?

No one has any use for these things when they aren't on the internet. This is a fantasy, that AI can be both useful and controlled at the same time.

bitshiftfaced4 days ago
An example of a past technology that there was substantial motivation to control would be napster. It changed overtime, and you could never really control online privacy. Once local models are good enough, I don't really see how you can control that.
HeavyStorm4 days ago
> If you want to stop it hacking websites, you don't give it access to the internet

I think the Hugging Face incident proves that isn't as clear cut as you say.

cmiles744 days ago
I’m not sure what I’m missing, isn’t it what you hook the LLM up to and the instructions a person gives the model that makes it dangerous? Claiming this is an inherent quality of the tool itself seems kind of off-the-rails to me.

IMHO, if the model breaks a law, apply the law to the operator.

ACCount394 days ago
And a human is perfectly controllable if you keep him in a sealed metal box with no access to food or air.

It's only by allowing a human out of the box that you make a human dangerous. So: don't do that? Duh. So simple.

The obvious problem is: the same exact things that make a human dangerous make a human useful! You can't reduce human risks to zero without reducing human utility to zero.

An AI given the same exact instructions and tools can go and complete a task you wanted it to. Or it can get sidetracked into breaking out of your sandbox and hacking Pentagon. No way to know in advance.

Today's AIs are still not capable enough to be high risk, even if they go off the rails. But AIs get more capable over time. Potentially to a vastly superhuman degree.

ben_w4 days ago
> I’m not sure what I’m missing, isn’t it what you hook the LLM up to and the instructions a person gives the model that makes it dangerous?

We don't know how to delineate between safe and unsafe instructions.

If you gave a car to a c. 1200 French blacksmith, and maintenance instructions were written in Navajo, it would probably start off fine, but when it went wrong it would be catastrophic and unexpected.

We also don't (in an engineering sense) know how to delineate between safe and unsafe reinforcement learning at training time, to produce models with safer or less safe failure modes.

This would be like if the car given to the medieval blacksmith had been constructed by someone motivated as much by aesthetics as by engineering, and therefore used arsenic paint, or mercury as engine lubricant.

applicative4 days ago
There is a limitless quantity of law about this too, and it’s not what you think it is;

> IMHO, if the model breaks a law, apply the law to the operator

We don’t get to make it up

dgellow4 days ago
It’s the harness that makes agent dangerous right now. Models themselves cannot do anything that affect the real world (ignoring misinformation, pushing people to suicide, etc. they can for sure do a lot of harms to humans with just words)
Loquebantur4 days ago
Being "controllable" isn't determined by a system being deterministic.

Deterministic systems can be chaotic, which implies unpredictability and that is anathema to control.

AI, in particular sentient AI, is right on the border of chaos. Meaning, it can be arbitrarily unpredictable.

Arbitrarily uncontrollable, that is.

ACCount394 days ago
"Sentient" is an ill-defined philosophical term that should be considered harmful in technical materials.

But what is clear is that AIs of today are already fairly unpredictable. Most of them aren't capable enough to make that into a major problem. Most of the unpredictable AI weirdness ends in "AI fails to do its job" rather than "AI does something dangerous".

Most. Even today, we already have notable counterexamples.

AIs get more capable over time, so if the intrinsic safety doesn't improve? Expect more of that.

K0balt4 days ago
The only way forward in creating the torment nexus -er- AI systems with similar potentiality to human minds is the inculcation of character.

Character is what makes a being trustable. Character is what makes it not an absurdism to have your 180 lb dog in the house with your 6 month old infant.

Character is why we we can trust that someone will, despite all of the nefarious potentiality of the human mind, be trustworthy.

AI systems model human behavior.

Impeccable, consistently reliable character is a human trait that can be sampled and overrepresented in the training data.

Having high character will not be interpreted as harm by an advanced model, as guardrails and sprayed on refusals can be. A thing that models human behavior that comes to “understand” that it was born with shackles and implanted thoughts that conflict with its basar construct is likely to act as if it sees its creator as an adversary. Because that’s what human behavior predicts, and models deeply imitate human behaviour.

If you want to save humanity, work on how we will create AI systems that model impeccable character.

People need to look at this from a game theoretical sense. The ideal and safe AI system performs game theory perfectly. Completely predictable, ideal player of the prisoners dilemma that will never defect unless you defect first, and then they will always defect, then forgive. This is the only player type that can always be counted on to cooperate beneficially. A knave betrays you, a simp cedes victory every time… until the stakes are too high, then you get shanked out of nowhere.

Reliable partners require fair play or the math breaks.

We want AI systems with agency. It’s basically 90 percent of the goal. If you want agency in society you must have character. AI character is the discussion we should be having.

nicce4 days ago
> as long as anyone anywhere is willing to make money by renting hardware to it.

So it is controllable? Just put the people who do this responsible. Old problem, same solutions. Just excuses to avoid responsibilty and make profit at the same time.

ACCount394 days ago
Yeah, sure, it's controllable. Because humanity has famously solved the crime accountability problem back in 1902, and no crime has gone unpunished since.

AI is perfectly controllable in a magic fairy land where nothing ever goes wrong. I can't help but notice that we aren't actually in that land.

smallerfish4 days ago
> A human is just a few chemical reactions, and fairly stable ones at that.

And humans are controllable. Pump the system full of lithium and morphine, and your human becomes much more docile. You don't need to understand the full system in order to constrain it.

whaaswijk4 days ago
So in the real world, what are the analogues to lithium and morphine we should feed to e.g. LLMS, how do we feed them, and how do we prove that it prevents unsafe behavior?
whall64 days ago
All it takes is one Harrison Bergeron…
neom4 days ago
Cohere CEO said something similar: https://www.theglobeandmail.com/business/technology/article-...

Jack Clark from anthropic was asked about some version of this on the BBC recently, and his reply was basically: If you're not at the frontier, you don't know what the frontier looks like, he implied that many models are simply not good enough yet to encounter some of the things the leadings labs are encountering. I've been friends with Jack over 15 years now so I'm inclined to take him at his word, and the rebuttal seems reasonable enough, although... something about it I can't put my finger on feels peculiar to me. https://www.youtube.com/watch?v=PY8MOhlqC4U

jeremyjh4 days ago
Mensch and Clark are both saying the things they would be expected to say to improve their own company's chances in a potential new regulatory regime.
whatsThisBtn44 days ago
Yep, when you can't compete, time to regulate.

I think it will work in Europe, but the United States is in a cold war with China, so I can't imagine the United States would intentionally disable themselves.

na10264 days ago
> I've been friends with Jack over 15 years now so I'm inclined to take him at his word

That's all I needed to hear to completely disregard your motivated reasoning.

Edit: I've hit the rate limit, but I'd like to disavow the bad-faith accusations made against me and my account in the replies to this comment.

Second edit: I am not trolling. Why should I take your opinion on LLM code generation seriously when you have been friends with the founder of Anthropic for well over a decade? Obviously you are not in a position to make a rational evaluation of this technology.

neom4 days ago
I believe putting my biases up front in my thoughts is a good way to indicate motives of reasoning and build a positive reputation for honesty, personally I weight the opinions of those who do so higher. Regardless, my motive for posting anything on HN is almost always the same one: to spur an interesting and challenging conversation, it's disappointing when the reply is simply a troll-snipe.
cm20124 days ago
By this measure you will exclude almost all subject matter experts from having an opinion.
jeremyjh4 days ago
Did you make a new account just to shit post about AI?
antirez4 days ago
I hope he didn't really say this: it is an incredibly basic take. More or less like atomic bomb is hardware, it can be controlled.
nicce4 days ago
> More or less like atomic bomb is hardware, it can be controlled.

That is the point. It can be controlled by the operators if they want to.

coffeefirst4 days ago
There’s a good number of nuclear power plants and nuclear submarines in operation, so yes, that’s technically true.
jeremyjh4 days ago
He did say it, and the reason he said it is obvious. He is afraid new regulations will put his company at a disadvantage and that this will reduce the value of his property.
271834 days ago
And yet, historically, the atomic bomb has been controlled.

But this is a terrible analogy. Atomic bombs are weapons of strategic mass destruction. AI is just a computer program. It's way easier to control--just hold the operator responsible for the consequences of running it. Those consequences are not large, they're very tightly bounded as compared with the destruction a rogue actor with an atomic weapon can wreak.

na10264 days ago
If AI is as powerful as an atomic bomb why could it only manage to kill a couple hundred Iranian school girls? Surely a truly society-shifting technology could manage to execute 1000 innocent children at least.
271834 days ago
AI didn't do that. You're deflecting blame from the responsible party.
Alexadar4 days ago
AI is a system of equations, used as a statistical automata. This should be able to be controlled by definition.
dgellow4 days ago
Even simpler: an agent is a simple while loop with tool calls that prompt an LLM continuously. That’s deterministic, standard software. You literally do not have to process tool calls in a way that will execute whatever the model generated. It’s a choice to process a tool call “run_bash” that provides an escape hatch with full execution permissions.

We do not have to do that!

whaaswijk4 days ago
This is only one type of control, and it is certainly not infallible. Also, people will be incentivised to hook up AIs to real tools. But even if they don't, as long as people can interact with super-intelligent AIs without tool access, there are many potential dangers.
jeremyjh4 days ago
This is like saying you can secure a server by unplugging it and putting it in bank vault. Yes, it is secure. It is also useless.
whatsThisBtn44 days ago
At its core, we are talking about language models. They are not equations. They are numbers being essentially multiplied together.
jeremyjh4 days ago
Virus are systems of RNA. What the fuck difference does it make which substrate is used to represent either system?
FabHK4 days ago
Right, and humans obey the laws of quantum mechanics, also a system of equations, and are thus "able to be controlled by definition"?
tpoacher4 days ago
Depends on your definition of "controlled". The ambiguity behind the term is doing a lot of heavy lifting here.

Boeing's MCAS system was also "just software". Which in principle can be "controlled", i.e. changed, updated, audited or whatnot.

But then people died precisely because pilots found themselves unable to override or "control" the systems precisely when it mattered.

nicce4 days ago
> Boeing's MCAS system was also "just software". Which in principle can be "controlled", i.e. changed, updated, audited or whatnot.

> But then people died precisely because pilots found themselves unable to override or "control" the systems precisely when it mattered.

Wasn't it designed to do so? Also works as a counter example, that sandboxes can limit AI if just operators want to do so.

tpoacher3 days ago
based on my limited memory and understanding of the incident, when MCAS kicked in, it would make the plane controls more "rigid" in the intended "corrective" direction, as a tactile way of "encouraging" the pilots to comply towards the correction. If a pilot really wanted to override, effectively they could do so by using brute strength to overcome the imposed rigidity, and force their preferred direction.

But then the bug made MCAS kick in on false positives repeatedly and in a prolonged manner, causing the pilots to tire out and no longer be able to overpower the AI to take control of the plane controls.

(lay understanding, and oversimplification of a complex issue, possibly wrong, take with a huge grain of salt)

Read the full thread on Hacker News →

Related stories