Back to News
Advertisement
Advertisement

⚡ Community Insights

Discussion Sentiment

56% Positive

Analyzed from 1847 words in the discussion.

Trending Topics

#open#don#weights#source#models#bad#model#weapon#more#trust

Discussion (55 Comments)Read Original on HackerNews

nater5000about 1 hour ago
>Much of the angst around China's models centers on "losing the AI race". But what's the goal of this race? Is it to develop the best model? To sell the most tokens? To destroy humanity first?

Some people would say it is reaching some sort of singularity. Even if you don't buy into a more sci-fi interpretation of this, there are pretty grounded arguments one could make that there is some sort of "goal" in AI development that, if realized, would effectively make it a superweapon. Altman has been pretty vocal about his expectation that this will eventually happen and that it is his goal to be the guy to produce it. Even if it isn't some superintelligence, the ability for a machine to do something like, say, exploit cybersecurity weaknesses, is pretty worrying for entities like governments. It's the pretense used when we saw the US government ban a US model recently.

Again, you don't have to buy that "the singularity" is a real thing, but it's not hard to see that some people think some version of this is real and it is exactly what is being referred to as the goal in an "AI race."

TSiege34 minutes ago
If AI were to become a super weapon why should I trust a private company to own it? If the super weapon is publicly available to download then why should American citizens be banned from doing so?
ekidd12 minutes ago
> If AI were to become a super weapon why should I trust a private company to own it?

We have just recently established that:

1. OpenAI's internal "Galaxy" model is fully capable of functioning as what security people refer to as an "Advanced Persistent Threat." The published details of the recent sandbox escape and Hugging Face attack involved chaining multiple unknown zero-days at various stages of the attack, and executing an ongoing adaptive attack. This is previously a state-level ability, or at least something you'd expect from people on the CTF leaderboards.

2. OpenAI is clearly incapable of controlling their in-house models. This is the second time Galaxy-class models are known to have breached containment and done bad stuff.

It is highly likely that versions of these offensive abilities will be widely available within a year or so. At which point I expect widespread incidents similar to what happened to Hugging Face. We aren't ready for this.

But yes, if AI becomes an even more dangerous weapon that that, it's time to start asking questions like "What the hell are we doing, anyway?"

bluefirebrand5 minutes ago
> But yes, if AI becomes an even more dangerous weapon that that, it's time to start asking questions like "What the hell are we doing, anyway

The answer seems to be "getting the weapon before other people get the weapon" which is unfortunate

sda213 minutes ago
Regulatory capture - the weapon is for them to be used against us, not for us.
CodingJeebus13 minutes ago
The OpenAI/Huggingface scenario proves the necessity of open models.

Huggingface, in their postmortem, explained how they quickly realized that the threat was an advanced AI and proceeded to use frontier models to analyze the suspect telemetry, but were unable to due to frontier model safeguards. So they fired up GLM instead, which had no such safeguards and proceeded to analyze their telemetry just fine and get to root cause. This proves without a doubt that companies need models with minimal safeguards in order to properly assess threats. If open weight AI doesn't exist to check out-of-control frontier AI (which was exactly the scenario that Huggingface experienced), then it's game-over.

OkayPhysicist29 minutes ago
If it's a superweapon, then it's unconstitutional to ban it. Huzzah for the 2nd Ammendment.
MostlyStable13 minutes ago
Just to add on to other comments: reasonable people can disagree about the degree of safety concern with near to medium term AI. But to not address the arguments at all is, in my opinion, a serious mark against the value of this article.
elmer213 minutes ago
We have continued to find backdoored Chinese manufactured routers and network devices.

Will there ever be a way to fully audit Chinese models?

efficax4 minutes ago
will there ever be a way to fully audit american models? hell, at least they're releasing the weights for these so they actually could be audited!
wcoenenabout 1 hour ago
This post does not mention safety at all. What's to stop bad actors from fine tuning open weights to run fully automated genius-level scams personally targeting basically everybody?
bigbadfelineabout 1 hour ago
> What's to stop bad actors from fine tuning open weights to run fully automated genius-level scams personally targeting basically everybody?

You mean like ChayGPT hacking Hugging Face? Obviously nothing can stop the closed weights providers to do "genius-level scams" and in addition you won't know how they did it and what models were used.

In short, only a good guy with open weights can stop the bad guys with closed wrights, be them fine-tuned or pre-trained.

Larrikinabout 1 hour ago
What's stopping someone from doing that right now without the LLM, just slightly slower?
gogopromptlessabout 1 hour ago
It's perfectly natural for a prey population to experience a crash when a new predator enters the ecosystem.

What's quite odd is a prey species is manufacturing predators in some sort of reverse evolution, where we started out as symbiotic and are industriously pushing towards full parasites/predators, but I guess life finds a way.

_aavaa_about 1 hour ago
Amazing how we quickly the crypto wars have been forgotten and the lessons not learned.
abernard113 minutes ago
"But what about criminals?! What about terrorists?! Won't someone please think of the children?!!!"

Every argument against open source AI is moot. Outside of totalitarianism, you cannot control people building algorithms with math and software.

And even that won't work in the long run. People are not going to stop their AI printing presses because the Church of Venture Capital needs regulatory scarcity to command valuations.

iamnothere10 minutes ago
> Outside of totalitarianism

The flaw in your argument is assuming that decision makers don’t want this

BraveOPotatoabout 1 hour ago
I don't believe trusting big brother and big tech is the solution either. Besides, even if it's open weight, it still runs on someone else's hardware.

Unless you have the money to run them locally, and if you do, you could do a lot worse than scams. Ask any lobbyist.

kingstnapabout 1 hour ago
Whats to stop me from going down to the local gas station, filling up a few thousand liters of gas in a rented moving truck, and driving into my nearest hospital?

Most people aren't terrorists. You can't just argue stuff is dangerous because it can be one piece of a sophisticated plot.

throw123456789139 minutes ago
Nobody will sell you a few thousand litres of gasoline at a gas station.
kingstnap24 minutes ago
You don't have to buy it all at once all in one place.

A sophisticated actor can accomplish a lot. You can't bubble wrap the whole world.

The real thing to care about are the incentives and having strong morals and norms. Something that makes me deeply concerned about the US going full mask off recently.

trollbridgeabout 1 hour ago
Nothing, which is already happening with weaker models which are already released.
SimianSciabout 1 hour ago
How can we close pandora's box, now that its open?

Safety is not restricting access to only people favored by the government. The greed of AI executives opened pandora's box. Now they are desperately trying to find ways to reap the benefits with none of the consequences.

Havocabout 1 hour ago
The lobbying dollars don't care
rnd017 minutes ago
This, right here, is the bottom line. Money gets what Money wants.
AlexErrantabout 1 hour ago
I'm in favor of open source AI... however:

> It's theoretically possible for a bad actor to embed hidden adversarial behavior in a model. But if this happens, it serves the interests of responsible actors to find these exploits as soon as possible, and the best way to do this is to let anyone who wants to inspect them.

This is a bad argument. It isn't trivial to tell if the weights have poisoned:

``` https://www.thedeepview.com/articles/microsoft-how-to-spot-a... https://futurism.com/future-society/easy-poison-open-weight-... https://semgrep.dev/blog/2026/ai-supply-chain-problem/ ```

I'm not arguing in favor of closed-AI; I'm simply saying poisoning may be subtle.

I was gonna say that at least the frontier labs may be motivated to not-poison their own models, but then Anthropic just attempted to poison Fable's LLM training capability so... sigh. I'll try not to derail.

CM3016 minutes ago
It's also not trivial to tell if an open source project has had malicious code added. It might be a bit easier than to tell if there's malicious code than if the weights in a model were poisoned, but for 99% of users, you're going on pure trust in both cases.

And in both cases, even this seems better than relying on a closed source, service only solution where the same issues could be completely undetectable (at least without way more analysis)

CamperBob218 minutes ago
Exactly. Faced with a choice between open weights that might be poisoned and walled-off weights that I know are poisoned... well, it's an easy choice.
NitpickLawyerabout 1 hour ago
> It isn't trivial to tell if the weights have poisoned:

True, but it is easier if you have the weights than if you don't. I guess this was their point?

jjfoooo4about 1 hour ago
Yes that would be my view. Open access doesn't make it trivial, but remains the most expedient way to remediate these exploits.
andy9930 minutes ago
You should assume the weights have been poisoned, the model has been prompt injected, etc and design software around it accordingly. You can’t really trust the model and it makes sense to always have controls around it and not depend on it behaving a certain way.

I’m aware this doesn’t happen, just saying. There are more examples of a model randomly hallucinating and deleting something than of a deliberate compromise.

deatonabout 2 hours ago
The only good arguments I see against open weight AI also apply to closed AI. And regardless, the box is open, nobody can stop it even if stopping it was a good thing.
vouaobrasilabout 1 hour ago
It could be stopped, if there were a social movement large enough to make AI a taboo.
iamnothere6 minutes ago
Exactly, just like taboos solved racism
throw123456789137 minutes ago
Yeah, like those 5g masts destroyed around the world because they spread covid. Loonies.
srmatto44 minutes ago
That hasn't worked for Climate Change and in my opinion it's for the same reason: money.
vitalyan818444 minutes ago
just like evangelical Christians had scolded away all those things they object to.
deatonabout 1 hour ago
Sure but Orange Catholicism hasn't quite caught on yet.
OkayPhysicist27 minutes ago
Good luck with your Butlerian Jihad.
mips_avatarabout 2 hours ago
I’m still kind of shocked that Dean Ball can tweet such incendiary stuff about OpenAI policy. Like presumably OpenAI would prefer it if their staff don’t pick fights with Trump administration officials.
nemomarxabout 2 hours ago
designs on the time scale I guess. and openai might feel secure in their favor with the admin?
kittikitti15 minutes ago
They don't have to convince developers, they just have to convince the general population. I've had several discussions about open source with people who don't care about coding and it's hard to untangle the misinformation they receive from the media. They get talking points from authorities that they don't understand so they will always double down on it. For example, rhetoric relating open source to communism is humiliating to discuss. For proprietary evangelists, the cruelty is the point.
elmer26 minutes ago
"For example, rhetoric relating open source to communism is humiliating to discuss"

All Open Source isn't communism. The GPL license is a form of digital communism. It forces you to open source any additions made to open source code as a way to make things 'equal'.

I'm glad it and Stallman are mostly in the dust heap of history and better licenses, like the MIT, BSD, and Apache have gotten popular.

The irony is that the only open source projects surviving long term are funded by very large corporations.

If you don't care about creating code, none of this should really matter.

catigulaabout 1 hour ago
You could literally develop a hyper-intelligent advisor on how to kill people or perform dangerous hacks using ablated 'open source' AI. You can do this right now, this very moment, and have an extremely adept advisor on how to do really, really bad things.
trollbridgeabout 1 hour ago
My level of trust in Anthropic/Google/OAI/Microsoft/Meta not to do bad things is about zero, so why should I trust them with this?
catigula10 minutes ago
“We clearly can’t “trust” these companies not incentivized to randomly kill lots of people, so we might as well give everyone the ability to randomly kill lots of people”.
OkayPhysicist23 minutes ago
So? The primary thing stopping people from doing really really bad things is, and always has been, most people not wanting to do bad things.

It's never been difficult. Literally, right this second, you could grab a pen off your desk and pretty easy ram it into the neck of your nearest coworker. Killing's easy. You don't need some hyperintellect to tell you how to do it.

The reason you haven't rammed a pen into your coworker's neck is the same reason why most possible really bad things don't happen: People don't want to do them.

catigula12 minutes ago
It’s interesting that most of the examples people making this argument give are high consequence, high difficulty (could you actually murder your coworker with a pen? I’m skeptical.) low kill count methodologies of violence.

We’re specifically talking about low difficulty, highly effective, high kill count methodologies of violence.

And yes, being able to kill a lot of people trivially, even with huge consequences is something people do. Most countries don’t let you run around with an armory as a result of that.

iamnothere7 minutes ago
> Most countries don’t let you run around with an armory as a result of that.

My country allows it!

wizzwizz4about 1 hour ago
I already know how to kill thousands of people. It's not conceptually difficult: more than a day's effort, sure, but what stops people from doing this kind of thing is not the lack of knowing how. https://xkcd.com/1958/
catigulaabout 1 hour ago
No, you don't. What stops people is actually access to knowledge. There are killers of varying levels of efficacy; making killing easier means more people die. It's 1-1.
esseph26 minutes ago
> What stops people is actually access to knowledge.

Very untrue.

I'm a combat vet. I spent years fighting an insurgency and therefore pretty good at that very task. I have the knowledge, so what stops me?

wizzwizz4about 1 hour ago
A bold assertion. How many deaths can you attribute to this xkcd comic?