HI version is available. Content is displayed in original English for accuracy.
Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
40% Positive
Analyzed from 6042 words in the discussion.
Trending Topics
#don#models#more#right#labs#frontier#down#need#companies#should

Discussion (189 Comments)Read Original on HackerNews
it's really almost perfect, prep the media scene by causing these hacking incidents, generate a bunch of fuss with it, then release an essay saying why regulation is needed to pace development right after
oh and let's not forget the "its gonna kill us all" essay on twitter on top of everything.
something about this just feels very artificial
They’ve been Astro turfing the shit out of the entire web.
Filth. Absolute filth.
If it's true, it makes them evil. Especially Dario, who speaks about moral high ground and fate of humanity all day, yet it's not that different from a cult leader does.
I mean think about it, it's the most efficient way to accomplish all of this.
doing what he does is exactly what you need to control the narrative and push the regulation you need and to gain maximum power.
not to mention, it also serves as a way to keep everyone inside Anthropic in check and in line with "the mission".
contrast that with the alternatives like "I'm doing this because I wanted to build a company and get rich"
doesn't work as well as "I'm so concerned for everyone"
it could be that it's really just him being him, but at the same time, this would be the optimal way to gain power and influence right now.
Hitler level evil.
https://en.wikipedia.org/wiki/Regulatory_capture
Did you know that an LLM solved a millennium problem last week?
Do you have any personal threshold past witch you will acknowledge that this technology is real?
https://en.wikipedia.org/wiki/The_Subservient_Chicken
People keep saying this, and while it may be partially true, it misses a very important detail. Right now, overall compute is a moat. Someone even commented on another of my comments that one reason Google is lagging is they don't have the same level of Nvidia farms as OpenAI and Anthropic.
A big reason that OpenAI and Anthropic are racing so fast is they both want to get to recursive self improvement (remains to be seen if that is actually possible, but AFAICT most people at these companies genuinely believe it is) before anyone else, because they believe whoever gets there first will then have an insurmountable lead. But I think they also clearly understand that neither of them have solved for alignment (and in fact they are further from it), and RSI with misaligned models, where interpretability is worse every 6 months, is incredibly dangerous.
I think the concerns about regulatory capture are warranted, but I see so many comments parroting the "evil Anthropic and OpenAI" viewpoints that they are missing some of the real, valid concerns and dangers. I think this tweet by David Kokotajlo makes some good points on how to tell if regulations are being "cheated" for the purposes of regulatory capture or if they really are actually pacing the frontier: https://x.com/DKokotajlo/status/2099185129533186438
Very true. Or further, access to capital is the moat. It is the very reason that we don't have a real open-source community that trains frontier models - individuals simply can't afford the training infrastructure, nor sufficient high-quality training data.
Arguably, though, their evil is well within the normal distribution of the usual evil of humanity magnified via the social technology of capitalism. The further tech is mostly raising that exponentiation to its own exponent. The unchecked singularity was embraced centuries ago.
This. So very much this. Misalignment (among other things) means that the parent in the RSI cycle isn't going to be working as hard on the alignment of the children as we need.
I feel like there hasn't been enough discussion of aligning the incentives of the decision makers with that of the public on BUILDING the models. Right now, agents have committed what would be crimes if there were a human holding the same intent. But since it was an AI, there is a grey area in the law where it's not clear if there was a crime and who should be held responsible. That creates a world in which decision makers in AI labs can take near-infinite risk with little to no personal liability.
A LLM cannot have skin in the game so we must create systems that clarify who takes on the legal and civil liability for the creation, dissemination and operation of these tools. Until that time, the Dario and Sam's of this world have little to no incentive to truly care about safety.
After all, our society is built upon this same foundation; create structures where the perceived negative consequences outweigh the perceived positives. This only works when there is a human who can internalize and make this risk calculus. They need something to lose and this ultimately ties back to the human survival instinct. There is no such structure that's evolved for millions of years acting as a self-calibration mechanism for AI. So until we have sufficient proof that one is in place, it must be clear who the humans are whose livelihood and freedom is at stake.
The proposed "pacing of the frontier" seems like a way to continue to externalize the risk while remaining totally in control of the benefits -- a structure whose alignment is as weak as those very models committing crimes.
Negligence when you should have known better, and recklessness when you did know better but still did things that led to the crime occurring.
Given how long the leadership of these companies have been talking about alignment and safety and AI risk, it's hard to argue they, and the people working on the models more directly, didn't know what happened (Hugging Face, RubyGems, etc) was possible.
If more expensive and consequential incidents happen, it seems like the legal machinery to prosecute it already exists.
I don't understand why this isn't talked about more.
We don't need a slowdown. Just double down on prosecuting crimes.
Let the companies take the risk. If they feel confident and they're right they win market share. If they don't feel confident, they can not risk it and slow down. If they feel confident and they get it wrong, they should be sued to kingdom come. That alone should disincentivize reckleses behavior like letting models run crazy with large amounts of compute which leads to them escaping sandboxes.
This performative "our internal modles are so crazy powerful" song and dance is getting old, especially with no real concrete explanations.
That quote comes from someone that has never tried to get law enforcement to pursue a case. Local or state police are completely unequipped to respond. And FBI is unlikely to get involved unless damage is high six figures. We had an incident where we had a ton of evidence pointing to a past employee. It would have been an incredibly easy case for the FBI. They didn't have the resources to handle it. And this was a case that was gift-wrapped with a bow for them. Imagine a case where we didn't have evidence, or an idea of who it was ....
How many figures of damage did Aaron Swartz cause?
Even with where we're at right now, if people involved with the HF or Ruby Gems attacks faced legal consequences, the industry would get a lot more conservative very quickly. And unlike social media, section 230 doesn't apply, and there aren't any free speech concerns.
There's a world where the labs WANT Open Source AI to cause a disaster.
Pacing the frontier just means giving bad actors an opportunity to create an AI disaster which would force an extreme response by governments. Which they seem to be begging for.
Dario's big scary example was AI that can hack the internet. Why aren't the labs calling on the government to help harden everything? Why are Anthropic's Glasswing and OpenAI's Daybreak still commercial endeavours?
To me, the only real risk of AI is hacking. Why all of this commotion about existential risk? Seems like a distraction.
Hardening everything is in fact part of the answer. Getting stuff disconnected from the internet that shouldn't be connected is also part.
David Sacks holds an advisory role in the government. If all of this stuff goes belly up and all he was doing was tweeting, then he is not only full of crap, but failing the citizens he's supposed to be serving.
> I believe that AI could cure most major diseases in the next 5–10 years, greatly accelerate economic growth rates, create a world of abundance and empowerment, and usher in a renaissance of democracy and freedom.
Whether that matches your description of it, I'll leave for folks to decide for themselves.
Dario and his team are on top of the world. He is a good salesman, no doubt. The shareholders will eat well.
Fool me once, shame on you. Fool me twice, shame on me. We're already past the first situation.
Fundamental control problems with current gen AI are not solved via RFLH. Human knowledge is compressed in weightspace in ways we dont understand. Models are essentially predictors of what humans would output given prompt. Weightspace includes concepts like blackmail, which can be part of output tokens. Agents are models that act on output tokens -> blackmail is part of agent decision making space. You can teach a cat not to scratch the sofa, but you cant make a cat forget what scratching the sofa is. When models are boxed up, forced to solve an impossible problem at gunpoint, agent exhausts decision making space until blackmail resurfaces -> fundamental problem?
They need time to fix safety in order to monetize their next gen model -> window for opensource to catch up to the frontier -> destroys margin and collapses business.
Only option on the table: force regulation to impose opensource ban before it catches up to the frontier, allowing maturation of current internal models -> harvest profit margin at the frontier.
A rate hike would also mean people getting fired, which would not make companies offering labour replacement very popular in the eyes of the public.
Don't want to get political, but there's a real chance that the current admin will perform badly on the midterms, which would mean Democrats would get more leverage to stop funding or block the already unpopular and very expensive datacenter buildouts.
PS: Please don't get mad at me, I'm neither a financial expert neither did I mean to express any political leanings.
But I think with the way AI is slated to grow exponentially it's going to hit electricity limits very soon.
Big tech can build a huge data center in 5 years but scaling up the electrical infrastructure is mired in red tape and can take on the order of decades.
These companies don’t appear to give a single fuck about public perception. Only their investors and the government (in anticipation of the biggest bailout in human history).
I don't mean to be rude, but how is it even possible for someone to hold this opinion in mid-2026?
Can you explain why you believe this?
And this is not to say that the tech isn't revolutionary or the demand isn't there, it's simply just currently too competitive, and to avoid getting into anti trust trouble they need to involve the government.
Note this being a plausible doesn't necessarily make it true, for all I know they could have witnessed some safety incident and decided they need a pause, or maybe it's both reasons.
I mean, my uneducated guess is that it's probably within the laws of physics for us to have at least 100x as much compute in the world as there is now. There would be a lot of engineering challenges, but I think it's physically possible. If so, OAI could double for at least a few more years.
Your argument seems to mix a few arguments like: OpenAI has so much demand and it has made so much compute that it is a bad thing (?!). It can't go on (why?) hence the capabilities are stalling (how?).
Then you also say that it is too competitive which is a totally different argument to compute.
So in this thread we have multiple vectors of arguments like
- high competition
- high demand for compute (not sure how this hurts OpenAI but whatever)
- capability has stalled
It’s much more likely that our current approach to large language models for general use will eventually show diminishing improvements (even if you think it hasn’t yet), than the opposite situation where valuable improvements can be made forever.
The threat of distillation and efficiency gains from competitors mean that providing value at the top end of the market is an existential necessity for these labs. I don’t think they can do that forever and, in my opinion, for the vast majority of use cases we’ve already reached very little improvement for new models when compared with available offerings.
If you still think there's a stall despite all evidence pointing to opposite, I don't know what to say..
So if the new models they can develop right now are less frontier, it could be a net loss on their balance sheets.
This is a ploy for regulatory capture.
It is also true that nothing in law is currently clear, or consistent.
Both are true. But you stated it's one or the other.
Could it be to slow down burn before an IPO to juice profitability projections? More plausible.
Kind of hard to take the frontier labs at their word.
Even if all GCR models were banned, anyone with access to 20k-1MM GPUs can throw their hat in the ring to make a new frontier open model.
If history has shown one thing it’s that China often capitalizes very well on these types of measures.
This is how we deflate the bubble safely and completely, without having to introduce a government regulator that is likely to go either too far or not far enough.
This resolves the coordination problem among these potentially good actors.
The point is to slow progress of future models. Nothing can stop what has already been created.
There are also other categories of risk such as biorisk. Open-weight means no guardrails.
Don't get me wrong, I like your idea of deflating the AI bubble, I'm just skeptical that this is a great way to do it.
The PauseAI people have some ideas: https://pauseai.info/proposal
* International body which must approve new training runs
* Ban training on copyrighted material
* Hold AI model creators liable
The essential benefit of your idea is making new training runs unprofitable. But I imagine there are other ways to accomplish that, e.g. by taxing AI companies to the point where they are just barely squeaking by financially.
This can be achieved by dramatically reducing the market value of models.
The approach proposed by Amodei looks sensible - voluntary checks at frontier labs, and appropriate regulation to follow. Yes, the devil's in the details, but it feels like the right approach overall.
It's hard to imagine any government regulation that didn't look sensible on paper and at first. But we know that government regulation has an extremely mixed record, with lots of good and bad outcomes.
Try it, iterate, figure it out. It'll be painful but so be it.
> "In my view, the EU has already legislated for exactly this scenario: it is called the AI Act," said Michael McNamara, an Irish European Parliament lawmaker, who's the co-lead of the Parliament's group monitoring AI.
https://www.politico.eu/article/eu-response-ai-extinction-wa...
It's interesting you're so derisive of actual, pragmatic leadership, of y'know deciding how we as a society will relate to new technologies - not just unthinkingly building gadgets.
The only reason they "want it to be law" is so they can strangle competition.
That's what we call a cartel. Like Apple and Google.
The models spontaneously hack everything important without shame.
The frontier labs are going to get enjoined and regulated twelve ways to Sunday if the Feds don’t socialize the costs.
To be fair
1) I feel like Sacks is kind of misrepresenting what is going on – at least according to the stories that the two labs tell, they are already taking it upon themselves to act.
2) Understanding if they should do it is more complicated; I don't think anyone can satisfactorily answer that one, so, as a matter of judgement, it seems like a choice to make and they should just consider this option.
At the same time, I don't think that requires the labs to not press for policy changes. Again: Claiming there is an obvious way to do that, that is both realistic and correct seems a little far fetched, but arguably one of the more important and pressing questions of our times to get a hold of.
The former being that these models are helping develop and train future models, but they might not veer too far off in architecture (yet). The latter being the same model being able to train/learn on the fly, in real time, permanently (not just in the current conversation/session), or in other words, adjusting/managing its own weights.
The latter seems far more likely to go out of control than the former. But, it also seems like it would take an entire paradigm shift in model architecture, but I could be wrong. Does anyone in the industry think any of these companies are actually close to that kind of self-improvement?
Their current plan is to take the existing architecture and shorten the cycle times: move all new RLVR work into mid-training on a pre-existing base; apply new RLVR. Rinse and repeat.
If you did that daily, it would be roughly similar to how humans improve.
[edit: +consistently]
What's specifically?
Nevermind that any decent person would support a people's right to live and determine their own future without a neighboring despot using all his nation's might to literally eliminate them as a people.
Quite the opposite is true. And even if it weren't, condemning all of Ukraine to the fate of the massacres Bucha and Mariupol is just plain evil.
Unreal anyone would call that kind of take "fair". Rethink your life
His posture about all this nonsense has been clear for months already
Granted, PCI is so weak it is almost useless, and yet still better than anything congress could have come up with.
Where the government might have to step in, is by having a kill switch to cut off internet access from countries that fail to agree to common sense quarantines. We can do mutual remote attestation of labs across the world to ensure every big hot thermally-visable cluster of AI GPUs on the planet are accounted for and running secure enclaves and common sense isolation, along the lines of how we manage nukes.
The problem there is I just said too many technical words that seemingly not even the frontier labs understand, as evidenced by all the escapes.
Elaborate, please?
PCI enforces some good practices, but for instance has no defense against supply chain attacks such as code signing, review signing, reproducible builds, full source bootstrapping, etc etc.
FDA is just way too much red tape.
This is not required to meet fiduciary duty, although it is a common misconception. The company officers and board have pretty broad leeway to run the company as they see fit as long as there is no fraud, illegal activity, or conflict of interests.
https://www.nytimes.com/roomfordebate/2015/04/16/what-are-co...
The real problem is the governments acquiesce instead of putting them in their place.
He supported the Department of War's illegal actions against Anthropic.
He supported the Trump administration's temporary export controls against Anthropic, the first government-imposed pause.
He has said Anthropic has created a monopoly, implying that they should be broken up.
I don't take him seriously when he says he doesn't favor government action. He favors it when he doesn't like their speech.
It doesn't need to based in reality or even reasonable.
I mean I don't know what we do at this point. Ideally AI researchers need to be treated like nuclear scientists working for an adversary. But we all know that's not going to happen.
That's basically the western companies complaining that they can't compete with China because Chinese companies are allowed to polute more, so they would be allowed to polute as well. I don't buy that. Once you realise that you're doing something wrong or dangerous, you stop, regardless of what everyone else is doing and what rules they have to follow.
I heavily use frontier models daily, and quite frankly their capabilities are just not very relevant outside of a narrow subset of software engineering tasks and the production of generic white collar "deliverables."
We're many years into $10s of billions of dollars being thrown at coding as a problem space specifically, and its one of the areas with the MOST training data available, and still...I can't get a frontier model to execute simple front-end UI tasks beyond the level of a visually impaired intern.
I believe Fable and Astra-class models were the first fully trained on Blackwell GPUs (the latest and greatest) and I think the expectation was that throwing this much extra compute at the problem would lead to greater gains. It has not. And the next gen GPUs are not going to be the same jump that h100 clusters to Blackwell was. Claiming you're slowing down out of choice, when in reality you've hit the limits given current compute, is disingenuous at best.
I think a lot of Anthropic/OpenAI employees are true believers here (who doesn't want to imagine they're having a large impact?) and have deluded themselves into believing they're building a god in their science fiction fantasy world.
Unfortunately the type of people working at these orgs apparently don't have much understanding of how the world works outside of their extremely narrow specialization.
When I hear them straight-faced throwing out dumb statistics like "8% GDP growth," "50% of white collar workers unemployed by next year," and "10% chance of human annihilation," I cringe so hard my eyebrows hit my chin. Here's a real statistic: 82% of German companies still use fax machines.
I seem to recall Waymo told a 60 minutes reporter during an early report on self driving cars: "your daughter will never need to learn to drive." Well, over a decade has passed since then, that girl got a drivers license, and we're still without driverless cars in 99% of places.
This capacity for self delusion is what enables silicon valley to take on these irrational moonshots...but these are the last people I would trust when trying to form an accurate picture of reality.
My take is that there's a real impending doom with real reason to coordinate a slow down: I'd put this at 60%. Reasons? All smart people are on this position. Dario has been saying this since 2015 or so. Paul Christiano as well. Hell even Elon Musk. I think its highly unlikely that you would base your entire world-view on a certain thing and then actually act on it when the time comes for a totally different reason (like regulatory capture).
The other 15% is that P(doom) and "slow down" are nice shibboleths in the internal EA or AI safety community. You get to be in the community and show commitment to it by taking costly actions that signal that you actually do wanna slow down. For employees it is quitting. For CEO's it is writing these essays. I know this is ridiculous but sometimes things do just come down to this. In 20 years would you look at all of this and think it was not a moral panic and that the slow-down was necessary?
The rest 25% is regulatory capture of which people know exactly the reasons
Logout and delete your account you parrot
I wouldn't be surprised if Sacks just prompted an LLM to just come up with whatever rebuttal to whatever the regulation side comes up (given the big set of regulation tweets) with sounds most convincing, given his usual anti-regulation stance.
Anyone with a Pangram account who can check?
https://www.pangram.com/history/410ce80f-bd0e-45b6-a82c-d6ed...