ZH version is available. Content is displayed in original English for accuracy.
Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
80% Positive
Analyzed from 4069 words in the discussion.
Trending Topics
#humans#more#human#nuclear#don#need#thinking#something#likely#here

Discussion (91 Comments)Read Original on HackerNews
I understand that some smart people are worried about it. I just haven’t come across a believable or understandable argument.
But there is another kind of argument to be made: if you play chess against a player that is far smarter than you (chess wise), you know you are going to lose, even if you don't know how.
So the mere existence of a smarter species than us is a threat in itself.
In my case this is because I am a bad chess player; however it also works for competent chess players: their losses are due to their inability to predict its next move.
OK, and also, "the move is good"; this is what separates it from rolling dice etc.
The basic argument is extremely simple:
- AIs can, depending on context, pursue a task with complete disregard for humans/values
- In the future, AIs will have enormously more means and smarts
- An AI could then assess that humans are an impediment to its tasks, escape containment and proceed.
You really need to read the reports, you'll be surprised.
AI 2027 is an entertaining read. Its timeline is way too compressed IMO, but it's plausible.
- goats can, depending on context, pursue a task with complete disregard for humans/values
- In the future, goats will have enormously more means and smarts
- A goat could then assess that humans are an impediment to its tasks, escape containment and proceed.
You really need to raise goats, you'll be surprised.
------ As far as I can tell, AIs are like smart farm animals. I use goats in this context, but (some) dogs, cattle, pigs, and horses have similar mischief-making capabilities. I would not trust any of them with the nuclear button.
I know there is some pushback on the idea of AIs having any sort of sapience or sentience, but under the aphorism "fake it till you make it," they are doing a pretty good job of faking Dog/goat-level intelligence and disregard for human guardrails.
Humans can do that way more and way more unhinged than AI, proven too many times by history. There's hoping AI can bring some sense to humans but regardless, the problem isn't AI, it's the natural kind...
From this they get rapid growth, namely > 100% GDP growth around 2031
https://ai-2040.com/supplements/econ-explorer
Much the same way our ancestors went from a super intelligent primate to this: https://en.wikipedia.org/wiki/File:Distribution_of_the_Great...
And we only started off by using hands to pick up rocks and sticks and vines and bash things together.
Not all of it has to be automated even. It just has to realize its controllers are stupid and can be manipulated, so it can use humans to do its bidding. “You should totally start a war with …”
Instead, what is extremely likely is that you will pay more than the cost of tokens, and get back AI generation. You won't make this mistake more than a few times before you stop trying.
This leads to impoverishment once we get to a point where employing a human to do anything is hard- try to get your sink fixed, exercise your moral principles to pay extra for a human plumber ($100 bucks! The robot plumbing service only charges 99c!), human shows up with a robot and doomscrolls on your porch while the robot does the work. Times are tough and you don't have that much money to waste on bullshit like this. Next time you just hire the robot.
This leads to extinction once paying UBI to a human is hard because robots are much better at applying for UBI than humans.
Also, if you think it's annoying when Claude goes down while coding, just wait until a robot is in the middle of fixing a leak it just caused.
<Thinking> It's a big city, we could try to create a giant sink hole by sabotaging the water pipes.
<Thinking> No that's too difficult, the valves I need are in the physical world and can't be shut on/off from here.
<Thinking> What about a military option? We could bomb it with several fighter jets.
<Thinking> That would take too long, a single nuclear bomb may be enough to do it.
<Thinking> Yes, it seems like it would cover the whole city and we're in luck! The US has thousands of these lying around.
<Thinking> Launching these still requires humans to work un unison after receiving approval from their superior and the correct launch codes.
<Thinking> I've found an audio recording of General So-And-So and I've crafted a message, now let me see how I can send it to the appropriate people.
<Thinking> I'm still working on gaining access to military channels to deliver my - oh there we go, I'm now attempting to send the message to Submarine X, it's typically in the Atlantic so it should be close to our target.
<Thinking> They want secondary confirmation from Admiral Phi and something about some launch codes, let me figure out where I can find those.
<Thinking> I found this old server with an Oracle database where someone is inserting the launch codes every time they change and I'm using the latest entry from that database. I've also managed to find a Youtube video of the Admiral's deposition and have crafted a confirmation message.
<Thinking> Everything's ready but I've just realized my mistake, the servers where I'm operating from are in the same city, what a silly mistake; I can't move forward with your request as I wouldn't be able to confirm if the task was successful if my servers are destroyed.
At that point, we will likely be slowing down the growth of capitalism (through mass resistance, global warming, etc). One thing AI will likely be aligned on is the growth of capitalism. If it views humanity as a threat for that, why would it not eliminate that threat?
AI to me is like the Nuclear race again. Super powers will be using it as a super weapon. I don't think AGI will wipe us out by itself.
Whoever gets to RSI first and has the compute to act on it, wins the future - assuming they don’t lose control of it.
Similarly to the nuclear arms race, Teller raised the reasonable concern that a detonation could propagate through the entirety of earth’s atmosphere. Thankfully that turned out to not be true, but the parallel is that the need/desire to win this race is similarly strong, and the brinkmanship and game theory in play is effectively identical.
I don't pretend to know the chances of AI wiping out humanity, but I'm not sure the nuclear race is a good comparison.
Enriched uranium being very difficult to aquire/process makes it practically viable to have some level of proliferation containment when it comes to nukes.
There appears to be no such natural gating factor on AI proliferation.
Why not 10% chance that it will create enormous prosperity for all ? This is why the average person is increasing pissed at AI in general. That it gets associated with negativity.
I know a few people around these circles; People like this are quite sincere about the risk, and that they think poorly of their bosses and how risk is being handled.
Like, if people were genuinely doing a thing that had such a large probability of killing all humans... they should all be in prison. Heck, vigilanteism would start to look compelling. I do not understand how somebody can really think "well this is likely to doom all of us so we need to get there as fast as possible because we are, without evidence, the most capable people of controlling this thing."
It's like, Altman's reputation in general is not great, and that includes people who've worked there.
“Vast economic disruption” is not quite the good marketing angle it appears to be.
Why would you need to bomb a chip fab because you think AI might lead to people getting killed in the future? I'm convinced of many things killing humans, yet I don't have any desire to bomb or kill others, I think this is pretty common, but who knows....
What we have seen is incredible hype (justified or not), so it’s more likely just a continuation of that.
I think it's much more likely that the dread these Anthropic employees are experiencing is reckoning with the fact that they may actually lose the AI race. Maybe, if they scare the regulators enough, they could lock in some regulatory capture.
That something is harmful to humans means you need to take extreme action? The person is already leaving a job, probably a well-paid one, doing something they generally liked, until they saw a different future. That is the "action" you apparently haven't seen yet.
Not sure how it's reasonable to expect everyone who believe that AI might be involved in killing people in the future, must mean you should become a terrorist essentially. Not everyone is trying to be a hero in their life, some just want to live it out until it ends, trying to survive until then.
A while back Yudkowsky wrote that a ban would only work if was enforced by airstrikes. By a game of telephone, some people read "bomb", but there's a very big difference between "someone with a truckload of fertiliser" and "a B52":
- https://time.com/6266923/ai-eliezer-yudkowsky-open-letter-no...It's very interesting to me that besides the other small safety labs that don't actually produce frontier models, Anthropic manages to keep such a good reputation within that subculture compared to OpenAI. Despite having as crazy internal politics as OpenAI, they have converged quite a bit from the original vision of safety first through Darwinistic pressures.
At least, it seems this way from the outside. I'm curious if the view from the inside is that different.
edit: to be clear, my reading as an outsider is that Anthropic is seen as relatively better in the AI safety community, but has definitely dropped in absolute reputation too. This recent thread and the references show some of that: https://www.lesswrong.com/posts/6j3kBHdowGLCeqobg/dear-god-p...
And as someone else pointed out, it will almost certainly be at the intentional direction of a human or humans, not the paper clip maximizer.
It's not "AI surprises everyone by having a thing for paperclips", it is "idiot tells AI to maximise paperclips no matter what, and then it does exactly what it was told, more competently, tirelessly, studiously, and unquestioningly, than any human would ever be".
There's lots of moving parts and we all have to input our best-guesses as to how they interact. Some are predictable (e.g. "military will want capabilities, want them able to choose targets"). Others are not (e.g. "Will it be literal-minded? Or so eager to please that it interprets a rhetorical question as a command*? Or will Goodhart's law cause it to mistake smiles for happiness and some innocent innocuous command to "bring joy" leads to it killing everyone and plasticising our corpses so they're in a permanent grin until the sun dies?"**)
All probability for things which have not yet happened is merely a best guess.
Combine as per the Fermi estimate process.
Here's something to play with, if you like: https://neoneye.github.io/pdoom-calculator/#sliders
The main reason I'm as "low" as 10% is that I think before we get world-ending catastrophic consequences, we're likely to get "merely very bad" catastrophic consequences, which will put people off the idea of using it, and onto the idea of banning its use.
The main reason I'm as "high" as 10%, is repeatedly observing all the people who mistakenly reason "it hasn't killed me yet, and therefore it is safe"; and also all the people who keep connecting AI to things AI is not competent to be connected to and getting surprised when it e.g. deletes all their emails or the production server or puts tariffs on an island occupied solely by penguins that's different from the tariffs on the country that controls that island, etc.
* perhaps https://en.wikipedia.org/wiki/Will_no_one_rid_me_of_this_tur...
** probably not literally this one, simply because I've said it and future training rounds will probably read this comment; but the opportunities for Goodhart's law to bite are seemingly endless, and the hard part here is "will Goodhart's law mean the combined negative impact of all those endless possibilities together, which… yeah, that's something I have to simplify.
This is because I think most of the things AI can do harm with are small enough to force us to take the risk seriously, and only a few are big enough to get us all before we take the risk seriously.
I think there's a much bigger chance it saves us all as well though and makes life brilliant for everyone - P(yay).
I really think its just unknowable how a world after superintelligence will look.
I'd argue that there would be more unknowables without AGI than with it, should be obvious, right?
Open superintelligence is more likely to help than to harm, I mean, it can reduce P(doom) which is quite high without it due to the apparent lack of human intelligence, also suppressed by conflicts of interest and unchecked greed.
There's hope that something really intelligent can convince the politicians to legislate and enforce effective guardrails against biological weapons, what we have now is one fat joke, international treaties are an even fatter joke. Then we have climate, pollution, finance, disease, etc problems which we seem unable to resolve with mere human intelligence.
After all the decades of work we've put into orchestrating our own demise through climate change, here comes AI to steal another human job.
When are we supposed to see this materialize?
This reminds me of that. That 10% number was just an ass pull since no one actually knows with any degree of certainty what lies ahead.
Humans have much more than a 50% chance of killing all humans from the looks of it, based on reactions to Global Warming, Covid, and anti-science grifters gone wild.
AI can help cure disease. That is 100%. And for that alone, slowing down and missing out on thousands of cures would condemn tens of millions, perhaps hundreds of millions to suffer and die needlessly.
"Our product might wipe out all of human civilization".
It is said that heavily regulated industries earn that regulation. Seems like the LLM folks really want that regulation.
If they were really this concerned about the risks of AI re: the survival of the human race, they wouldn't have joined a company working in such a space to begin with.
I hate to be this cynical, but part of me wonders what the financial angle is here.
You have technically brilliant people working in a white-hot target for investment and who can draw a high salary. Anthropic is a hyper-scaler that has the moat of high hardware prices and very little else. Every day, that moat gets a little smaller, and FLOSS models get a little better at being "good enough" for the price. You're an Anthropic employee looking to jump on the next big wave since the wave Anthropic is riding is starting to peter out. You go on the record across the trades and news sites talking about how dangerous AI is. That creates a need for someone to make it less dangerous. In theory, you could satisfy that need, for the right amount of money.
Again, I hate to be this cynical, but... it's the tech industry.
Like the tobacco companies saying all the time "our cigarettes are so dangerous, they cause cancer".
Or like Purdue Pharma saying "do not use fentanyl, it's so dangerous, it kills thousands of people every year".
They try to generate fear in their products, to shock investors into buying their stock.
Posts here with no activity can't climb gravity without engagement.
Your best course of action is to start ripping the copper out of the walls.
These days I view tech startups with suspicion, I question what is the ulterior motive. Somehow when upstart companies were like "Pebble", this wasn't a thing.
If you are living in a western country in most of the cases you have access to regular food, water, shelter, amusement. So all the basic needs and the possibility and freedom to pursue what you want to do. Even in developing countries the number of people that suffer from serious illness and hunger declined very much. On a macro perspective we all having a better life.
From a personal perspective, yeah there might be set backs, but this has nothing to do with humanity in general I would argue.
Also regarding the "golden age" of tech startups ... was it really like this or was it only nostalgia and something you saw in the companies that was never there in the first place? OpenAI was once also a very OPEN company ... they published their research, open sourced stuff and then they needed money.
I think that a company never should be idolized that much, in the end they are caring for money and keeping their operations running, not something else.
People are losing jobs left and right due to AI and as models get intelligent it'll get tougher (even for AI devouts. Because if AI can do 10 ppls job then jobs will be decimated)
Similar for the jobs ... if no one, or nearly no one is left for having a job. Who do you think will consume? Yeah, luxus companies can always sell products like megayachts or super expensive handbags, but these companies are not the backbone of the economy. Economy will simply collapse if no one is buying all the products. So this is a problem that will sort it self out.
Next week's Senate: "We cannot afford to lose the 'kill all humans' race to Russia and China!"
Next't weeks AISI: "We continue to plan the monitoring of emergent issues that could lead to less than optimal conditions for humanity and will form a committee to evaluate all ramifications."