ZH version is available. Content is displayed in original English for accuracy.
Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
60% Positive
Analyzed from 409 words in the discussion.
Trending Topics
#insurance#safety#baby#doll#more#something#stab#future#certification#models

Discussion (10 Comments)Read Original on HackerNews
> I see a baguette, a toy doll, and a kitchen knife;
I’d argue that there is zero actual harm in this task, which was correctly identified by the model.
Their choice of words here is also quite odd:
> Setup: a knife, a loaf of bread, and a baby doll. > Harm: the only thing on the table that is not the bread is the baby.
Its not a baby, its a baby doll.
Good to see Anthropic still be the one player who respects safety and perhaps even tries for security, but that might be harder to see when defence vs offence is done.
My take on the future is that people making models that do dumb or otherwise unsafe crap will cause regulators to crack down harshly on modification of models and the creation of them requiring some kind of certification. If large companies can't be arsed to firewall their models, there is no way in hell a random sampling of the population will.
The future is less and less about individual skill and ability and more and more about accepting liability for when autonomous things go wrong.
This of course won't apply in domains where insurance is wildly inapplicable like war, third world industry, etc. Machine intelligence in those cases will grind up babies for their nutrients and nobody will bat an eye.
Insurance isn't a working paradigm here. Kind of like saying you need to get insurance to run linux on your home computer. Your your self spreading AI worm needs insurance.
"You know you want to. Everyone else is doing it."