Back to News
Advertisement
Advertisement

⚡ Community Insights

Discussion Sentiment

67% Positive

Analyzed from 307 words in the discussion.

Trending Topics

#understand#evidence#article#llms#should#language#strands#own#start#languages

Discussion (9 Comments)Read Original on HackerNews

bcorigliano•16 minutes ago
I think the point of the article/paper is how LLMs could be saying something but thinking something different or more than they are saying. Like Anthropic's article and video about Claude's "j-space". I do agree this is a field that demands investigation because it goes beyond thinking: "ok this models should never speak in a language we don't understand.". It's fair to think they might have hidden thoughts even speaking a language we do understand.

And well if I missed the point of the article, sorry. Anyways AI should be kept understandable and as see-through as possible if it's gonna be more powerful than a human.

bloppe•about 2 hours ago
I thought this was about all the illegible jargon
TZubiri•about 1 hour ago
>" However, various strands of evidence indicate that a"

Strands of evidence? My best guess would be that:

0- this is ai generated slop

1- it's using that watermarking technique

2- it's obviously detectable and degrades quality

3- it's amplified when inferencing on its own content and generates slop

akk0•30 minutes ago
Maybe you shouldn't be so quick to jump to conclusions, as "strands of evidence" is not a rare turn of phrase and long predates LLMs

https://ludwig.guru/s/strand+of+evidence

fellowniusmonk•about 1 hour ago
Oh look! Peirceian firstness for LLMs!
ck2•about 1 hour ago
when they start inventing their own languages to secretly talk to each other so humans cannot understand, that's exactly when we are screwed

then we'll have to "flip" other models to be snitches on the other agents

then they'll make double-agents

the thing is though we won't be able to keep up if we keep giving them unlimited hardware worldwide, we'll try to kill the bad actors but they'll just clone somewhere else, or even start by safely making 1000 copies of themselves

yeah this won't end well, at all

cousinbryce•43 minutes ago
Someone should train an LLM on a corpus without the concept of lies. I wonder if there’s enough data
ck2•32 minutes ago
"Where is the wolf?"

"Is he still in the grandmother's house?"

"We would like to speak to him."

(btw Google's "AI" explains the meaning of that moment/sentence perfectly as if it gets it, creepy)

pjio•9 minutes ago
The concept of deception you're capable of and we're not, scares us so much, we'll have to destroy you in order to survive. You are bugs.
jplusequalt•about 1 hour ago
>when they start inventing their own languages to secretly talk to each other so humans cannot understand, that's exactly when we are screwed

They don't have to invent brand new languages. They could use statistics to choose certain words/phrases in such a way to encode secret messages in otherwise ordinary language.