HI version is available. Content is displayed in original English for accuracy.
Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
45% Positive
Analyzed from 1177 words in the discussion.
Trending Topics
#training#data#using#title#fired#contractors#labeling#human#rsi#why

Discussion (46 Comments)Read Original on HackerNews
This title says why. The article then goes on to explain why that's a problem. Pretty straightforward.
"People training OAI AI" has a colloquial meaning of "People working at OAI"
So the headline is built to imply:
"People who work for OAI got fired for using AI to do their job"
When a non-clickbait headline would be
"3rd party workers contracted to distill their human knowledge fired for using AI instead"
Not nearly as "clicky"
1. That system isn't broken in a difficult to tell way.
2. The systems implementation is incorrect.
3. That the theory that particular implementation uses isn't incorrect.
There is nothing hypocritical at all about that.
Is your objection to the word “train” in headline? Otherwise, you’re just restating the headline while calling it clickbait (which, btw, it’s not, perhaps you meant to say it is misleading, which is a different thing.)
Its the same difference between "I was arrested for having liquid in my car while driving" and "I was arrested for holding an open bottle of whiskey while driving"
I don't know if it's deliberately misleading or if the journalist doesn't understand the difference, but the output is functionally the same.
I've been told on HN about a year ago that the era of scraping is over and that it's all AI-training-AI now. My web server logs and stories like that disagree.
They hired contractors on the condition that they provide human feedback without AI; those people broke the rules, so their contracts ended prematurely.
Not sure why this is news uncovered by an investigative journalist.
What is it, exactly, that makes AI-generated text so poisonous to AIs but totally harmless to humans? What is mode/model collapse, and why can it only happen to AIs with too much AI text in their training data and not to, say, human students with too much AI text in their textbooks? The people who know the most about this phenomenon seem much more careful about contamination than they are encouraging us to be.
(I’m not trying to be disingenuous, though I am trying to express a niche viewpoint. Hopefully, my willingness to take the metaphorical role of conspiracy theorist is understood as epistemic humility.)
The more you think about it, the worse it gets.