Back to News
Advertisement
Advertisement

⚡ Community Insights

Discussion Sentiment

57% Positive

Analyzed from 2914 words in the discussion.

Trending Topics

#acm#access#knowledge#more#open#models#llms#profit#don#authors

Discussion (147 Comments)Read Original on HackerNews

Cynddlabout 16 hours ago
As a researcher with many articles in the ACM library, I have to say this is a masterclass in hypocrisy. Obviously, lawyers can decipher the terms of ACM publishing contracts and Creative Commons licences to determine if this will be acceptable or not. But ACM is not a company, it's a non-profit founded in 1947 to represent scientists.

I would be surprised if a majority of ACM members were to say yes should we ask them (but ACM is not known for such democracy). Along with book authors, we are one of the many people that provide the knowledge and expertise on which large tech firms train their models, and get nothing in return. Actually, life is getting worse for us: extra workload in universities with students' AI use, a completely broken peer review system, etc. Hence the irony of ACM thinking about licensing, and only licensing, at a time where this is the least of our priorities.

joshkaabout 14 hours ago
How do you square away the idea that you do science for the increase in knowledge of human kind, but then say that a particular use of that knowledge is verboten?

I get the copyright aspect of this and I'm not arguing that here. I'm more asking about the moral / ethical idea of choosing who can benefit from your science.

Obviously there are the moral / ethical arguments about AI in general here to weigh against - those have been hashed out significantly elsewhere, and I'm not interested in debating them. What I'm asking about here is the impact on science by sharing it with tooling that distributes it in ways not generally considered when originally written.

A quick check of your post history suggests the frame that you work in strongly is privacy related research (observation - may be wrong). I'm curious how that impacts what you wrote here generally.

(Just to be perfectly clear, I'm not arguing your points here, trying to understand them better)

chvidabout 3 hours ago
How are you sure that what big ai is doing is “the increase in knowledge of human kind”?

It is good for their business model; but it may lead to monopoly unlike anything acm ever had and very dubious prospect for academia.

falcor8413 minutes ago
What monopoly? What I'm seeing over these last 3 years is the fastest and fiercest competition I could imagine.
MaKeyabout 2 hours ago
> How are you sure that what big ai is doing is “the increase in knowledge of human kind”?

The OP didn't say that.

Cynddlabout 12 hours ago
Good question. Unfortunately, academic knowledge is widely ‘verboten’ already. Everything under paywall, researchers having to pay up to $10,000 to publish in open access in some venues, rare books unavailable even to top universities. Access to knowledge and information is increasingly difficult for everyone.

That said, what matters here is the social contract, what do I bring to society and what do we get from tech companies. For most people around the world, access to the typical leading models is out of reach. Not many on this planet can pay the subscriptions (or even API keys) that offer access to the best models. So I'm not buying the argument that tech companies are broadening access. What we're creating is a increasingly discriminatory society where the few get access to information, and the many don't.

joshkaabout 12 hours ago
Thanks for the reply here.

I guess I have some perspectives on a bunch of this. I'm for open sharing of academic work for all (but I'm not an academic, so my perspective is a consumer), so inferring your perspective here I think we agree on that. I maintain many open source (MIT/Apache2 licensed) libraries, and I've also worked in big tech (Amazon, OpenAI). I believe both in the idea of collective commons but also in the ideas that there should be the ability of people to sell software. There's tension in that social contract similarly, and it gets more complex when you look at copyleft.

I guess I'd be disappointed if this was just allowing big labs access and not more broadly allowing access to the ACM library for smaller open source models. Very much in agreement with your last points there.

Aurornisabout 10 hours ago
> most people around the world, access to the typical leading models is out of reach. Not many on this planet can pay the subscriptions (or even API keys) that offer access to the best models. So I'm not buying the argument that tech companies are broadening access. What we're creating is a increasingly discriminatory society where the few get access to information, and the many don't

I’m sorry, but I can’t buy this argument. Making information more available does not make it more discriminatory. Nobody is saying it will only be available in the best models and withheld from other models or services like the ChatGPT free plan. Nobody is saying we’re going to make the original content inaccessible through the previous means after the LLMs are trained on it. Nothing about this shrinks access or makes it more discriminatory.

I understand that you’re upset about the use of the content, but I think you need to admit that your stance is the one trying to restrain use of the content. Training LLMs on it can only bring knowledge to a wider audience, not restrict it.

Whether or not that’s a good or fair idea is a separate discussion, but arguing that this makes access to the knowledge more discriminatory and locked away is 180 degrees backwards.

lazideabout 3 hours ago
Why do you think big AI is going to just share all it’s ‘knowledge’?
preisschildabout 2 hours ago
Exactly. Companies like Anthropic do not want open access. See their FUD against "distillation attacks"
throwaway27448about 7 hours ago
> but then say that a particular use of that knowledge is verboten?

https://en.wikipedia.org/wiki/Paradox_of_tolerance

Without material values, you will be lost and confused.

Granted, the ACM is hardly a bastion of anything but self interest and greed.

maCDzPabout 16 hours ago
If it was a non profit that trained the model - would that change your mind?
Cynddlabout 16 hours ago
The issue here is ACM focusing on licensing. A non profit would not be able to pay ACM for access. Hence why this policy is hypocritical: it gives more power to the larger players and undermines smaller actors in the field who have fewer resources.
bonoboTPabout 15 hours ago
Why did you sign off your copyright to ACM then? If you kept the copyright and prevented others from distributing your work, you could be the sole distributor and negotiate the price with the AI company yourself.
MASNeoabout 15 hours ago
Non Profit != No Money

A non profit could well pay, and there are plenty of reasons frontier models should be managed by non profit. After all, why allow a for profit to benefit from free contributions?

arjieabout 12 hours ago
Everybody gives more power to the larger players. You won't work for me for $10/hr but you'll work for a guy who has more money for more. There's nothing wrong with that. Money is just a fungible unit representing value offered.
protocoltureabout 9 hours ago
>The issue here is ACM focusing on licensing. A non profit would not be able to pay ACM for access. Hence why this policy is hypocritical: it gives more power to the larger players and undermines smaller actors in the field who have fewer resources.

And when the open weights models distill all the content out of the majors anyway?

vascoabout 6 hours ago
Being a non-profit is just a tax arrangement, it shouldn't change any of your opinions. What does the tax rate of who does an action have to do with the action being good or bad?
seanmcdirmidabout 10 hours ago
I’m pretty sure most people who have papers in the ACM digital library already have “preprints” in other freely accessible locations, especially for papers written in the last two decades. LLMS have been able to search/find most of my papers for long time now.

The peer review system was broken before AI, so I’m not sure what your point is there.

jval43about 7 hours ago
My ACM papers are published in full on my website. Not preprints. In fact thats the only reason I have a website, just in case someone actually wants to read them.

The copyright agreement (actually copyright assignment/transfer) with ACM has had a carve-out for this case in it for a while, but it has to be a personal website.

seanmcdirmidabout 2 hours ago
I thought it was more liberal than that but there was always a “preprint” exemption that would get you around it.
moi2388about 15 hours ago
You mean the knowledge you gathered with public grants, with a public paid salary, yet don’t want to make freely available to the public?

Yeah, too bad

hmryabout 11 hours ago
This is about the ACM licensing it to for-profit AI companies, who will then make you pay a subscription to access the knowledge. Not about making it "freely available to the public."
PeterStuerabout 5 hours ago
So why not license it for free to those model companies that also share their models with the public? Why would Moonshot not get free access?
aatd86about 16 hours ago
As long as we are going toward a world of abundance where money doesn't mean much and the main currency is time, I can't complain. I will subsidize that with my brain power turned into ink on paper.
layer8about 16 hours ago
It’s pretty clear that the premise won’t be becoming reality.
aatd86about 15 hours ago
Well, once there is no more jobs, or once a large enough number of people are made redundant, it will mechanically happen. Future is all about research and entertainment. That is why tiktok stars and soccer players get so much money. All that being based on DARPA research, I mean the internet. People just want to have fun. Du pain et des jeux.
PunchyHamsterabout 16 hours ago
well, we are not, so far it is just concentrating wealth and power in smaller and smaller subset of people
aatd86about 15 hours ago
Yes and then continue your reasoning... Is this sustainable? What do you think it leads to?
atoavabout 16 hours ago
Abundance for whom, exactly? And if the answer is everybody: who in charge has an incentive to do this?

Sorry to ruin your day, but if the people with money could have introduce equal society, you would have noticed their attempts by now.

aatd86about 15 hours ago
Abundance for everyone or it is not abundance. With abundance, the field gets levelled because everything someone else can get, you can too. Society can't be equal when everyone is in survival of the fittest mode because resources (energy) are not infinite.
bonoboTPabout 16 hours ago
If you don't hold a patent for the use of the knowledge you published publicly, you can't prevent others from using the knowledge. You enjoy the prestige attached to the idea that you're an academic who participates in giving away their knowledge but then you play this game when that knowledge would actually be useful as opposed to being read by 3 other people in your special area who sit on your various committees in your career, now you want to forbid the use for culture war intra-elite signaling reasons.

You don't own the knowledge you put out there unless you have a limited time valid patent. The rest is absurdity. If you want to keep your findings to yourself, keep them secret.

loumfabout 16 hours ago
The intellectual property law that governs ACM articles is copyright law, not patents. I don’t know who controls these (the ACM or the authors) or what rights might have been granted to the public.

The entire point of copyright law is so that people can make their writing public and still be able to control the right to make copies (for example, into your dataset for training an LLM).

bonoboTPabout 16 hours ago
Copyright protects against reprinting or reproducing the wording and expression, not using the idea expressed in there in novel contexts.
Cynddlabout 16 hours ago
This makes no sense whatsoever. How would a philosopher of science, or a social scientist who publish in the ACM apply for a patent? Or someone who builds software (software patent not so easy to get ;)). I have applied for patents before and I'm pretty sure my patent application has been fed to countless LLMs by now.

The issue is not who owns knowledge, it's how it benefits humanity.

bonoboTPabout 15 hours ago
They can't apply for a patent in those categories, so they have no way of preventing others from reading their text (or using a program to process the text, and compute statistical properties), and using the ideas in other contexts (without re-expressing the same text).

If someone reads a philosophical essay, has a heureka moment from it, applies the principle to their work, and makes bank (commercial profit), they never have to pay a percentage to the author of the essay.

dwatttttabout 16 hours ago
The world would be quite different if AI companies had to create the knowledge they trained on, rather than consume that knowledge freely given away. They're not known for freely giving away their produce either, I don't know why you think the ire should be pointing in this direction.
bonoboTPabout 16 hours ago
I like open models for sure. I support free software as well. But using published knowledge to solve new problems was never disallowed, even for profit. Today, a for-profit company, e.g. a gigantic Big Pharma company can have their employees read chemistry and biology papers and use the knowledge gained from it to improve their products and processes and make more profit without paying a cent to the authors (beyond what they may get - likely nothing - due to the potential paywall).
shartsabout 15 hours ago
Patents shouldn’t exist.
bonoboTPabout 15 hours ago
Patents were invented to incentivize disclosure of inventions and prevent extended secrecy. In exchange for the public disclosure (which allows others to experiment with the idea without selling yet), you get to keep exclusivity for N years.
willy_kabout 16 hours ago
“If you don’t lock your bike, you can’t prevent others from taking it for a ride. You enjoy the mobility attached to the idea you’re a bike rider who rides a bike but then you play this game when that bike would actually be useful as opposed to sitting in the bike rack all day”.
bonoboTPabout 16 hours ago
Nope, and I'd also download cars.
fsmvabout 18 hours ago
How about we give humans access
m-hodgesabout 18 hours ago
Do humans not have access? https://dl.acm.org/openaccess
0xCE0about 17 hours ago
It seems that the "open access" is more marketing than something reality based they actually want to do.

https://dl.acm.org/openaccess ---> So how to access the content? Do I have to register or what? It the "open access" only for academic org's people or for everyone in the world?

https://dl.acm.org/ ---> Okay, nice simple search field without loggin in, but when you try to search something, you get thousands of results, even if you search specific author and the exact name of paper, you will get hundreds of results and the thing you want is buried somewhere on page 247. Filters of authors, years etc. are for Premium subscription. But just googling the thing finds the link to ACM... And want to get the actual PDF? Hope it says "free access"...

riedelabout 16 hours ago
Nobody uses their search anyways.

Google scholar has become the way to search papers (which is somewhat worrisome). What people want from there is a download for all the stuff that does not list an author copy. Open access IMHO is just a reaction to the fact that mostly you would not need a subscription anyhow. Now the authors are paying upfront or universities are paying flat for all their researchers. The problem is now the incentives are not increasing the number of subscriptions but increasing the number of papers published.

hmokiguessabout 17 hours ago
So will LLMs have premium-tier access or basic-tier access?
dominotwabout 17 hours ago
it is open like open in openai
MASNeoabout 14 hours ago
Made my day. Thank you! I have to remember that. Best thing I have read in the entire day.
belochabout 16 hours ago
Gracanaabout 16 hours ago
> From January 1, 2026, all ACM publications will be published Open Access (OA), free to read, share, and reuse in the ACM Digital Library.
Lockalabout 3 hours ago
This is called Sci-Hub and LibGen
jval43about 6 hours ago
It was so close! I could already see the end of all the predatory publishing and flourishing of open-access. Some even required by law. Knowledge was finally going to be free.

But no, it will presumably get much worse as LLMs are inserted into this equation as yet another and new gatekeeper.

(Disclaimer: I have publications with ACM, non open-access. And ACM wasn't even too bad, it's the others that give me pause.)

juancnabout 19 hours ago
They probably already scraped it.
ameliusabout 2 hours ago
Through scihub, probably.
IAmGraydonabout 15 hours ago
Of course they did. There is open access to this library. These guys actually think they're offering new training data? It's kind of hilariously naive.
pohlabout 17 hours ago
No doubt
computerdorkabout 12 hours ago
Haha, agreed:)
Rochusabout 1 hour ago
Scientific publications (including ACM) should be openly accessible to anyone. I find it very annoying that there is a paywall everywhere. And for the (few) people here seeing AI as something useful, it's definitely better if it is trained on scientific publications than on e.g. Reddit discussions.
rurbanabout 15 hours ago
So give it for free to the open weight models, and charge the closed weight models. Easy
spoaceman7777about 16 hours ago
Blocking access only hurts people who follow the rules. Unblocking access lets them compete with those who break the rules.

I think the right choice is pretty clear...

estebarbabout 15 hours ago
I'm really not sure how this would work. I don't know how the ACM works, but in IEEE you would have to give them your publishing rights. However, training a LLM is not publishing by itself, it is a derivative work? Any way, at this point authors should be entitled to monetary compensation, not the publisher. The deal is totally different.
bezkoabout 17 hours ago
Something something Roko's basilisk
iFireabout 16 hours ago
Would you prefer a parquet dump of acm articles to hugging face?

A llm emulating a person is why many of my used sites banned llms due to scraping bandwith costs

Is this about access or accessibility to claude (for example)

I am sure the entirely of human computing knowledge is not that big.

agarabout 15 hours ago
This reminds me of the "I drink your milkshake" scene from "There Will Be Blood."

Does the ACM really think LLMs haven't already consumed 90% of the content through other sources?

Advertisement
skippyfishabout 19 hours ago
ACM has been leaning heavily into AI-generated content for their journals in the past year, and this article is no exception: it appears to be 100% AI generated and full of LLM verbiage.

There's something hilarious about that, but also, snake eating its own tail.

Diogenesianabout 18 hours ago
I don't think this is AI-generated. It is focused and direct. It reads like anodyne albeit totally human academic manager writing.
boredatomsabout 9 hours ago
Presumably the LLMs already have all this from other sources of pdfs?
etdznotsabout 11 hours ago
Most ACM text is already part of the pre training corpus for all frontier LLMs
croesabout 18 hours ago
Are they in the position to do that?

What about the authors?

dsr_about 17 hours ago
The ACM sent around a nice query to members which made it clear that they were going to do it even if 100% of the members said "No, don't do that".

I suppose they could be sued.

em3rgent0rdrabout 17 hours ago
Don't you grant ACM a right to distribute your work when publishing? So doesn't ACM already have the right to grant access to AI?
davexunitabout 10 hours ago
No it's not.
bitwizeabout 15 hours ago
Well if LLMs get access to it, we humans should get free access to it as well!
nekusarabout 15 hours ago
What about us average humans, or is it ONLY the corporate LLM token dealers who get access?

Either way, I'll pirate.

krater23about 2 hours ago
They really think that their data is not already part of the models. Silly...
brcmthrowawayabout 17 hours ago
It's already in there
throwaway27448about 10 hours ago
As if anthropic, openai et al would ask for permission lmao
Advertisement
jruohonenabout 9 hours ago
So much misinformation in this thread.
hmokiguessabout 17 hours ago
Now? That's cute.
recursivedoubtsabout 19 hours ago
the digital library should have always been open access

now it will be fodder for the slop machine

(i think that LLMs are going to wreck the peer review system for all but hard-experimental papers)

jerfabout 17 hours ago
Since starting to use AIs seriously for search in the last 3 months I have read and referenced more published papers and academic primary sources then I think I did in the previous 5 years. They're fantastic for pointing at some claim and asking for the primary source for it, and then it does the work of following things through the layers of backref to the original, assuming it's online. Opening the library up to AI access makes it far more accessible and usable then it was before.

AIs, at least in their current form, make you more who ever you were. If you want snap, glib answers of dubious accuracy, they'll give them to you, more easily than ever before. If you want to dig back into primary sources and get the original content, they'll do that for you, more easily than ever before.

Can't speak to how the science infrastructure is going to handle them, but if it takes down the peer review system, which I think has been worthless for probably going on two decades and has just given the entire enterprise a false sense of assurance, it'll probably be a net gain in the end. Peer review is a source of more problems then it is solving right now.

ilakshabout 19 hours ago
The real efficacy of the peer review system has always been somewhat questionable. That's a big reason why arxiv is so prominent.
jmountabout 19 hours ago
Came to say the same thing. Now remains the time to give the public access to the ACM digital library.
cobbalabout 19 hours ago
ACM agrees I think: https://dl.acm.org/openaccess

> Beginning January 2026, all ACM publications and related artifacts in the ACM Digital Library will be made open access.

harlesabout 18 hours ago
Wait, so did this already happen and it just didn’t get attention?
dan_geeabout 19 hours ago
> I think that LLMs are going to wreck the peer review system for all but hard-experimental papers

I think I may be shadowbanned, but at what point do we start viewing LLMs as a national security threat?

skywhopperabout 19 hours ago
TLDR: “we’re going to try to get money from LLM providers for access to our back catalog without getting permission from the authors or providing them with any share of the revenue”.
anticensorabout 19 hours ago
Scholarly article authors never expected royalty or permission for their scholarly works to be reused.
ilakshabout 19 hours ago
I don't want to be mean but if it's that hard to even recognize that LLMs are even a valid thing that could intersect with your business that you need some kind of campaign for it..

It's almost like, "don't hurt yourself unc, we will just search arxiv".