ES version is available. Content is displayed in original English for accuracy.
Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
56% Positive
Analyzed from 2394 words in the discussion.
Trending Topics
#open#models#chinese#source#model#openai#weight#don#companies#microsoft

Discussion (59 Comments)Read Original on HackerNews
Pure game theory, a bunch of losers in the AI space groveling at the white house to try to stop their competition.
And besides, it seems to me like the ethical choice. Given how these models are, in some very real sense, mechanical plagiators, built on the generosity of creators past and present, some of them now in danger of being replaced by the machine. I think the least the labs can do is open these models up. These and other such considerations were the reason OpenAI started with that name. Of course, it was questionable that those ideals would survive the encounter with generational wealth. Just look up what the founders of Google were saying about advertising when they were two students tinkering at an as of yet unproven tech. Same thing for OpenAI, self-interest speaks that much louder when there's real money on the table.
It just boggles the mind that people now make excuses for their all-too-predictable about-turn.
To me, "open these models up. " must mean provide all the data and supporting documentation required to reproduce the model. That would be "open". Postgres is open because you can download all the data and supporting documentation and reproduce the binary yourself. However, being able to only download a postgres binary would make it no longer open.
Additional training on top of an open-weight model sounds analogous to writing mods for minecraft. You may change some behavior but that doesn't make minecraft "open".
Note: Just pointing out the comment intent and nothing else
It would seem as if the community either isn't doing that or is relying on the Chinese to do that.
> Distillation ... reflects a long tradition of learning from, building upon, and improving existing technologies, a tradition that has helped drive innovation since the rise of the open-source software movement. By contrast, unlawful efforts to extract value from closed models raise legitimate concerns. Those concerns should be addressed through targeted legal and commercial frameworks rather than sweeping restrictions on techniques that play an important role in AI innovation.
Sounds like they are saying "please protect our IP theft" that created closed weight frontier models in case we arbitrarily decide to close our models. But don't get rid of distillations in general so that we can all also keep benefiting from open models. We don't want to lose the ability to benefit from the work of others as we launder IP into closed models.
Surprised Linux Foundation kept their name on it with that.
i've said this in other comments but the fact that these companies are trying to align with open-source when there's no "source" included with their models is pretty damning. I think it lays bare the absence of any kind of noble or righteous motive with respect to distillation.
https://www.microsoft.com/en-us/corporate-responsibility/top...
This is the correct stance, hopefully this is the stance that prevails.
It makes sense to see Meta, the startups, and the VC firms on this list. And it makes sense to see OpenAI, Anthropic, and Google all missing. Microsoft is somewhat unexpected to me. I don't think of them as having a focus on open weight models (no more than Google), and they have a large stake in OpenAI. Maybe they're looking at this from the angle of Azure providing compute. But then why is Amazon missing, when it has AWS?
And of course, we don't see DeepSeek, Moonshot, or Z.ai on here. The letter is about American technological leadership after all. But then we _do_ see the French company Mistral! Mistral who released [a playbook](https://europe.mistral.ai/) for Europe becoming a self-reliant AI powerhouse.
I assume all of this is in the context of influencing the Trump administration's thinking on Chinese open weights models. But the letter is really about _American_ open weights models. The call to action is all about keeping the American open weights ecosystem competitive. I didn't even realize that was under threat, so I feel like I'm missing something.
OpenAI and Anthropic can bring American open-weight model providers like Meta to court but not the Chinese. Chinese companies breaking American IP law (whether the law is correct or not) are basically immune. This has been going on forever, I guess enough is at stake finally for the federal government to care.
(a) sellers of hardware (Nvidia, IBM) and those who rent hardware out to others (Amazon, Microsoft)
They want open models because they don't have to license it to run it on their hardware, so they can offer lower cost and inspect what they are offering to clients
(b) open model creators (Arcee, Perplexity, Mistral, IBM, Microsoft)
...who would be unable to work if open-weight models are banned
(c) Software-heavy companies ( Mozilla, The Linux Foundation , ServiceNow, Microsoft, IBM, heck all of them)
...who want models that are cheap to run because that reduces the cost of running an LLM (whether you are trying to replace a software developer or just assist them, the value of a reduced cost LLM is directionally the same)
China’s open-weights AI strategy is winning
https://news.ycombinator.com/item?id=48979269
The arguments against open source AI are bad
https://news.ycombinator.com/item?id=49024643
AI Kill Switch Act: Official Bill Text by Reps. Lieu and Moran
https://news.ycombinator.com/item?id=49028757
Microsoft, NVIDIA, Meta, Palantir, IBM...They have all been actively hostile to open source for decades, and have a history of embracing it only when convenient and profitable.
Microsoft benefited from its close partnership with OpenAI for years right up until it went sour. Where was this enthusiasm for open weights then?
Meta was developing open models and then abandoned that effort in search for profits. Muse Spark is now fully closed.
All these companies have the resouces to train and release frontier open weights models today, but choose not to. So spare me the marketing and virtue signaling.
If I change one weight in Kimi, is it still a Chinese model?
If I fine-tune Kimi on pro-America nationalistic freedom loving anti-Chinese propaganda, is it still a Chinese model?
If I distill Kimi from one server to the next without ever directly transferring any weights, is it still a Chinese model?
If Claude accidentally trains on some of Kimi's output, is Claude then considered to be a Chinese model?
If China's next open-weight model is released secretly through a European company, is it still a Chinese model?
I don't know that the US government has a unified position on open source AI. But it's not like this isn't a discussion happening among policymakers.
Ungovernable makes intuitive sense, but I don’t understand how open weight models could be decelerationist or slow the development of the frontier. Was there an argument for those positions?
In other words, if AI models become a commodity, the "AI leaders" will not be able to burn the cash like they do today, thus resulting in a "slower" (one could say instead "more sustainable") development of AI models.
Is the US planning to invade China and take away their models? Because otherwise China is still going to have the models even if the US bans them. I don't think they're going to be going with the US position on this one.
[0] https://en.wikipedia.org/wiki/Operation_Aurora
Hey nvidia, what about making your full set of linux drivers open source?
I mean, I do agree that being open is important but you're hijacking the narrative here, and I suspect being open have nothing to do with any of this.
Start with the outcome you believe will be the most in line with the spirit and traditions of the open source community. This is precisely what won't happen.
It won't necessarily be the inverse. It could be, of course, but it's also likely to be a compromise between the two.
For a topical example: The companies doing "AI" layoffs aren't doing so out of a sense of having failed their staff that helped carry them so far. Often, these announcements come on the heels of record-breaking profits. They're doing it out of contempt for workers.
See https://www.cnet.com/tech/services-and-software/cory-doctoro... for Doctorow's take on centaurs vs. reverse centaurs. Broadly speaking: C-suite business leadership wants reverse centaurs. Workers want centaurs. Unless you have a union (or some other collective bargaining power structure that I'm not aware of), the C-suite calls the shots.
The same can be said of the current administration. They don't give a fuck what most of us want on any given issue. They're captured by an aggressive ideology. Their only incentive is to be able to spin their decision to make themselves look good; to posture as "strong" leadership.
there's nothing open source about an open weight model. It's more like the binary blobs Linux fought so hard against. If you can't compile, or in this case train, from source then you're missing the "source" part of open source.
While your point is valid on its own merits, it isn't relevant here.