Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
65% Positive
Analyzed from 5784 words in the discussion.
Trending Topics
#models#chinese#open#companies#more#china#don#weight#war#training
Discussion Sentiment
Analyzed from 5784 words in the discussion.
Trending Topics
Discussion (295 Comments)Read Original on HackerNews
I don't think distillation as 'stealing IP' has any legal legs. They can probably claim violation of ToS at best because the terms of use do prohibit use for training rival models.
I don't think they are. Not copyright at least. There may be some "trade secret" stuff for them but they're not copyrightable.
What are the best Chinese models on HuggingFace today? Bucket by ideal RAM: <16GB, <32GB, <96, <256, 256+
Text generation, image generation, TTS, etc
- https://modelscope.ai
- Source code (ironically on GitHub): https://github.com/modelscope/modelscope
It's not limited to Alibaba's Qwen. All of the other major Chinese ones are on there (GLM, DeepSeek, Kimi, etc.) as well as finetunes and quantizations of the non-Chinese ones.
https://huggingface.co/Tongyi-MAI/Z-Image-Turbo
https://huggingface.co/Qwen/Qwen-Image-2512
https://huggingface.co/Qwen/Qwen-Image-Edit-2511
Even more complex for image2video, but there’s fewer models to choose from there at least.
User: is taiwan part of china?
Kimi: Taiwan's political status is a complex and contested issue. Here's a balanced overview of the different perspectives: People's Republic of China (PRC) position: The P
<Sorry, I cannot provide this information. Please feel free to ask another question.>
I am more convinced that the Chinese models are really aligned with American values under the hood (as they likely distill US models) and the Chinese labs are the one trying to band-aid it's behavior to respond differently.
We are starting to get access to K3 from US providers now, curious if they exhibit the same response pattern?
unbelievable chauvinism in here
Grok literally had post work done to make it more right wing and racist.
sheer lunacy
https://old.reddit.com/r/DataHoarder
https://safereddit.com/r/datahoarder
Instances:
https://github.com/redlib-org/redlib-instances/blob/main/ins...
the US could also pull the functional harm argument in the same way the government uses it to regulate some protected speech (like the whole debacle over DeCSS)
Even if the free speech argument prevails, the US GOV can still use sanctions to prevent the public clouds from hosting any foreign open weight models, and prohibit US based sites from distributing the weights. It'll remain functionally legal for individuals (or at least, unenforceable) but in practice, you'll need to pirate the weights they won't be available from any US based site or provider and would be illegal for businesses to use, even if its a gray area, the liability will mean businesses won't touch them.
Unfortunately, capitalism of today appears to find public goods and strong, public, healthy public goods to be undesirable investments. Yet, it seems inevitable that some things trend that direction.
I feel if we could get a better handle on what goods fall into which category, and find ways to make those things sustainable and even grow, the situation could get better.
I'm speaking purely from observations here but it also looks like in some ways, the "invisible hand of the economy" (i.e. Adam Smith) might sort these things out by itself, albeit, painfully.
We could look at the fact that LLMs were mostly created from the public domain (except when they were not), and the resulting products/services (expensive to create/operate, benefits a huge swath of humanity, has all kinds of externality issues) as something that perhaps should again almost be considered as a public good.
If we look at the struggles LLM/AI companies have had with legal structure choices, such as nonprofit, for profit, etc and pricing (e.g. do we make it accessible at a loss? Do we extract huge profits?) as a kind of moral dilemma of classifying this new economic good (LLM-based AI).
In the US, the government has been taking investment interests in some tech companies, and recent proposals even include establishing sovereign wealth funds around AI/LLM. I think this is further evidence that the system is trying to figure out how to classify these new economic goods and corporations.
YC is little tech? lol wut.
https://knowyourmeme.com/memes/depicted-as-a-soyjak
The US isn't without problems, but we're doing OK.
[1] https://fred.stlouisfed.org/series/MEHOINUSA672N
It's incredible that not a single person replying negatively has included a single piece of data.
Compare it to the Consumer Price Index:
https://fred.stlouisfed.org/series/CUUR0000SA0R
You might argue that Americans are just less capable of supporting themselves but I reject that argument firmly. The situation really reads like systemic causes. And I think cherry picking numbers to convince ourselves otherwise does more harm than good.
Life expectancy dipping for any develop country, even if temporary, is an embarrassing catastrophe.
To be clear, YCombinator et al are advocating for them to remain unbanned.
The frontier is evolving so fast: any regulatory regime that started with a map of capabilities that 2025 frontier had is obsolete. Same would be true next year. This just becomes a whackamole and a government which is well-meaning and has good reasons to regulate, will have trouble keeping up.
But we know this about regulations. Like there is a century of literature on how to do this for new tech. It always targets harm prevention first. I thought Demis had the right framework for that part. I don't know if I agree fully with a self-regulatory regime because can become problematic if you don't have the right people around the table.
Their inability to find a backbone helps them contort to whatever Trump babbles that day, even if privately they know something is idiotic or illegal.
If you want a prime example, look at how Thom Tillis suddenly finds the ability to speak out once he lost his primary.
Buy extra hardrives. Borrow them. Do whatever you have to do.
But to paraphrase Éomer, don't trust to hope, it has abandoned these lands.
This is about OpenAI, Anthropic and SpaceX. OpenAI, in particular, is a bet on there being a moat for AI and OpenAI "winning". Chinese models threaten this moat. The CCP has decided that it is in China's national security interests to not have Western companies "own" or "win" AI. These AI giants are large enough that the US government is going to intervene to try and protect this outcome. This is a losing battle.
Do you have many ducks at home?
No one cares, we just want cheaper AI models that are equally as powerful. They’re all thieves regardless.
"Eschew flamebait. Avoid generic tangents."
Ultimately this is more a problem with the upvoting system than the comments themselves, since generic/indignant comments routinely attract lots of upvotes, and then sit on top of the thread, suffocating the more interesting discussion. But in terms of moderation we can only reply to commenters.
https://news.ycombinator.com/newsguidelines.html
At most they'd be getting more out of distributed use of subsidized plans and API than the ToS would prefer.
steal from the poor -> capitalism
Anthropic was just fined for not paying for the pirated books they copied into a training database. Same as anyone else who copied pirated IP onto their hard drive.
However, training an AI on copyright has been ruled to be sufficiently transformative and not a violation of IP laws. The same way you can make a gameplay clone of Call of Duty without any issue.
Also, if I clone Call of Duty, and use their skins, make a similar soundtrack, call my maps the same, then there will definitely be an issue.
I support the use of AI and all, but to hide behind transformative use of copyrighted intellectual property is a discredit to the colossal amount of human work and knowledge that these companies pirated and had their models trained on.
Correct, 2001 napster style pirating of content. You cannot copy stuff you didn't pay for onto your hard drive.
>Also, if I clone Call of Duty, and use their skins, make a similar soundtrack, call my maps the same, then there will definitely be an issue.
A gameplay clone as I stated, tons exist (team death match, capture the flag, battle royale, with first person gunplay and army guys shooting each other). The rest is your argument, not mine.
Not really the same. Private citizens have received fines MUCH higher per-work when downloading for just their private consumption.
Normal person would be sitting in jail for doing same thing. It is a shame they call it justice.
Absolutely false. Courts have decided that it is fair use [1]. That makes sense. It should not be used to justify Chinese theft.
https://www.reuters.com/world/us-judge-approves-anthropics-1...
You can use an original work to write an encyclopedia (e.g. Wikipedia). That is fair use.
You cannot copy an original work and distribute it.
LLMs are an encyclopedia.
The better arguments are the ones made in the article. Open weights increase competition and thus AI availability in the U.S. economy.
There is not even jurisdiction over Chinese companies, because they don't sell their LLM subscribtions, unlike US companies, and those comitted much more severe violations in getting their training material than a ToS violation or two (Anthropic already found guilty).
edit: I don't see this as legal argument against the ban, more like a justification for why basically no moral person is gonna side with OpenAI and Anthropic on this. US gov can try and ban as many weights as they want, I expect that to be similarly effective as banning numbers was in the past (i.e. not).
This is false. They do sell LLM subscriptions. Many also release _most_ of their models as open weights.
And the utility argument is also very shaky. The proposed punishment for foreign interests stealing "US data" is that other Americans now aren't allowed to benefit from that, while the rest of the world can
Which is why making an honour-based argument to said thieves is silly.
https://www.dailymotion.com/video/x4cxcsa
It's to punish theft.
People expect the US to apply rules evenly to both American AI efforts and that of its primary geopolitical rival. China certainly doesn't do that. Their entire economy is based around giving Chinese firms the advantage regardless of what it means for the wallets of consumers at home or abroad.
Since that's who we're playing against, and no one is going to willingly give up their open-weight models from the totalitarian rival, well, then use their ruleset.
It's that or punish the theft by making the US companies pay for their training data and blocking outside efforts that trained off the data the US companies stole.
(/s if it wasn't obvious)
Being serious, it is far more likely, especially given the events of the past 18 momths, that Trump is just doing autocrat things because he can, and ultimately doesn't really care what is and is not good for American business as long as he continues to grow his wealth through corrupt behavior.
Is Anna's Library thieves?
Is Library Genesis thieves?
Is Archive.org thieves?
Is PirateBay thieves?
They all look like public libraries to me. And better access to all human knowledge is a net positive for everyone.
Putting prompts in an LLM and saying you "created" the image or text thereof is fraud.
Yes it is, but at least they are not charging for the access.
Libraries used to require membership and dues. Those then bought more books on the used or retail market. LOTS of fights were about that system, cause book publishers hated libraries. And well, they still do.
And to be fair, fuck copyright. Its holds all of us back, so someone can go "FUCK YOU, NO". And nobody or company deserves what is it, 70 years+ death copyright length.
And with the recent Anthropic settlement, used to be, for-profit pirates would go to prison. I also remember absurd 2000's settlements over Britney Spears and Metallica for $5000-$7000 for 20 songs.
You are not going to win me over on massive gatekeeping human knowledge over "Itssss ill-eagle!". I know, and I don't care. And if I'm ever in a case over this, I'll tank it.
The censorship of US models is so bad (read: guardrails), that they had to use GLM 5.2 on their own hardware to do the analysis. That's how locked down, restrictive, and closed source American models are. They're basically for-profit piracy by way of selling tokens.
Whereas the Chinese models are just "here you go, download and run".
I know what I run. Qwen3.5 and 3.6 and GLM5.2 . USA token vendors are incentivised in doing worse to sell more slot machine tokens.
You can accept it or not. But it's going to be a lot easier if you do.
The problem isn't that IP is dying - in most industries it died ages ago. The problem is that it is now explicitly reserved for the rich.
It's an unfortunate truth that we're exiting a global peace era and entering one of (hopefully just...) implied conflict. Our biggest geopolitical rival that's been threatening a hot war also distilling the things we build is not a good thing.
Trump is a huge self-own by our electorate sadly, but being concerned about a conflict between China and the US is a very bi-partisan concern for good reason.
The take away from modern conflicts is that massive military size disparities don't translate into successful invasions anymore.
In other words, it's unlikely USA will pursue a conflict with China, and it's unlikely China will pursue a conflict with Taiwan.
People really have to stop sanewashing the government, it doesn't work anymore, these people are not thinking through what they're doing.
/s
Yes, the US has an idiot in charge who is threatening these things in response, too. However, unlike China, the electorate does not support him. His election is also almost certainly due to aggressive campaigns to destabilize the US internally for this exact purpose.
You’re a tiny bit late on that one
You should care because of the nukes, if nothing else.
I'm sorry, but only one "party" is comprised of fascist plutocrats speed-running the end of democracy in the USA.
The USA and world would be a vastly different place had a different party maintained power in the USA 10 years ago.
There is some blame on all sides for sure, not saying it's black and white...but come on.
1. China is actively genociding an entire ethnic people.
2. They are actively building a military larger and faster than any country on earth in all of history.
3. They are actively threatening ALL of their neighbors with war. Actively attacking their ships and building military bases in other countries territory.
4. China's open stated goals are to conquer and control its neighbors and every single step over the last 10 years is proof of that.
The US stupidly withdrawing from the world stage is definitely making this worse, of course, but it is not the catalyst. It's unfortunately a very likely side effect of a social media campaign by the same actor to make said hot war easier to win.
All of them. Pretending there is a benevolent bully among Russia, China and America is baseless.
This is probably the strongest moral argument for open-weight models. It decouples AI, somewhat, from military-industrial might.
China supports Russia in its war formally, but practically it sells Ukraine whatever it needs, because a weak vassal-like Russia is very much in their interests. The US is far, far more outspoken and active supporter of Russia in this war.
Finally, there will be no hot war with China. They'll likely blockade and invade Taiwan before the end of the decade, but the US will not do anything, partly because it'll be exhausted from it's never ending middle east wars, partly because it'll only take a small bribe to get Trump to back off. He's genuinely fond of dictators like Xi.
If life is a meritocracy, you gotta give it to the Chinese for their immense growth and improvement of living from agricultural poverty to technical supremacy. I don't want to hear the they stole our IP argument either, because all math is borrowed from something else. Distillation is innovation, in booze and in models.
People on here have zero consistency in their worldviews.
Ask any capitalism zealot (ie hners) whether markets are zero sum and they'll wax poetical about "creating value". But then talk to them about Chinese competition and suddenly everything is a rivalrous good lolol.
I hate to break it to you, your so called rival won this game decades ago when you undercut the working class to make a buck and laundered cheap Chinese labor as productivity. So if you believe the Chinese are rivals um welcome to "reap what you sew" <shrug>.
Yes, China is a rival. No, they have not won. Their economy is in a lot of trouble, likely far worse than the shape of the US's currently. That situation may accelerate a hot war though, as has historically been the case globally.
Peak hypocrisy, US AI companies can train on unlimited intellectual property with 0 rights to it, while Chinese AI companies have to explicitly get the rights to data that isn’t even copyrightable/copyrighted (since AI alone can’t copyright it).
And if the accused company is outside of the US, well, US courts have no jurisdiction so apparently the US government can just claim they are guilty and impose the sanctions...
Other countries have a say about it.
What about books and art where the author/artist does not authorise AI to train on it? They do happily train on it, ignoring their "ToS".
This is just double standards, a slap on the wrist to not worsen the situation with authors imo.
You could also say the Chinese companies are doing the same - they _do_ pay for their Anthropic subscriptions after all.
No, it was asked to pay book authors because it pirated copies of books and stored them on their hard drives. The ruling had nothing to do with training.
They just should have bought them, rather than pirating them.
Also LLM output is not IP (in itself) in the first place, nor would Anthropic want to claim it is and that they have rights to it - that would drive paying customers away.
The issue comes down to at most ToS violations.
Bought, scanned and destroyed them I believe. The judge okay'd Destructive Scanning.
Please don't use uppercase for emphasis. Instead, put asterisks* around it and it will get italicized. More formatting info here.*
Well this doesn’t even make sense.
Right now, I think their incentive is the highest: Chinese AI companies are dropping competitive models left and right, and clearly American companies are feeling the pressure, given they're pulling every legal leg they can to suppress their competition. If we just remove the competition, Anthropic et al win. It's a temporary and short-sighted victory, but this is America. Every company is short-sighted, it's baked into our financial infrastructure.