RU version is available. Content is displayed in original English for accuracy.
Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
76% Positive
Analyzed from 877 words in the discussion.
Trending Topics
#video#models#seconds#model#those#same#frame#still#weights#together

Discussion (26 Comments)Read Original on HackerNews
Is this a common approach to reducing weights with "no loss in output quality", assuming this is true? Seems almost too simple to work. If this is doable, would this be applicable to LLMs as well?
Neat with native frame-to-frame generation, but wonder how easy it is to "link" together clips at the intersection, typically the models kind of lose the "momentum" across these stiches, being able to merge things with frame-to-frame between clips might help with this it feels like.
I remember a paper which was posted on HN a few weeks ago where somebody implemented KAN networks in FPGAs, since those can readily be approximated as LUTs.
Is MiniMax H3 capable of logical / technical reasoning, or is it purely art oriented?
As far as I can tell, the current ComfyUI nodes don't even do compilation, and I haven't looked into what attention mechanism they're using, but I'm sure with time these durations will come down even more.
Pretty cool.
But assuming you have a 16GB 3060, how long would it take to generate a 15 second clip?
The only one that looks "off" is the beverage ad video during the can opening clip, it still has that "AI smoothening" effect. Good thing this can be done pretty well using traditional rendering.
I feel like for a good while now we'll transition into a process that uses traditional "close-up" rendering/shots + AI generated wide-shots or quick cuts.
Exciting, but also troubling. This being open-weights is a massive win for the community though.
I remember reading a report where people running AI-model instagram account were using insanely long and detailed prompts about the setting, lighting, makeup, pose, disposition, clothing, etc. about their models. Presumably with some reference image of the face / body to remain consistent across images.
It‘s not clear to me whether a sufficiently detailed prompt can generate actually interesting video with a natural ”texture” (for lack of a better word).
I suspect it'll be quite a while until AI gets a good enough aesthetic sense to do this, as even with static HTML websites humans can easily see that it's AI slop.
There is some debate on the license for those in the US, UK, EU, plus… no comment other than whew those samples though!
You just have to pinkie promise you won't make disney mad and they will send you a licence https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/Q...
This is AGI.
As with the arts, 99.9% of people can't use these models to express vision, get attention, or achieve distribution.
The game is the same as it has always been. You still need hard work, taste, something important to say, the ability to articulate it, good timing, and luck.
Nothing has changed. We can just build faster.
What this does enable is for more to be created that caters to a wider variety of interests. It disrupts existing structures of capital allocation, production, and distribution and gives new players a chance to reshape the game.
The bar will rise and people will still be running at the same pace on the treadmill. There will be more to see, but less time to see it.