DE version is available. Content is displayed in original English for accuracy.
Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
100% Positive
Analyzed from 294 words in the discussion.
Trending Topics
#run#huggingface#page#countdown#https#news#hours#meta#googly#claudy

Discussion (11 Comments)Read Original on HackerNews
Edit: for anyone curious, I saw the huggingface countdown page for Kimi K3, which gave me a hunch that led me to the countdown page for Qwen3.8-27B. As of ~11am EST today, the countdown page said 2 days and 2 hours to release.
113 and assumming close to constant load and the gpus taking turns as it does on my machine with two gpus is around 113gpus*113W = 11KW... just to hold and run that one instance... and only god knows the cooling requirements...
I guess just someone on Meta/Googly/Claudy will say... oh nice, let's download it and run it...
How did the unsloth guy did to process a file that size?
https://huggingface.co/unsloth/Qwen3.8-2.4T-A95B-GGUF
another 50k for the machine to put your 8 h200s in
Its pricey obviously but the 500-1M price tag is not out of the realm for most large companies.
You can run it for like <$50/hr. You don't need to be Meta/Googly/Claudy for that. For batch inference it makes sense to run it yourself(high latency, high throughput, predefined workload for few hours or day).