Back to News
Advertisement
Advertisement

⚡ Community Insights

Discussion Sentiment

67% Positive

Analyzed from 299 words in the discussion.

Trending Topics

#model#run#mac#price#pricing#inference#blackwell#studio#lot#com

Discussion (7 Comments)Read Original on HackerNews

apimade•about 1 hour ago
I guess the question is; is this the late-cycle cash-out (aka harvest pricing) akin to Sun Microsystems at the dot-com peak -- or is it a repeat of the crypto pricing hijinks we've already seen from Nvidia, where they're just exploiting the lack of supply?

That's a 90% uplift since the original pricing, in a market that is already showing signs of seizing.

With Apple offering leasing options, CXMT knee-capping Samsung, all of the big tech players on a run to outspend on CapEx by the end of the year..

We're at a point where local models are exceptionally capable, model-on-silicon dies like Taalas (recently acquired by AMD) may be cutting inference cost substantially for the 90% of work we do day-to-day (similarly Alibaba's T-Head division with open-model-forward inference chips being produced domestically in China).

If we shifted all of the design/planning to cloud models like Fable 5/Sol 5.6 Ultra, and day-to-day operational inference to these chips -- it's quite likely we'll squash usage to single-digit percentages of what we're currently using. But we should also expect the model providers to take a similar approach.

In any case -- I'm keen on demand destruction, both as a consumer and someone without skin in the game.

areoform•about 2 hours ago

    > RTX Pro 6000 Blackwell has 96GB of GDDR7 VRAM
A mac studio with 96GB unified memory costs, $5,299.00.

https://www.apple.com/shop/buy-mac/mac-studio/m3-ultra-chip-...

LLMs can re-write and cross-translate software really well. Why does CUDA still have a $11k price premium?

xyzzy123•about 2 hours ago
You can get 5-10x the perf out of the blackwell under the right workloads. It has faster VRAM (> 2x) and can do a lot more matmuls (>> 10x).

They're both good value (or crazy expensive) depending on how you look at it.

It depends how you price the ability to run a particular model at all, vs run the model quickly and serve several parallel streams.

MiroslavPokorny•about 2 hours ago
Because a lot of people are buying them at that price.
swiftcoder•about 2 hours ago
The Blackwell is going to be a lot faster for the same memory capacity than a Mac Studio, surely?
JackSlateur•about 2 hours ago

  LLMs can re-write and cross-translate software really well.
Is that was true, claude would be written in Rust;
lolive•about 2 hours ago
Just when Cyberpunk gets almost bug-free, the price for a decent Gpu to run it smoothly is skyrocketing. #doooh