Back to News
Advertisement
Advertisement

⚡ Community Insights

Discussion Sentiment

87% Positive

Analyzed from 464 words in the discussion.

Trending Topics

#ultrafast#speed#frontier#faster#fast#models#openai#gpt#mode#fable

Discussion (15 Comments)Read Original on HackerNews

iamcoder18•40 minutes ago
I've been waiting so long for something amazing to come out of the OpenAI and Cerebras collaboration.

> In our evaluations, GPT-5.6 Sol on Ultrafast mode answered all 2,500 HLE questions in 11 hours and 11 minutes. Claude Fable 5 needed 78 hours and 27 minutes, more than three days of continuous compute, to arrive at the same conclusions. In other words, Ultrafast worked through the frontier of human knowledge in a single working day, achieving comparable accuracy nearly 7Ă— faster.

This is actually insane.

Hopefully the release ultrafast of Terra and Luna too.

piyh•35 minutes ago
Feels like the 90's again where single threaded speed is improving fast. ASICs and wafer scale rather than node shrinks, but end result to me the consumer feels the same.
wxw•about 1 hour ago
> Compared with output speeds reported by Artificial Analysis GPT-5.6 Sol on Ultrafast mode runs 11x faster than Fable 5, and 5x faster than Opus 4.8 on Fast mode.

Awesome work. I'm personally very excited for faster models/inference.

I think speed is underrated to some degree in the current conversation. For a while, I was using Cursor's Composer quite a lot, even over frontier models, just because of how darn fast it was.

arw0n•15 minutes ago
What do you need speed for? That's a genuine question, I feel like the limiting factor already is my creativity, attention span and budget. And I'm not even yet optimizing cost by batching things like review to slow local models over night, or schedule tasks to take full advantage of my subscriptions.
kilroy123•about 1 hour ago
I've been using DeepSeek flash a lot this week to try it out. Now, I deeply want the smart frontier models to be just as fast.
GodelNumbering•about 1 hour ago
The corresponding OpenAI post https://openai.com/index/previewing-ultrafast/

There is no pricing info, which could mean it's "if you have to ask..." territory or they are simply gauging interest before deciding

rirze•about 1 hour ago
They're expanding access to companies that apply for the program and explain their use cases. So it's very real but limited imo.
WarmWash•33 minutes ago
The stake in the side of cerebras has always been that the economics are pretty poor.

Who knows if they will subsidizes it to mitigate sticker shock, but it's a safe assumption that it will be scarily expensive. However if you are in a "cost is no obstacle, speed is god" position, it will likely be pure magic.

fcarraldo•26 minutes ago
Can anyone explain why Cerberus needs to be _fast_ instead of _cheap_?

I don't think I understand why they aren't leveraging the increased speed to do batching to serve more customers at a "normal" tok/s.

Is the limitation, even on cerberus, still that the cache can only serve so many concurrent sessions over time? Is there no scaling advantage? I genuinely do not understand how any of this works.

thraway3837•about 1 hour ago
This is really cool. Someone here commented about similarity between this and hardware advancements for AV encode/decode.

I think it's only a matter of time before miniaturization can have a thumbnail sized user-replaceable accessory that contains the LLM built onto the hardware. I admit I don't know how any of that works, but would be amazing to experience. Fully local, fully offline, ultra fast local inference better than any personal computing product.

christkv•42 minutes ago
https://chatjimmy.ai/ Is that. Company behind it just got acquired by AMD
storus•38 minutes ago
Wow, that's even faster than diffusion LLMs but with the Fable-level quality! Congrats!
crazysim•about 1 hour ago
GPT 5.6 Luna Ultrafast when?
HawtAds•about 1 hour ago
Their dinner plate chips are impressive.
scotty79•38 minutes ago
I swear that now frontier AI stuff comes out few times a week.
poly2it•about 1 hour ago
I guess Gemini 3.7 Flash is no longer at the pareto frontier of speed to intelligence.
odo1242•about 1 hour ago
Well, there’s still price