Back to News
Advertisement
Advertisement

⚡ Community Insights

Discussion Sentiment

74% Positive

Analyzed from 1382 words in the discussion.

Trending Topics

#better#human#science#training#spent#more#anthropic#claude#where#point

Discussion (9 Comments)Read Original on HackerNews

mccoyb•34 minutes ago
Notwithstanding the duplication of these posts across social media, OpenAI, and Anthropic ("it's all the model, they just tell it to keep going" if you beat anything with RL enough ... it's going to do the thing)

Here's my issue with this post:

> True to the spirit of the challenge, they didn’t use millions of dollars in computer power. They used Fable 5.1, working within Claude Science, a platform scientists can pay to use.

Okay, billions of dollars have been poured into these agentic LMs, right? Each training run to get the next increment is costing millions of dollars?

This feels like an obvious jab at Navier-Stokes, but where we get to shift the numbers around to hide where the compute actually is being spent ... compute is being spent. It's either being spent in amortization to make the search smarter ahead of time, during training, or its being spent after.

Also love: scientists get to pay Anthropic to work within their special science harness to do science. That's exactly what I dreamed of doing when I pursued physics in undergrad, one or two companies holding the keys to "progress" for a monthly subscription price.

ElevenLathe•21 minutes ago
> Also love: scientists get to pay Anthropic to work within their special science harness to do science. That's exactly what I dreamed of doing when I pursued physics in undergrad, one or two companies holding the keys to "progress" for a monthly subscription price.

I get this anxiety, and am largely an AI skeptic, but at one point there were only a handful of computers in the world too (same for batteries, or engines, or crucibles, or stills -- it goes way back), and the organizations that had them had a stranglehold on progress in the field, as did the small number of companies who knew how to make them. It got better as they got cheaper and more plentiful.

I guess my point is that there are more important anxieties to feed when it comes to LLMs and the current state of the world.

teiferer•1 minute ago
> I get this anxiety, and am largely an AI skeptic, but at one point there were only a handful of computers in the world too

I don't see how that makes the state of affairs any better.

tomrod•4 minutes ago
> Also love: scientists get to pay Anthropic to work within their special science harness to do science. That's exactly what I dreamed of doing when I pursued physics in undergrad, one or two companies holding the keys to "progress" for a monthly subscription price.

The moat is very limited. Harnesses aren't crazy hard to engineer. Open models are quite capable.

Majromax•21 minutes ago
> This feels like an obvious jab at Navier-Stokes, but where we get to shift the numbers around to hide where the compute actually is being spent ... compute is being spent. It's either being spent in amortization to make the search smarter ahead of time, during training, or its being spent after.

I think that argument is recursive? These posts aren't very complicated for either of us, but they're written on devices that are fabricated with billions of dollars of semiconductor equipment. At what point do we just acknowledge that we stand on the shoulders of giants?

To me, the distinguishing factor is that the expense not special-purpose but upfront. The model here is trained without foreknowledge of what problems it will solve. Solutions like nine loops are genuine expressions of a pre-existing model capability, even if that capability has not pre-existed for very long.

Glemmlko•18 minutes ago
Divide the compute of your argument (pre-training, training, post-training) through everything this model now can do and it will not look that bad at all.
wat10000•22 minutes ago
The training costs get rolled into the usage prices. It's not hiding anything to talk about the cost of usage without the cost of training, any more than I'd be hiding something by talking about the price of a $1,000 CPU without mentioning the tens of billions of dollars in R&D and infrastructure needed to create it.
sejje•20 minutes ago
What a weird thing to be butthurt about. You don't have to pay them. Do it the old fashioned way, with elbow grease and pots of coffee. Or invest in local AI.

Or embrace the future and realize that you couldn't imagine everything that would unfold, when you pursued your undergrad.

How you gonna get mad about all this?

skyberrys•29 minutes ago
What is it about telling Claude "I'm only a biological human and now I have to go to sleep" that will get it to keep cranking away for hours. I have good success with similar statements and usually come back off several hours to find that Claude has managed to babysit several computational runs.

It's interesting to note that the expensive part of this experiment was the Claude operations expense. For me I find that Claude is a small fraction of my cost with most of the bill attributed to computers to run simulations instead of the AI to monitor and tweak the simulations.

ball_of_lint•13 minutes ago
I think it's mostly just what it is on it's face; You're telling it to be independent for a while. The default prompts start in an interactive mode where it checks in and verifies requirements a lot.
6thbit•32 minutes ago

  > As it turned out, the result wasn’t all that far away for humans either. A few days after I heard from Anthropic, we heard from Song He, an amplitudeologist at the Chinese Academy of Sciences in Beijing. Song’s group had already gotten the majority of the result. They’d used some AI assistance, based on GPT-6, but not the kind of one-shot almost human-less approach Anthropic used.

Please have AI come up with something no human is also about to solve?

This gets me wondering why ai labs aren't proposing their own millenium prize type challenges.

qarl•18 minutes ago
> Please have AI come up with something no human is also about to solve?

No human-only was close to Navier-Stokes. The team that was close was also using AI.

6thbit•14 minutes ago
Fair point, and i think that's fine. AI labs could just not jump themselves into such problems and let humans have a go at them, whether ai-assisted or not. As in this case with a regular budget one researcher could've access to.
danvayn•about 1 hour ago
Cool article.

>While it’s possible that this is just a much more AI-friendly problem, I don’t think it’s just that: I think the technology has genuinely gotten better.

It has gotten better imo. The blog author mentions 2 cases at the beginning -- users who think that AI will be capped and those who think it will be uncapped. From my perspective, both are technically right -- AI is capped or technically has usually reached some sort of cap, until human innovation improves it. AI doesn't really improve itself on a grand scale so much as humans improve it.

In other words, AI can and does iteratively improve, but every single ceiling we've spotted and broken through so far came from human ingenuity or effort. It will likely continue to require it, regardless of how much it can do on it's own. In that regard, it seems as though all of this will inevitably be "uncapped, until it reaches a cap, and then likely it will eventually be uncapped by humans (again)". Because of this, AI will never perfectly fit neatly into an 'uncapped' or 'capped' bucket, as long as time continues moving and we continue solving issues as they crop up.

utopiah•9 minutes ago
I'm an AI skeptic, or even anti-AI, and even I would have a hard time to admit it hasn't gotten any better.

In fact I will even admit it, yes AI has gotten better... but also how could it NOT? It has literally all the resources it can has, namely attention from everyone, everybody talks about it, a lot of workers, from state of the art researchers to annotators, all the material resources, from dedicated chips to networking to water to electricity, the entire available dataset of Human written and said thoughts, literally everything published and thus categorized.

AI literally has everything humanity can provide, it better be "better".

teiferer•5 minutes ago
> The blog author

I also liked the article and the author before this post. But let's be honest that this was paid work as part of Anthropic's PR campaign, not just a random blog post.

CuriouslyC•19 minutes ago
AI can totally identify optimizations within a paradigm without human intervention. What it isn't good at now is deciding to try totally new paradigms. Humans aren't great at this either, though it's more cultural in our case.
ipsod•30 minutes ago
I feel like usage of AI alone should be enough to uncap it, so long as AI companies do the predictable thing and steal everything from everyone.

There are thousands of people figuring out how to make AI better for their particular thing. Way more than that walking the AI through taking things from problem to solution.

parl_match•34 minutes ago
> I’ve heard from smart, well-informed people who are confident that AI is a few years away from superintelligence, and that superintelligence will be capable of truly terrifying things. And I’ve heard from smart, well-informed people who are equally confident that LLM-based AI is close to a ceiling, that models like Claude won’t even be able to do impressive work in physics, let alone conquer the world.

This is framed as an "or", as if they're contradictory.

IMO, Both of these statements are true.

ryeights•14 minutes ago
i.e., we're entering RSI and LLMs will help to build their more capable successors under a new paradigm? Fair point, but most people saying #2 surely mean to say LLM development has been a dead end and hasn't gotten us substantially closer to AGI
parl_match•10 minutes ago
normally i'd agree, for "most people"

but the author explicitly distinguished between "AI" and "LLM-based AI", and they work for an AI company, where making these distinctions are really important

mrkn1•about 1 hour ago
I'm sure there's a simple answer to this, and I'm probably just missing it, but what happened with the supergravity one?

And how many runs did it take before this one? They say "in one shot" with nothing more than "keep going." But we only see the successful run, reported by the people who ran it.

atemerev•38 minutes ago
Just recently I used Fable 5.1 and Astra to solve and prove a long-standing mathematical problem I was always interested in: the shoreline search problem (a ship in total fog is at unknown distance from the shore [infinite line], what is the optimal trajectory?) A particular kind of logarithmic spiral was conjectured 33 years ago; a few days ago, I have obtained the proof. How routine it has become.

https://arxiv.org/abs/2609.24454

skyberrys•27 minutes ago
Cool, that was an interesting read on its own. I viewed the html version and skimmed it, I appreciated your illustrations.
onlyrealcuzzo•26 minutes ago
What if the shore isn't a clean line?
atemerev•6 minutes ago
saadn92•40 minutes ago
AI can now do really hard science homework all by itself without messing it up, which is kinda cool.