Back to News
Advertisement
Advertisement

⚡ Community Insights

Discussion Sentiment

69% Positive

Analyzed from 1388 words in the discussion.

Trending Topics

#code#something#model#debug#why#where#models#freedom#never#need

Discussion (21 Comments)Read Original on HackerNews

onion2k6 minutes ago
One of the first things I tell the junior/mid-level developers I mentor is "You can't debug something just by reading the code." We all have a mental model of how our code works, and it's usually a bit wrong. Bugs are the real world manifestations of those mistakes. When you read the code it's all filtered through your model, and that makes you blind to seeing why something unexpected happened. In order to debug something you have to be able to put the system in the state where the bug happens to see why it occurred.

LLMs generally only debug systems by reading the code with whatever information you give them in a prompt. The image in the article is meta-prompt - the prompt is whatever comes from the vision model the AI happens to use to 'understand' the red circle annotation. That won't work. To successfully debug what's going on it will need much better state information. Has the 'shelf' been explained to is? Is the contrast and lack of shadows in the image messing up the vision model? Why isn't the 'lid' in the image? And so on.

LLMs are clever but they're not magical. Treat them like a naive junior dev. Give them enough data about the state of something to understand it properly.

hspeiserabout 2 hours ago
I completely understand this. I’ve worked on robot hands and 5.6/Fable 5 were practically useless at helping me debug anything visually.

What I have found super useful actually is having models make a interactive 3d viewer in which I use move / highlight / paint (soft body painting directly onto the geometry for issues and different colors mean different failures). This gives a much better way to communicate the physical relationships and positions that are hard to get across in a labeled screenshot.

Its for sure still a lot of manual work so the "seeing-eye dog" description definitely holds. But I have found that after a couple of examples with the extra context the model gets much better at handling the problem and becomes useful.

hypfer32 minutes ago
I mean there's a reason why we're doing MoCap for video games. If computers were good at this, we wouldn't be needing that. But actual motion and all seems to be much more complex than the systems can predict, apparently.

Also.. uh.. isn't this.. good? I thought AI was to steal all our jobs.

___

Beside that, kinda weird self-description.

Isn't the computer executing your commands and you're just filling in where it cannot do that?

Being that dog implies that the computer is in the driver seat.

I mean it's supposed to be a joke I guess, but I read it as one that leaks internal metadata which seems to be incorrectly calibrated.

bitwize36 minutes ago
> All that’s left is the dumbest workflow possible: I fire up the debug viewer myself, look around for weird mistakes, then take a screenshot and tell the language model how badly it messed up this time. Eventually I just decided to do all the debugging work myself, so I would at least get to do the fun part too.

I call this "thanoscoding" for two reasons:

1. "Fine, I'll do it myself"

2. In the past I found I have to "snap away" the mess the LLM made in order to start afresh from a known good state (generally with git reset). But that was 1-2 generations ago when it comes to models. GPT6 Astra probably does things right the first time, 90% of the time.

SadErnabout 2 hours ago
I like to think of training and improving AIs as bringing freedom to the world. The useless toil and labor associated with rebuilding the same solutions into different contexts is finally at an end.

Coding was never the reward. Acting as a translator for a machine is far worse than allowing the machine to solve the mundane parts and leave you with bigger building blocks to play with.

After 25 years I have to confess I hated being a software engineer. It felt like grinding in a video game.

Now I can finally create and innovate at the speed of thought, and I'm very grateful to have this technology now.

anonzzzies42 minutes ago
For me code still is the reward; I made and make my fortunes with boring code people on HN say no one needs or wants, my hobby is, and has been for 45 years, writing, perfecting and optimizing code manually until I find it perfect. I have been working for 10 years on a programming language and OS (with niche 2 dbs that are now prod quality and we use) and those are the best moment of my day, I couldn’t care less what anyone thinks of it or ever uses it. But the lessons learned do flow into LLMs to write the boring code and I can say that our million$ paying LoB code can handle 100s req/s on shite hardware and even if the LLM did a crap job. Which is almost 100s times more than the client will ever needs.
burnoutdvabout 1 hour ago
You got nothing.

Giant corpos hold every sliver of your so called freedom, you dont innovate, you repeat what others created before you. You use a tool that shackles your thoughts and creativity in a never before seen way, what you perceive is an illusion of liberty that is no present. Without others you are nothing and they can take it away any seconds, you are an addict, not an innovator.

hypfer14 minutes ago
I can see where you're coming from, and I share the sentiment and resentment to a large degree, but I think you might be not doing reality justice.

It is true that closed weights models are a big issue. It is also true that LLM-generated solutions usually drift towards a median.

But that is not the dead end you think it might be.

Most coding work is repetitive boilerplate, and most typing is just.. well.. typing. Miserable work I too did not really enjoy. What I did enjoy were the end results, and that was just a necessary step to get there.

Now that's less the case than it was before we had LLMs, and for that I too am glad.

___

I think the article headline might've primed you (and me, fwiw) to reading the comment you're replying to as passive. But if you just look at the words of it, that might not actually be the case.

post-it33 minutes ago
The same can be said of stuff like cloud storage and email and social media. And yet here we are.
burnoutdv26 minutes ago
Your point is true to some extend, I personally host my own "cloud" storage..it comes with drawbacks, same is true for email..especially email, social media is a society constructs, by definition its reliant on others but many will argue that its not worth keeping anyway (although I wonder how one stays human in this world then, the meat space seems with so many barrieres in this age).

I feel very much like a luddite..or some artisan of a bygone age. The guy above laments how he never enjoyed coding, this is an interesting sentiment, I know quite a few people who enjoy the craft itself, myself included. The ability to form words that have meaning, that create something from nothing. Sure, the words are just a tool, but its something _I_ can master, not some abstract wish machine that may change its functionality tommorow. Obviously one can argue that the computer itself is in this case the one that creates and not me, its not magic that just works with me, but the machine I got obeys me and me alone..another reason why personal computing is important.

21asdffdsa1227 minutes ago
You lost leverage. Now be great full for the crumbs. If there remain any.
grebcabout 2 hours ago
Viva la revolution!

Serious question - what do you do for fun? I find fishing with friends enjoyable, and using my hands tidying up the old place I bought. It’s not innovating software I’ll never use but I’ll never tire of writing the same old ASP that delivers my clients the results they’re after.

altmanaltmanabout 1 hour ago
> I like to think of training and improving AIs as bringing freedom to the world.

Yes, when I think of OpenAI and Anthrophic and Google and Meta and any AI labs and their intentions, i cry a single tear for how these great instituitions are working so hard to bring freedom for humanity.

Walfabout 1 hour ago
Freedom is slavery!

AI owned by a few companies is more likely to put the majority right back to serfdom. You're a privileged fool to believe that freedom is the likely outcome of the current stampede.

With a username that appears to cheer for a sociopath¹, I doubt reason will convince you.

¹ https://futurism.com/artificial-intelligence/sources-sam-alt...

Barbing39 minutes ago
Their /s was implied

(Like, Meta’s intentions? lol)

flyinglizardabout 1 hour ago
I share the sentiment but the question here is how long can we maintain the balance point where the human in the agentic loop is required. It might be a window lasting only a few years, or for the foreseeable future; I think the answer lies the opaque compute economics of the frontier lab: how well models keep scaling and how economically sustainable is serving those models under the current market conditions.
post-it32 minutes ago
If my job gets automated, I'll find something else to do. I wouldn't have wanted lamplighters to succeed in preventing electrification, so it would be unfair for me to prevent the automation of my job if it can be done.
anonzzzies28 minutes ago
I use Astra to drive Fable; I drivel into my phone while walking in the forest and it builds. I don’t need to check; I do as clients need to pay, but it always is great. And it surprises me with things I did not know were possible even (never encountered them before so why would I know). We are at the point where our clients send voice messages and they get what they want without humans basically. This costs 10+10 max2 subs but that’s nothing compared to hiring people. We didn’t fire anyone; we just have 100+ more clients and make almost 50x more money. It’s boring but great as long as it lasts, we are already where you say for what enterprises generally need for the boring parts. That’s 99%.
flyinglizard20 minutes ago
I find I still add value, but I don't know if my value is real. Am I just biased and expect things to be done a certain way and penalize the model for doing something different? and am I providing the model with enough high quality context to align with my expectations in one-shot?
preommrabout 2 hours ago
These people need to take a vacation and come back in a few months when the vision models get better/cheaper.

Astra is already good at taking screenshots and acting on it (part of the agi claims).