Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
58% Positive
Analyzed from 1432 words in the discussion.
Trending Topics
#openai#don#agent#service#customer#code#company#maybe#product#trust
Discussion Sentiment
Analyzed from 1432 words in the discussion.
Trending Topics
Discussion (42 Comments)Read Original on HackerNews
If you call a bar in SF or a concierge in Vegas, you'll often get a voice chat agent. There's no terms of service you agree to. In aggregate, it's a pretty rich dataset on openai's end.
What happens to your voice, your queries, your PII?
This reads poorly to me. So they’ve proven the models CAN work, but they also say in the next line that they CANNOT do high value work in production.
No other comments from me but that first two sentence opener should have been massaged a bit. I’m sure they have a PR / messaging person (agent?) though, so maybe I’m reading too far into it.
AI support has been terrible so far in my experience and tends to over optimize for a happy path easily solved on a company website rather than solve the issue that made me call them on the phone that is likely a system bug or other issue.
This issue you mentioned of the in-built sycophancy being at odds with the customer service path, is the same reason I cant trust these sorts of systems in my own environment. I see the value of having a human staff the position of customer service being that they can often delineate the line between "good acceptable service" and "violates company policy" much more keenly than what ive seen from LLMs.
I guess my question is, what is a safe AI "process percentile" for a company that lets it recover if the AI goes down or goes haywire? No one knows (?), but it's worth thinking about.
I am also curious just how much direction the AI will need given the environment is dynamic. Humans generally align to the company's goals because they are rewarded (paid) to, and need to to survive. Relationships are built on unspoken and spoken communication. Unspoken could be cultural, hierarchical, environmental pressures, etc. As an extension of that, companies build relationships with other companies by sending (essentially) diplomats and ambassadors to each other (the article mentioned using AI for sales). -- All this to say: if you have to constantly explain the job to someone, they are not fit for the job.
We'll see, I guess.
Is this the correct reading? If so, I'm not that impressed.
Otherwise, I see this as potentially useful for firms that dont have global pressence and want to service a support system across multiple timezones where it might be difficult to staff a night-shift. However, issues of trust and quality will always be something to contend with here.
That’s the first I’ve heard of it. Has anyone tried it?
This has a lot of words with very little information. Having codex make changes to code because someone made a request to IT is insane, even if it needs approval.
Idk man, I feel like I am taking crazy pills.
I use codex everyday but I plan the work and codex performs the work bit by bit, which allows me to review every piece of code. I would not sleep well not knowing what is in production.
The problem with "someone makes a request, code changes happen automatically, and all someone else has to do for that to be committed is mash approve" doesn't strike me as a way to create a maintainable code-base.
The changes may even work, and may technically fulfill the request, but the agent cannot know the design intent or process intent. Today there may be a programmer sitting between the request and approval, but what about tomorrow? Might it be a non-technical or barely technical middle manager?
I don't want to laugh too hard this afternoon. But more seriously, I'm not really sure who this product is for in a way an existing tool can't handle? Like if a business really wants to go deep on agentic workflows for customer service, what is going to make them reach for this?
Is this some sort of joke? Like. That’s the proof. If it can do reliable work then that’s the proof. Contradiction in sentence one. If your pipeline breaks 80% of the time but sometimes the magic token lottery gives you something remotely useful, that’s NOT proof. The ceiling for these companies is so low it’s unbelievable. What happened to products that worked first, and worried about the rest later? Can we have an inspired ad article for that instead?
or again it's simply OpenAI looking for PMF to justify their valuations ?
I think it's the latter and will fail again just like Sora.
Did it use OpenAI presence for that?
I'd bet in 2 weeks nobody will remember this
Not sure if they're just trying to find what the abstractions are worth focusing, or if it will stay that way, but certainly cool to see how much they're trying out.
Makes sense to me. The cost of shipping products is lower than ever and they have lots of talented folks who want to ship things. No one knows what could be the next killer app.
However, reality is different and unpredictable.