RU version is available. Content is displayed in original English for accuracy.
Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
65% Positive
Analyzed from 3764 words in the discussion.
Trending Topics
#law#legal#case#more#lawyers#astra#models#openai#https#money

Discussion (159 Comments)Read Original on HackerNews
> The top is a three-way tie: Muse Spark 1.3 Max, Claude Opus 5, and Claude Fable 5.1 all reach 55.29% all-pass accuracy, a clear ~6-point step ahead of the next model. [Astra for Law reached 54.0%]
> Under partial-credit scoring, Claude Opus 5 reaches 90.58% weighted pass rate but 55.29% under strict all-pass grading, where every rubric check must pass. The gap shows models often get most of an answer right but fail on one or two required elements. [Astra for law reached 90.0%]
https://www.vals.ai/benchmarks/legal_research
Its just like code I suppose, if you can read and understand and validate, you can use it to scale and otherwise it could end up being a vibe effort.
> API customers including Harvey and Legora will be able to build on Astra for Law, bringing this intelligence into their own products and workflows.
In other words: "no, no, we're not eating our children to prep for the IPO. Don't worry."
OpenAI didn't need to name Legora and Harvey in the second paragraph of the launch post.
They are pre-empting the obvious interpretation of Astra for Law: that moving this far up the legal stack puts them in direct competition with their biggest legal AI customers.
“Don't worry, they can build on us” is a pretty conspicuous message to include on launch day.
They have clearly thought about some pessimistic outcomes.
This is everything OpenAI have to say about privacy in this announcement. No guarantees. No promises. Just a pinky-swear promise.
Anyone trusting them–or a lawyer who relies on them–for legal work deserves what they get.
Very much a work in progress, only federal and state so far, no municipal codes yet, and no case law yet. Big hole, I know. Also working on making the search ranking work better.
Found that to be very interesting
[0]https://www.technologyreview.com/2026/06/04/1138391/courts-c...
the cost of making a legal argument can collapse while the cost of reaching an enforceable, legitimate decision may go higher, which will gate the "justice" system even more.
https://www.ftrain.com/nanolaw
You dropped this /s
I don’t know in the USA but in France, if it’s deemed that you launched a lawsuit knowing very well it wouldn’t succeed, you are susceptible to get a 10k€ fine. Even jail in serious cases.
https://artificialanalysis.ai/models/gpt-6-astra?omniscience...
Don't be too sure about that. [0]
0: https://www.damiencharlotin.com/hallucinations/
I've talked to a lawyer about how they handle this. They do indeed double-check everything, since it'd be embarrassing (or worse) to send hallucinated statements to opposing council or to the court. They still find the assembly a huge time saver
But based on stories in the news on the subject, not everyone has this same level of diligence
I have had some success using frontier models from the last 6ish months, but only when I can break up my work into discrete and verifiable tasks. For example, I had ~15k pages of discovery I needed to dig through for a summary judgment motion. Instead of just asking Claude to find the best evidence, I asked it first to run a clean, high quality OCR pass (it was almost entirely PDFs). Then I had it generate embeddings and write some reusable python scripts to make keyword and semantic searching easy for agents. While I was writing the brief, I would routinely ask my agent (Claude Code) to use both keyword and semantic searching to find the best evidence supporting whatever assertion I was trying to make. I trusted it because there were traces I could follow.
In other cases/situations, I’ve tried just giving a model access to all the docs and saying “write a brief arguing X,” but it’s always terrible at this. It writes briefs with lots of evocative jargon and rhetorical flourish, but a low signal-to-noise ratio.
Again, I’m sure others’ experiences differ based on workflow, legal area, etc.
You. Don't take legal advice from a word calculator.
That will be a decacorn product or more.
https://commonpaper.com/standards
I don't think you can. What I get from this article is that this is not a product they're going to sell to average consumers.
And yes, even contracts drafted for millions of $ have oversights and unlawful or unenforceable terms.
Then again, nobody will have money to buy anything at this rate, so in all liklihood, this is a total non-issue.
It’s cleaner.
People confuse slop with "bad", but slop isn't bad per se, it only becomes bad when real effort was required.
they need to pick a lane and optimize for it. coz at their size they can't serve the application layer (a.i startups who can fine-tune models will eat their lunch)
if they gonna do a consumer play - then go ham on that.
otherwise they're gonna get caught in the dreaded middle valley.
Provisioning and contracts and data retention was just an extension to review of existing ones.
Nobody serious is going to risk sending sensible data to OpenAI/Anthropic, etc because "the benchmarks have shown +8% performance there and +2% there". Irrelevant.
You can afford to play silly buggers when you have dumpster trucks of money backing up to your door every day see also: Meta.
OpenAI doesn't have any of these things. They have products that they're paying for customers when the market they're in is rapidly converging on fighting for API reasoning as part of enterprise systems and fighting a race to the bottom for fickle consumer solutions that will be eaten by open source once they have to make money.
Maybe they had a brief window for dominance of information search (or maybe it was only ever going to last as long as Google releasing all their internal research) and maybe they had a brief moment of monopoly till Anthropic got going but theyre not in the same dominance position as Google.
Qwen 3.8 Max and Opus 4.8 score highest.
This was already always the case. If anything, making this more accessible will reduce the barrier to entry for whether or not it's worth your time to take on a case. Instead of 50 lawyers spending 100s of hours on a case, you can have 1 or 2 lawyers + Astra working on it and if there's a case you can add more real lawyers.
1) Lawyers are not as naive as software engineers and will fight being replaces by new laws.
2) If they are replaced, OpenAI will take a cut commensurate with the amount in dispute (OAI, please credit me for the idea in the IPO brochure).
I am curious what level of trust established law firms treat LLMs with.
It's analogous to crypto. Started from some noble anti-authoritarian ideas and morphed into machine that removes any friction for capital - whoever has the most money will keep gaining the most.
Hence the crypto analogy - it was also supposed to "democratize", but the opposite happaned - it only further empowered the most powerful. Imagine legal case so purposefully complex that only those with access to best models have chances to participate and win the dispute.
Is the play here a set of specialized harnesses using their best general model?
https://aeon.co/essays/what-made-law-into-a-white-collar-swe...
In the real world, lawyers submit detailed bills and their clients examine them. If you don’t, that’s on you.
As much as I hate to see it. They are now threatening industries like Engineers, Game Developers, Accountants, 3D modelers, 3D animators, Video Production, Audio Production, Therapist, Tax Auditors, Journalists, Authors, Artists, Mathematicians, Product managers, Every type of analyst and pretty much any other job that can be done behind a computer screen.
We have big problems for humanity.
Real estate, food, energy, mobility.
That will be done with money. Or violence. Either way, a scary future to people when labor doesn't provide any value. Elon promises abundance, but what can he do against greed?
The biggest problem is that we're conditioned by a paradigm that frames these as problems.
If the benefits were shared across humanity, that could bring us closer to utopia. My worry is that we’ll instead end up with a handful of even wealthier billionaires and millions of people out of work.
Effectively yes, in the current forms. Those professions will likely evolve, but the traditional forms (ie writing code by hand, writing law filings by hand etc) are all dead.
there will still be writing initial and incremental prompts by hands, until and if LLMs surpass humans in all intellectual functions.
this is more likely to democratize the legal system by reducing the cost of a good legal team
As it stands, it seems far more likely to result in a wonderful life for a few, and an absolute catastrophe for most.
Same argument for plumbers. Everyone always jokes about what a good time it is to be a plumber. But what happens when all the software engineers turn to plumbing? Suddenly it's not such a good time to be a plumber anymore.
They also openly tell you what they are afraid of btw: collective worker power. something that is massively lacking in our industry, although i feel like it would be one of the easiest industries to unionize in terms of # of workers.
Interesting interview I just watched about how powerful and dangerous these "wishes" or "prophecies" are especially in the hands of the ultra-wealthy: https://www.youtube.com/watch?v=eR7grHa1NR0
If you automate lawyers out of a job, you can absolutely automate lawmakers out of jobs next. (Not that this would be a bad thing? Maybe pervasive agents for everyone can be the gateway drug to a "this time it's different!" workable direct democracy)