Maybe AI agents can be lawyers after all
Last month, I wrote about Mercor’s new benchmark measuring AI agents’ capabilities on professional tasks like law and corporate analysis. At the time, the scores were pretty dismal, with every major lab scoring under 25%, so we concluded lawyers were safe from AI displacement, at least for now.
But AI capabilities can change a lot in a couple of weeks.
This week’s release of Anthropic’s Opus 4.6 shook up the leaderboards, with Anthropic’s new model scoring just shy of 30% in one-shot trials, and an average of 45% when given a few more cracks at the problem. Notably, the release included a bunch of new agentic features, including “agent swarms,” which may have helped with this kind of multistep problem-solving.
Regardless, the score is a huge jump from the previous state-of-the-art, and a sign that progress on foundation models isn’t slowing down. Mercor CEO Brendan Foody, who was particularly impressed, said, “jumping from 18.4% to 29.8% in a few months is insane.”

Thirty percent is still a long way from 100%, so it’s not like lawyers need to be worried about getting replaced by machines next week. But they should be a lot less confident than they were last month!
You May Also Like
Europa winners should not get Champions League - Wenger
Veed.io Review - The AI Video Editor That Streamlines Your Workflow
Converge Bio raises $25M, backed by Bessemer and execs from Meta, OpenAI, Wiz
AIMERS (And Former Spectrum) Member Eunjun Announces Military Enlistment
How Priceline Express Deals Work
Runway vs. Pika Labs: Which AI Video Tool is Right for You?
AI SRE Resolve AI confirms $125M raise, unicorn valuation
IVE's Wonyoung Embroiled In Poor Attitude Controversy In Viral Post
Top SM Entertainment Idols Caught In Viral Dating Rumors With Each Other
CAF confirms Egypt as Hosts of CAF Women’s Champions League 2025 with the Final Draw Scheduled for Monday, 27 October in Cairo
Weekly Forecast: November 3-7, 2025 – Taurus Full Moon