We’ll see it more and more Btw Harvey Legal Benchmark is very broad more than « super focused ». 1,250 legal tasks, 24 legal practice areas, it was made to be one of the first large-scale realistic legal benchmark
MARKET TRENDS
-
Agent economy’s accelerating growth surpasses even inflated expectations
By
–
The agent economy is growing faster than expected, even when you take into account that it will grow faster than expected.
-

This alone convinces enterprises to host LLMs on-premise
By
–
I mean, look at this, this alone is enough to get every enterprise out there into hosting their LLMs on-premise
-
Shadow AI in Banking: Governance Lags, Forbes Highlights Dangers, SR 11-7 Excludes Gen AI.
By
–
Shadow AI is already inside your bank: in treasury, payments, correspondent banking. The governance to contain it isn't. A new Forbes piece captures exactly why that gap is dangerous. And SR 11-7's update carved gen AI out entirely. The clock is running.
-
AI agents generate 450% more traffic; enterprise networks unprepared
By
–
AI agents generate 450% more network traffic than humans.
— Helen Yu (@YuHelenYu) 3 juin 2026
Multiply that by always-on, machine-speed execution.
Most enterprise networks were never built for this.@Cisco declared that the Networking SuperCycle is here. And the agent swarm already has an off switch.#CiscoLive… pic.twitter.com/SfCusDFKlPAI agents generate 450% more network traffic than humans. Multiply that by always-on, machine-speed execution. Most enterprise networks were never built for this. @Cisco declared that the Networking SuperCycle is here. And the agent swarm already has an off switch. #CiscoLive
-

World’s first heterogeneous disaggregated inference cloud shown live at ComputeX
By
–
The world's first heterogenous disaggregated inference cloud was just shown running live at ComputeX. VC2 — backed by a $3.5B compute commitment to SambaNova from @Vista_Equity & @cambiumcapital — brings three chips together in production for the first time:
– NVIDIA B200 GPUs -

Scaling PEFT: Towards Million Personal Models of Trillion Parameters
By
–
"On the Scaling of PEFT: Towards Million Personal Models of Trillion Parameters" Right now LLM personalization mostly means prompts, memory, or retrieval on top of one shared assistant. This paper instead keeps one trillion-parameter base model shared, and give each user a tiny
-
Top Chinese AI labs Moonshot and DeepSeek dominate
By
–
S-Tier Chinese Labs: Moonshot and DeepSeek These 2 are levels above everyone else
-
Gary Marcus argues LLM token prices will decrease, not increase
By
–
This is really intellectually dishonest. I have not been arguing that LLM token prices are increasing (though the all you can eat buffet is over), I have been arguing the *opposite*, viz that they will go down (e.g., in my recent tweet on commodity pricing that got 1 million
-

Costs matter: Uber caps tokens at $1500 per dev per month
By
–
we are seeing costs start to matter! uber just set limits of $1500 in tokens per developer per month i think we're going to start seeing more of this, and LangSmith Gateway is a great way to implement it