
GitHub’s HydraFusion cuts AI coding costs in every benchmark. It only matches quality in one.
The vendor's benchmark table doesn't fully back its own quality claims — and it's not the only AI company selling that story.
Sean Michael Kerner
Cohere Parse 5 loses the benchmark on points. It wins on cost per page.
Cohere's Parse 5 model balances cost and structure, offering $1.50 per 1,000 pages for efficient document parsing without sacrificing essential layout.
Sean Michael Kerner
Enterprises are overpaying for simple AI queries — Snowflake's gateway now auto-routes to cut costs up to 3x
A smaller model tries the task first and calls in a bigger one only if it can't finish the job — that's how Snowflake's new routing feature decides.
Sean Michael Kerner
Enterprises with AI context layers report agent failures at more than twice the rate of those without one
A governed context layer is supposed to stop AI agents from confidently getting things wrong. New VB Pulse research shows the opposite: it's what makes the failure visible.
Sean Michael Kerner
Nvidia's Switchyard router reshuffles AI models mid-task, cutting task costs to a third in its own tests
Nvidia is betting the real fix for AI agent costs isn't a cheaper model or a smarter router alone, but owning both under one open license.
Sean Michael Kerner
Tencent's Team Memory shares AI agent memory across a team — with no governance yet for when it's wrong
Most AI agent memory tools so far help one agent remember more in one session. Tencent's Team Memory is built for a whole team to share the same context — and inherit the same mistakes.
Sean Michael Kerner
57% of enterprises traced a wrong AI answer to missing business context — Credible bets portable, open-source semantic code beats proprietary metadata
Credible's platform runs on Malloy, an open-source semantic language built at Google — designed so enterprises can take their AI context with them, no matter what they buy next.
Sean Michael Kerner
Stop adding more GPUs: Weka's new storage platform reduces load by caching 100% of an AI model's pre-calculated tokens
As context windows and multi-turn interactions grow, so does the GPU compute wasted recalculating work a model has already done. Augmented Memory Grid, a NeuralMesh 6 feature built specifically for this, is Weka's answer.
Sean Michael Kerner
A single AI agent conversation can look perfect and still be broken, leaders from LangChain, Conviva and CoreWeave said at VB Transform 2026
One shopper's chat with an AI agent looked fine on its own. But that category of conversations needed 3x more clarifying questions than baseline — a signal no single trace can show.
Sean Michael Kerner
At VB Transform 2026, Zillow's engineering chief said AI ROI numbers only hold up if you measure before you build
Toby Roberts explains what Zillow got right before it ever touched an AI rollout — and why most enterprises skip that step.
Sean Michael Kerner
Intuit scrapped its own AI agent architecture twice in four months. At VB Transform 2026, its AI VP called that the fast path
Natural-language handoffs between agents kept compounding errors. Intuit's AI VP breaks down the 60-day rebuild that fixed it, and what convinced engineering to start over.
Sean Michael Kerner
Agents think in milliseconds, legacy infrastructure doesn't. LinkedIn, Walmart and Zendesk shared how they closed the gap at VB Transform 2026
LinkedIn, Walmart and Zendesk each hit a different wall scaling AI agents. Their infrastructure leads break down what it took to move past it.
Sean Michael Kerner