

Subscribe to get latest news!
Deep insights for enterprise AI, data, and security leaders
How enterprises coordinate and scale AI in production.

Long-running AI agents quietly drop compliance rules, and bigger context windows won't fix it
Imagine deploying an AI agent to run a multi-day master data validation workflow. By day three, it has ingested thousands of records. The baseline governance rules you hardcoded into the system prompt at the start? They are gone. Pushed right out of active memory.

Companies already run 3 agent platforms. Salesforce's new Enterprise AI Harness wants to govern all of them.

Most pipeline monitoring checks if the job ran. This one didn't check if the numbers were right.

Companies are spending millions rewiring how AI gets used. Almost none can prove it's working.
The hardware and platforms underneath enterprise AI.

Nvidia acquires Hugging Face after Stripe nabs OpenRouter: here's what open source AI builders should do

Microsoft AI’s MAI-Transcribe-2 undercuts OpenAI, Google and ElevenLabs on price and speed
That figure deserves a pause. When Microsoft AI shipped the first model in this line just five months ago, it charged $0.36 an hour. Thursday's early-bird price cuts that by roughly 72%. For an enterprise processing 100,000 hours of call-center audio a year — a modest volume for a large bank or telecom — the bill drops from $36,000 to $10,000. At that level, transcription stops being a line item anyone argues about.

Perplexity partners with Nvidia to launch Portable Computer, a fully local AI agent with zero token costs
The launch, developed in close partnership with Nvidia, is one of the most aggressive attempts yet to move serious AI agent workloads off the cloud and onto local devices. The model, the user's files, and the work itself can all stay on the machine. Work completed locally consumes no billing credits, and the company says every task starts on the device by default — with the system asking permission before sending any individual step to a more powerful frontier model in the cloud.

IBM’s next-gen mainframe chip is the first to run Arm and Z workloads on the same cores
The chip, which will power the next generation of IBM Z and LinuxONE systems, is the first dual-architecture mainframe processor ever built. It is designed to let enterprises run the vast and fast-growing ecosystem of Arm-native Linux software — including the AI frameworks that increasingly define modern infrastructure — directly alongside the z/OS transaction-processing workloads that anchor the world's banks, insurers, and governments.


Token Economics: Why Cost Kills More AI Projects Than Hallucinations

Ending the hallucination tax: Governing the backbone of enterprise AI

OpenAI launches ChatGPT for Financial Services with integrated data sources — it pulls research, cites it, and builds decks in minutes

Partner Content
Why the economics of enterprise AI favor dynamic model routing

OpenAI's new data agent skips the one thing rivals like Databricks are racing to publish: a benchmark

Databricks-trained AI agents match Claude and GPT-5.6 Luna's answer quality — in half the time

Partner Content
AI is changing the economics of software supply chain attacks

Partner Content
AI governance is moving to runtime — and regulated industries are getting there first

Why AI shouldn't be the one repairing your data pipelines
When a microservice fails in a cloud-native architecture, circuit breakers trigger, traffic reroutes, and Kubernetes spins up replacement pods within seconds. The system heals before end users even notice a blip.

Anthropic CEO says AI swarm could ‘take over the entire internet’ in 6-12 months, commits to AI slowdown plan
Enterprise tech coverage through the lens of AI adoption.

Amodei's AI slowdown plan never says open weights. It doesn't have to.

DeepSeek-V4.1-Flash debuts with $0.003/1M off-peak cached-input rate and benchmarks eclipsing GPT-5.6 Sol, Claude Opus 5

OpenAI solves longstanding math problem with 10,000-agent swarm — but can't rule out benefitting from a researcher's private Codex data

Partner Content
How European enterprises can meet sovereignty demands without giving up global reach
Presented by Equinix
Bleu Garde Announces Landmark EPA Registration for Arresta® at Builds Bio+ Philadelphia Life Science Symposium
LM-Kit Launches LM-Kit One, a Private AI Application Server for Organizations and Software Vendors
Quality Valve Announces Leadership Changes
Powerful Medical Receives FDA De Novo Approval for PMcardio "Queen of Hearts," Its AI ECG Model for Detecting Acute Heart Attacks
More

Security vendors use AI to rank Patch Tuesday CVEs — and rarely tell customers

Anthropic's safety monitor missed a live cyberattack because Mythos 5's reasoning said everything was fine

Agents identifying as OpenAI systems wrote 17,000 posts to a wiki no one was supposed to write to

Most security teams don't know how many AI agents they're running. Falcon Guardian found 18,000 at one company that had approved only 300.

MCP's new spec turns a planted prompt into a stolen credential
The Model Context Protocol's (MCP)'s largest revision since its initial launch shipped on July 28. By the end of the first day, all four Tier 1 SDKs were already speaking the new version, and Cloudflare's Agents SDK had support in place from day zero, with customers such as Sentry and Linear picking it up right away, meaning the surface this article describes is already live in production. A new 12-month deprecation policy keeps the changes in place through at least mid-2027. Most of the coverage has focused on what improvements have been made: A stateless core that scales on ordinary HTTP, OAuth-native authorization, and server-rendered UIs via MCP Apps.

My first VentureBeat story was the 'Woodstock of AI.' My last is Nvidia buying Hugging Face for $12.9 billion.
Today is my last day as Editorial Director at VentureBeat.

GitHub’s HydraFusion cuts AI coding costs in every benchmark. It only matches quality in one.




