
Infrastructure
The hardware and platforms underneath enterprise AI: GPU procurement, cluster architecture, and the compute decisions that shape what an organization can build.


Microsoft AI’s MAI-Transcribe-2 undercuts OpenAI, Google and ElevenLabs on price and speed
That figure deserves a pause. When Microsoft AI shipped the first model in this line just five months ago, it charged $0.36 an hour. Thursday's early-bird price cuts that by roughly 72%. For an enterprise processing 100,000 hours of call-center audio a year — a modest volume for a large bank or telecom — the bill drops from $36,000 to $10,000. At that level, transcription stops being a line item anyone argues about.

Perplexity partners with Nvidia to launch Portable Computer, a fully local AI agent with zero token costs
The launch, developed in close partnership with Nvidia, is one of the most aggressive attempts yet to move serious AI agent workloads off the cloud and onto local devices. The model, the user's files, and the work itself can all stay on the machine. Work completed locally consumes no billing credits, and the company says every task starts on the device by default — with the system asking permission before sending any individual step to a more powerful frontier model in the cloud.

IBM’s next-gen mainframe chip is the first to run Arm and Z workloads on the same cores
The chip, which will power the next generation of IBM Z and LinuxONE systems, is the first dual-architecture mainframe processor ever built. It is designed to let enterprises run the vast and fast-growing ecosystem of Arm-native Linux software — including the AI frameworks that increasingly define modern infrastructure — directly alongside the z/OS transaction-processing workloads that anchor the world's banks, insurers, and governments.
Subscribe to get latest news!
Deep insights for enterprise AI, data, and security leaders

Serval’s super agent Catalyst creates roving background agents to identify and fix IT issues before they’re ticketed

Cursor launches Origin code hosting platform as GitHub outage exposes opening in AI coding race
The developer internet did what the developer internet does.

Mistral AI wants to build 1 gigawatt of European compute by 2030 — and lock in customers now

Partner Content
AI is exposing the limits of traditional network architecture

Bright Machines says its new hybrid robot cell could help solve a major AI infrastructure bottleneck

Microsoft launches new in-house AI models it says cut costs up to 89% versus OpenAI

Poolside drops Laguna S 2.1, an open-weight coding model that beats rivals 10x its size
The model, Laguna S 2.1, is a 118-billion-parameter Mixture-of-Experts (MoE) system that activates only 8 billion parameters per token, supports a context window of up to 1 million tokens, and — according to benchmarks published by the company — matches or beats open models several times its size on agentic coding tasks. The weights are available immediately on Hugging Face under the permissive OpenMDW-1.1 license.
