Infrastructure

The hardware and platforms underneath enterprise AI: GPU procurement, cluster architecture, and the compute decisions that shape what an organization can build.

Nuneybits Vector art of retro desktop glowing Windows CRT trans 6b3a9b28-b30e-446d-b9ec-ec5d8a85e9cd

Microsoft AI’s MAI-Transcribe-2 undercuts OpenAI, Google and ElevenLabs on price and speed

That figure deserves a pause. When Microsoft AI shipped the first model in this line just five months ago, it charged $0.36 an hour. Thursday's early-bird price cuts that by roughly 72%. For an enterprise processing 100,000 hours of call-center audio a year — a modest volume for a large bank or telecom — the bill drops from $36,000 to $10,000. At that level, transcription stops being a line item anyone argues about.

Nuneybits Vector art of cloud dissolving into Nvidia green pixe 585344e3-0291-425f-9c47-895b98c353ff

Perplexity partners with Nvidia to launch Portable Computer, a fully local AI agent with zero token costs

The launch, developed in close partnership with Nvidia, is one of the most aggressive attempts yet to move serious AI agent workloads off the cloud and onto local devices. The model, the user's files, and the work itself can all stay on the machine. Work completed locally consumes no billing credits, and the company says every task starts on the device by default — with the system asking permission before sending any individual step to a more powerful frontier model in the cloud.

Nuneybits Vector art of 1960s mainframe shaking hands with Arm b5beb6f8-52b1-484f-8c33-30f6ca975174

IBM’s next-gen mainframe chip is the first to run Arm and Z workloads on the same cores

The chip, which will power the next generation of IBM Z and LinuxONE systems, is the first dual-architecture mainframe processor ever built. It is designed to let enterprises run the vast and fast-growing ecosystem of Arm-native Linux software — including the AI frameworks that increasingly define modern infrastructure — directly alongside the z/OS transaction-processing workloads that anchor the world's banks, insurers, and governments.

Subscribe to get latest news!

Deep insights for enterprise AI, data, and security leaders

By submitting your email, you agree to our Terms and Privacy Notice.

Nuneybits Vector art of minimalist lines clean villas overlooki 8da15709-c192-4383-b807-87d0dbfcf8e4

Poolside drops Laguna S 2.1, an open-weight coding model that beats rivals 10x its size

The model, Laguna S 2.1, is a 118-billion-parameter Mixture-of-Experts (MoE) system that activates only 8 billion parameters per token, supports a context window of up to 1 million tokens, and — according to benchmarks published by the company — matches or beats open models several times its size on agentic coding tasks. The weights are available immediately on Hugging Face under the permissive OpenMDW-1.1 license.