
Monorepos make coding agents more useful — but compromised accounts are harder to contain
Hacktron reached OpenAI’s internal monorepo through a compromised Codex account. The bigger question: How much can one compromised account actually see?
Taryn Plumb
Inference chips are locked in years before the models they’ll run. Architect Labs says AI can design them in weeks
Two architects wrote the spec. The AI handled design, verification, firmware and kernels, and got each spec change back onto an FPGA in under 48 hours.
Taryn Plumb
Google's Dream-RSI cuts discovery-agent calls up to 162x by replaying searches it already ran
Teams tuning agent search loops pay twice for paths that already failed. Google researchers say the record of those failures is a simulator nobody was using.
Taryn Plumb
Google’s Gemini 3.8 Flash is built for agents, while its Cyber twin hunts vulnerabilities
Google's 3.8 Flash Cyber excels in vulnerability detection, offering cost-effective solutions for cybersecurity teams with advanced patching capabilities.
Taryn Plumb
One in five enterprises can't stop a runaway AI agent's spending in real time
Builders are running three AI orchestration platforms at once, driven by distrust in any single vendor's security controls. But a fifth still can't halt a runaway agent's spending.
Taryn Plumb
DeepSeek's top-ranked V4 Flash stumbles on real agent tasks as its prices surge
Analysts say the fix isn't avoiding DeepSeek entirely — it's knowing exactly which tasks to hand it and which to keep away from it.
Taryn Plumb
AI coding agents are blowing through budgets — Replit, Kilo Code, and Symbotic explain how they're managing it
One engineer's $600-a-day AI bill. A support automation that blew through the budget. Three companies share what they now track instead of raw spend.
Taryn Plumb
Enterprise AI agents can't talk to each other, can't be trusted with permissions, and can't be audited — 5 startups are already fixing that
One startup said it cut cyberattack containment time from seven hours to twelve minutes. See four other approaches to running AI agents safely at scale.
Taryn Plumb
Target SVP says its real AI moat isn't the models — it's everything built around them
One Target store's AI-driven inventory call looked like a mistake. Analysts let it run anyway — a small bet that's shaping how much autonomy AI earns next.
Taryn Plumb
Instacart's CTO says AI made the company stop worrying about tech debt
Instacart's engineers no longer read most of the code they ship. CTO Anirban Kundu explains why that's made tech debt a non-issue.
Taryn Plumb
Brex built its AI agent policy by watching what agents actually do, not by writing rules first
Brex's open-source proxy watches how AI agents actually behave, then drafts the security policy itself — leaving an LLM judge to review only the trickiest 3% of requests.
Taryn Plumb
The AI architecture that let Liberty Mutual shrug off the Fable 5 outage
How one 114-year-old insurer built an AI architecture flexible enough to shrug off a major model outage entirely.
Taryn Plumb