Operations | Monitoring | ITSM | DevOps | Cloud

Shipped: Codex spend tied to the work behind it

People run Codex on their own laptops. When Codex is signed in with a ChatGPT subscription, OpenAI’s own admin console shows who used it and how much: messages and credits. What it doesn’t show is what any of that usage was for, or how it compares to what your team spent on other AI tools. The CloudZero desktop agent for macOS installs on a Mac, sees the traffic from AI coding tools, and prices what those tools use.

Your FY27 plan deserves a real AI number, not a hedge

Budget season is starting and most finance teams are finding the AI line is the most evasive line on the page. You lived through the year. AI spend came in higher than planned and moved in ways nobody could foresee or forecast. And when the board asked what it produced, the honest answer probably was “we’re working on it.”

Bringing Third-Party Apps into Harness AI Chat: Our MCP Gateway for Distributed Enterprise Systems | Harness Blog

TLDR: When you work in Harness AI Chat, your work doesn't stop at Harness. Your pipelines live here, but the change you actually need to make might be a YAML file in GitHub, a Jira ticket, or a Confluence doc. So we built an MCP Gateway inside Harness that lets AI Chat reach those third-party apps for you: safely, under Harness's own access controls and secrets, and without dropped sessions across our distributed fleet. This is the story of what we built and why.

Better context, smarter testing: How to give your AI coding agent direct access to k6 docs

As testing workflows become more AI-assisted, fast access to accurate documentation matters more than ever. Whether you're writing a new load test, troubleshooting an issue, or having an AI agent generate a script for you, you need reliable guidance that keeps pace with the way you work. But most documentation still lives in a browser. Every time you or your agent needs to verify an API or look up a best practice, you're forced to leave your terminal or editor and interrupt your workflow.

Monitoring Oracle ASM with Custom Metrics | The Tony and Tonie Show Ep 49

Even small Oracle ASM issues can become big database problems. Here's how to spot the warning signs early. Tony and Tonie discuss how Redgate Monitor custom metrics help teams close a common monitoring gap: surfacing Oracle ASM health and performance issues before storage pressure, rebalancing problems, or disk group failures become database incidents.

Creating Escalation and Regular Groups in OnPage

Learn how to create and configure an Escalation Group in OnPage with this step-by-step how-to guide. This video walks through how to create an escalation group, a regular group, configure escalation intervals and factors, enable Round Robin, set failover OPIDs, and add a Fail Report email address. With escalation groups, OnPage can route critical alerts to team members in a predefined order and automatically move to the next responder when needed—helping ensure time-sensitive notifications don’t go unanswered.

How to Create and Import Contacts in the NEW OnPage Web Console

A step-by-step guide for OnPage’s new web management console, including how to create a single contact and how to create multiple contacts at once by importing an Excel spreadsheet. Feel free to comment below with any questions! Whether you’re in IT or healthcare, OnPage helps teams manage critical alerting and communication to ensure urgent messages reach the right people at the right time. If you’re not yet using OnPage or want to see how it works, request a demo or speak with a member of our team to learn more.

Creating an AfterHours OnCall Schedule

Learn how to create and configure an **on-call schedule in OnPage** with this step-by-step how-to guide. In this video, we walk through how to select a group, create a new schedule, assign on-call team members, set escalation priority, and configure coverage for specific days and times. Using OnPage’s on-call scheduling capabilities, teams can ensure the right responders are automatically available for critical alerts, incidents, messages or calls during designated coverage periods, including after-hours, weekends, and other shifts.

What an AI SRE agent actually finds when you point it at a broken Kubernetes cluster

‍ Most of the AI features that shipped into observability tools this year summarize alerts. You get a paragraph that restates the dashboard you were already looking at, and the agent never reads the cluster itself, because giving it cluster access is a security conversation nobody wanted to start. This walkthrough starts it.

Builder in the loop: what production agents were missing before AURA

Builder in the loop is a Mezmo interview series with the engineers, product leaders, and operators shaping AURA. Each installment looks past the product layer to explore the decisions, tradeoffs, and lessons involved in building agents for real production work. This installment features Mike Shearer, the engineer who built AURA and, until recently, its only developer. AI agents are easy to believe in when the task is small.