Boston, MA, USA
2016
  |  By David Aponovich
Only 22% of finance leaders can tie their AI spend to a business outcome, according to CloudZero’s 2026 finance survey. When AI ROI falls short, it usually isn’t because a company is doing too much AI. It’s that no one is watching which model runs which task, and that one choice accounts for a large part of the cost. Here’s why it happens.
  |  By Djo Lopez
Model Rightsizer is an open-source Claude Code sub-agent from CloudZero that scores each task on capability need versus cost pressure, then routes it to the smallest model that can handle it. In its first week, it cut Opus spend 75% while shifting 234x more work to Sonnet. Every Claude Code agent you run has to answer a question it usually never gets asked: does this task need the smartest model available, or are you paying Fable prices to rename a variable across three files?
  |  By Keith MacKenzie
Proving AI business value means sorting every AI investment into one of four buckets - revenue growth, cost avoidance, productivity gain, or risk reduction - then tracking spend at the unit level (per feature, customer, or team) so each dollar has a traceable return. Most companies measure one bucket well and leave the rest unattributed. That gap is why the same AI deployment can look like a $40M win and a public reversal at the same time.
  |  By Kaitlin Woo
For many teams, Budgets is the first page they open in the morning, and it should tell them where they stand at a glance. That’s easier now, with a redesigned Budgets page and a way to set a budget without leaving the app.
  |  By Lyne Carolyne
Codex pricing runs six tiers, from free to $200 a month, but the sticker price is not your real bill. OpenAI Codex pricing 2026 charges by the token, not the plan, a change that took effect in April. Plus is $20, Pro starts at $100, and everything past that depends on how many files you let the agent read. Most Codex pricing guides hand you a price list and call it done. That is like pricing a taxi ride by the door handle. The meter is what matters, and OpenAI put a real one on Codex this year.
  |  By Lyne Carolyne
Claude Opus 5 launched July 24, 2026 at $5 per million input tokens and $25 per million output tokens, identical to Opus 4.8. It delivers near Claude Fable 5 performance at half Fable's price and is now the default model on Claude Max. New effort settings let teams trade capability for token savings, which means two teams on identical pricing can now run up very different bills. Finance teams, that last part is your problem. Anthropic has shipped a model that costs exactly what the old one cost.
  |  By Kevin Lamb
A savings number tells you money is on the table, but it doesn’t tell you whether the finding holds up, what it’s based on, or what to do next. In that gap, recommendations pile up unactioned. When you’re staring at thousands of them, a title and a dollar figure isn’t enough to decide which are safe to act on.
  |  By Kaitlin Woo
Creating an API key used to mean sorting through categories organized around our internal structure, not how you’d use them, so finding everything you needed for a specific job meant guessing, or having someone on our team walk you through it. Now you can tell what each permission actually does at a glance.
  |  By Keith MacKenzie
AI cost management software gives finance and engineering one view of AI spend, ROI, and cost per inference. Here is how to evaluate, choose, and buy the right AI tool in 2026.
  |  By Keith MacKenzie
The best AI cost management tools in 2026 are CloudZero (best overall for connecting AI and cloud spend to business outcomes), Langfuse (best open-source LLM tracker), Portkey (best LLM gateway with cost controls), Datadog LLM Observability (best for teams already on Datadog), and CAST AI (best for Kubernetes AI infrastructure). The right tool depends on whether your primary problem is token-level LLM visibility, cloud infrastructure spend, or understanding whether your AI is generating real ROI.
  |  By CloudZero
CloudZero unveils our new logo and brand.

With CloudZero you get insights about your applications and systems, helping you manage operations at a scale that you’ve never had before. Our platform provides you with insights about every piece of your system, including the real cost of resources, resource utilization, reserved capacity and cost center efficiency.

With the accurate and trusted data provided by CloudZero you can minimize or eliminate under utilized resources, visualize costs for easy comprehension and oversee the entire software lifecycle. Nothing is out of view when using CloudZero’s Observability platform. From regional views to individual resources, you have insights at every level to help you keep your systems running smoothly.

How do we do it?

  • Collect and Normalize: CloudZero’s platform starts by collecting the data from your CloudWatch, CloudTrail, VPC Flowlogs, Lambda Data Events and Billing Data from every AWS account you connect. This part of the platform is isolated in its own account for security and has read-only access to the accounts you connect.
  • Populate the Stream: All of the data collected is normalized and the events, resources, statistics and billing data are organized into data streams which allow our platform to perform real-time analytics on all the data collected.
  • Find Meaning: Our algorithms take in the normalized data and perform complex analytics sifting through all the data to filter noise and enhance signal. We use Machine Learning on a large scale to learn what is valuable to surface.
  • Visualize Everything: The application provides opinionated visualizations of the insights determined by the platform’s AI. From regional system maps to single resources to cost of service broken down by team, CloudZero’s platform provides true observability to everyone in your organization.

Observability for Everyone. Add cost as a first-class metric and understand the financial effect of operational decisions.