Operations | Monitoring | ITSM | DevOps | Cloud

The latest News and Information on Cost Management and related technologies.

Why engineers ignore cloud cost governance (and fixes)

Discover why engineers ignore cloud cost governance and how to build developer cost accountability. Learn how Harness helps empower engineering teams. Engineers often overlook cloud costs due to friction in traditional FinOps tools and a lack of real-time visibility. By embedding automated guardrails and shift-left cost insights into developer workflows, organizations can drive accountability without slowing velocity.

Shipped: Get alerted when AI spend spikes, with the cause attached

AI spend now comes from every department, and it can double in a week without anyone deciding it should. The invoice arrives after the month closes. By then the usual response is a spend cap, which slows every team, including the ones getting real work done with AI. You need to know about spend that breaks its normal pattern while there’s still time to act. The alert should reach the person who can act on it, with proper context and detail.

Shipped: One CloudZero for everyone, starting October 1

On June 3, we made the new CloudZero experience the default for every customer. Since then, we’ve shipped around 30 improvements a week: side-by-side period comparisons in Explorer, budgets you can create and edit right in the app, threshold alerts on dashboard tiles, and Monitors, which flags AI and cloud spend that moves outside its normal pattern and shows you what changed. Pages load 28 to 61% faster. JavaScript execution is 85% faster.

AI cost allocation: how to attribute AI spend by team, product, and customer

AI cost allocation is the practice of attributing every dollar of AI spend to the team, product, feature, or customer that generated it. That spend includes API tokens, GPU compute, per-seat tools, and shared infrastructure. It's harder than cloud allocation because AI spend arrives untagged, spans vendors, and pools in shared resources. Four methods cover most cases: tag-based, key-based attribution, proportional split, and usage-telemetry.

Why Engineers Ignore Cloud Cost Optimization & Fixes

Learn why engineers ignore cloud cost optimization and how to build a culture of FinOps governance. See how Harness helps. Engineers often overlook cloud costs due to lack of visibility, fragmented tooling, and competing delivery priorities. Organizations can fix this by embedding FinOps guardrails into developer workflows and providing real-time cost feedback during build cycles.

Shipped: Every AI provider, one cost story

If you were anywhere near LinkedIn last week, you probably saw us launch AI Signals. We weren’t exactly quiet about it. (Press release, a couple of blog posts, and more social posts than we’d like to admit. Sorry about your feed.) We covered the why behind AI Signals already, but I wanted to actually walk you through what you’re seeing on the screen. Sooner or later someone asks what the company spent on AI last month.

Shipped: Don't ask an AI agent what its work will cost

If you set the budget for your team’s AI agent work, or answer to someone who does, you need a rough idea of what a job will cost before it starts. That’s hard to get. Stanford researchers found the same agent, given the same task, can use up to 30 times more tokens from one run to the next, and you usually find out afterward. Most developers just run the job.

Cost per AI outcome: tying AI spend to results

Cost per AI outcome is your total attributed AI spend divided by the business results it produced: resolved tickets, converted leads, merged pull requests. It includes the cost of failed attempts, sits at the top of the AI unit-cost ladder, and it's the number that makes vendor outcome pricing, ROI claims, and build-versus-buy decisions comparable.

Shipped: A customer support experience that starts with an answer

When you have a question about your cloud or AI spend, you want an answer quickly, not a ticket that disappears into a queue. Support should not mean waiting for business hours, repeating your account details to multiple people, or wondering whether anyone picked up your message. That changed this week for every CloudZero customer. You get answers to most product and account questions immediately, at any hour, and when your question needs a person, they already have context.

We stopped asking an LLM how much its own work would cost

There’s a specific kind of measurement problem worth naming precisely rather than dramatizing: this month we found that our model-routing agent was assigning a token budget to every unit of work, and that budget was noise in the strict sense. Fixing it meant improving a system that’s mostly right, not tearing one down.