Operations | Monitoring | ITSM | DevOps | Cloud

AI cost reduction: tactics that preserve performance

AI cost reduction means lowering what you spend to run AI (tokens, inference, and compute) without sacrificing quality. The highest-leverage tactics, prompt caching, batching, and routing easy work to smaller models, cut spend 50 to 90% by removing waste, not capability. Somewhere right now, a finance leader is opening an AI bill that has quietly tripled, with no new product to show for it. Nobody approved it. No single decision caused it.

Get Ship Done: Everything We Shipped in July 2026 | Harness Blog

Harness shipped 71 features in July, about one every 10 hours. That's more than June's 62, and the surge lines up with what AI is doing to the rest of the SDLC: coding agents are writing more of the code, test agents are now generating and running more of the tests by default, and every stage downstream: deployment, security, cost, and resilience has to absorb that pace without falling over.

Control Runtime Behavior with Config Management | Harness Blog

As organizations ship software faster than ever, runtime behavior changes are becoming just as frequent as code releases. Teams need a way to update application behavior without waiting for code deployments while maintaining visibility, governance, and control. ‍ Now available in beta, Config Management provides a governed runtime control plane that separates runtime configuration from application deployments, enabling organizations to deliver configuration changes instantly across environments.

DEX Data Is Too Valuable to Limit to IT

For many organizations, digital employee experience (DEX) is still viewed as an IT project. It measures endpoint health, identifies performance issues, and helps service desks resolve incidents faster. While that’s useful, it’s also far too small a vision. In order to get the greatest return from DEX, organizations have to stop treating it as exclusively an engineering capability and started treating it as an equally powerful intelligence capability.

Progress WhatsUp Gold Recognized as a SPARK Matrix Leader in Network Observability

We’re proud to share that the Progress WhatsUp Gold solution has been recognized as a Leader in the QKS Group SPARK Matrix: Network Observability report, ahead of other vendors such as SolarWinds, Paessler PRTG and LogicMonitor to name a few. The recognition highlights the WhatsUp Gold network monitoring capabilities that help organizations gain deeper visibility into complex network environments while delivering impactful benefits for customers.

Straight from Support: AI credits, student plans, and why your Mac fans are so loud

Every so often we sit down with someone from our support team and turn their week into a blog post. First up: Roberto Vizcarra, on four things generating tickets lately, AI credits, student plans, integrations, and Mac performance. Here’s what changed and what to do about it.

What Is Network Latency? Causes, How to Measure It, and Ways to Reduce It

Slow application complaints are among the hardest tickets in IT to close. The network gets blamed first; the dashboard shows nothing wrong, and the ticket bounces between teams for a week. Network latency sits at the centre of that argument more often than any other metric. Most dashboards report latency as a single average, and that average hides the slow requests people actually notice. A path can average 30 ms and still drop a call every ten minutes.

What Is a Vulnerability Scan? How It Works and What the Results Mean

How many machines in your environment are running software with a publicly documented security flaw right now? That figure comes from an asset inventory, and asset records age quickly once they are written. The gap is rarely about tooling budgets. Software inventory across a few hundred endpoints shifts every week, while the published catalogue of flaws in that software grows every single day. Manual inspection loses that race inside the first month.