Operations | Monitoring | ITSM | DevOps | Cloud

The latest News and Information on DevOps, CI/CD, Automation and related technologies.

Start Kubernetes monitoring in 5 minutes with Netdata

While Kubernetes (k8s) might simplify the way you deploy, scale, and load-balance your applications, not all clusters come with "batteries included" when it comes to monitoring. Doubly so for a monitoring stack that helps you actively troubleshoot issues with your cluster. You need robust Kubernetes monitoring, but you don’t want to spend a week setting it up, much less a single valuable day.

Challenges of growth engineering in a DevOps company

Growth engineering is a practice in which product, engineering, and design support a company’s growth efforts from within the product itself. Growth engineering has gained traction in consumer-facing companies. This practice has gained plenty of traction in the SaaS world over the last decade, to help support growth of self-serve users who often purchase services without the involvement of traditional sales teams.

Introducing Incident Timer

We’re excited to announce Incident Timer - a “days without an incident” timer for software teams to keep track of major engineering incidents. As the people behind Spike.sh, we keep discussing how to build a culture of reliability with our customers. We loved the idea of safety/accident timers in factories which kept track of major accidents. It's a simple and elegant way to keep safety on everybody’s minds.

IT Ops tax: Death by a thousand cuts

There are many hidden costs in running sub-optimal IT operations, that most organizations don’t consider. Enterprises often look at service downtime as their only KPI, but that is really only the tip of the iceberg. Without a properly operating incident management lifecycle, enterprises tend to support poorly performing services instead of fixing them.

What is DevOps?

What is DevOps? DevOps is a term for a cluster of concepts that has become a movement, “a cross-disciplinary practice dedicated to the study of building, evolving and operating, rapidly-changing resilient systems at scale.” (Jez Humble) The definition of DevOps is not agreed upon by everyone because of the complex processes attached to the term, however, the benefits to teams are universally agreed upon.

Why Finance Teams Love CloudZero (Even if It's Built for Engineering)

CloudZero is a platform that helps you understand cost — but that doesn't mean it's purely a finance tool. In fact, unlike most other cloud cost management and optimization solutions, it’s built for engineering. However, CloudZero still makes a lot of finance teams very happy. First of all, the work that engineering teams do while using CloudZero saves money, which every finance team appreciates.

Canonical completes Azure Arc Validation Program, helps increase user confidence in Arc enabled production Kubernetes

Microsoft Azure has just announced the details of its new Azure Arc Validation Program, aiming to further increase customer confidence in deploying Arc enabled Kubernetes in production workloads, and at scale.

How Grafana and Prometheus work together

Let us get an insight on how Grafana and Prometheus work together for monitoring metrics. Application monitoring is a crucial feature for any successful software offering. Application monitoring in its simplest form refers to collecting metrics on an application and using those metrics to gain an insight to improve the performance and efficiency of the application. Think of it as a cycle. Grafana and Prometheus are probably the most prominent tools in the application monitoring and analytics space.