Operations | Monitoring | ITSM | DevOps | Cloud

The latest News and Information on Monitoring for Websites, Applications, APIs, Infrastructure, and other technologies.

How to use Grafana Assistant with the AWS CloudWatch data source

Grafana Assistant meets AWS CloudWatch! In this video, Staff Software Engineer Ivana Huckova shows how to use Grafana Assistant, the AI agent built into Grafana Cloud, with the Amazon CloudWatch data source. Watch her query CloudWatch metrics and logs in plain language, build dashboards in seconds, and troubleshoot AWS resources — no query syntax required.

New AI Features in Playwright (Live-Webinar)

An AI agent that can't open a browser is just guessing. Stefan from Checkly shows how giving AI coding agents a real browser via Playwright enables reliable end-to-end test generation and debugging, closing the quality gap created by faster, AI-driven shipping. The session compares Playwright MCP vs the Playwright CLI for agent workflows, showing that thanks to MCP spec changes, lazy tool loading, and skills, the two are now effectively just different interfaces to the same tool, with no real token advantage either way.

5 Key Differences Between Playwright and Puppeteer

Hyperping· Uptime monitoring Know before your customers do. Monitor from 18 regions and route alerts by phone, SMS, Slack or email. Start free Playwright and Puppeteer solve the same problem, driving a real browser from Node.js, and Playwright was started by engineers who previously built Puppeteer. That shared ancestry makes the two APIs look similar at first glance, which is exactly why the differences below catch people off guard.

Answer any cost question faster with the Cloud Cost skill in Bits Chat

Managing cloud, AI, and SaaS costs means answering a steady stream of questions from finance, leadership, and engineering teams. What changed? Which team owns the spend? Was an increase expected? Are we still on track against the budget? When each answer requires moving between dashboards, filtering cost data by team or service, or manually correlating billing data with observability data, it can slow down investigations while costs continue to rise.

Best Citrix Monitoring Tools for End-to-End Performance and User Experience

It is 9:02 on Monday morning. Within minutes, hundreds of employees hit the logon button at once, and the helpdesk queue fills with the same three words: Citrix is slow. The team scrambles. Is it the Delivery Controller? Active Directory? The profile server? Storage? Without the right Citrix monitoring tools, that question can take hours to answer, and every minute is a roomful of people who cannot work.

Choosing the Right Application Performance Monitoring Tool for Modern IT

It is 2 a.m. The pager goes off. Checkout is down, customers are dropping, and you have five dashboards open across four tools that do not agree. One says the application is fine. Another says latency is climbing. None tells you why. This is the moment that decides whether you have the right application performance monitoring tool. Modern applications run across containers, cloud services, APIs, and infrastructure spanning on-premises and multiple clouds.

The July 2026 AWS CloudFront Outage: VPC Origins, Cascade Impact, and What Broke

On July 16, 2026, AWS experienced a disruption in its CloudFront service, which affected a large number of websites and applications. The outage was caused by a configuration loading failure in CloudFront's VPC Origins feature. This was AWS's most widely-felt outage after last year's outage on October 20th, which caused widespread damage.

What Is Alert Fatigue, and Why Do IT Teams Miss Critical Alerts?

Alert fatigue is one of the biggest reasons critical incidents get missed. In this video, learn what alert fatigue is, why it happens, and how reducing noisy, repetitive notifications helps IT teams respond faster to the alerts that actually matter. Whether you're an IT operations professional, SRE, DevOps engineer, NOC analyst, or IT manager, this video explains alert fatigue in simple terms and shares practical ways to reduce alert noise, prioritize critical issues, and improve incident response.

What Is CIS Compliance? A Complete Guide

Most security breaches begin with something small and avoidable. Think of a password still set to its default, a server port left open to the internet, or configuration drift that slowly undoes last year's careful setup. Gaps like these go unnoticed because nobody is assigned to look for them, and they can be very costly. According to IBM's 2025 Cost of a Data Breach report, the average data breach incident costs companies 4.44 million dollars.