Operations | Monitoring | ITSM | DevOps | Cloud

The latest News and Information on Incident Management, On-Call, Incident Response and related technologies.

Grafana & PagerDuty: Automate incident management with ServiceDesk Plus Cloud

������ �������� ������������ ���� ��������! This month, we're bringing you two new ManageEngine Marketplace extensions for ServiceDesk Plus Cloud that help bridge the gap between your monitoring tools and your service desk. With the Grafana extension, alerts automatically create and resolve tickets in ServiceDesk Plus Cloud—eliminating manual ticket creation and ensuring incidents are tracked the moment they're detected.

Office Relocation as an Ops Project: Runbooks, Rollback Plans, and Zero-Downtime Moves

Engineering teams that would never push a config change to production without review will happily move their entire company to a new building on the strength of a shared spreadsheet and a group chat. Then the first Monday in the new office arrives: the ISP install slipped two weeks, the badge system doesn't talk to the identity provider, on-call is paging someone whose desk is in a moving box, and the conference room where the incident bridge usually happens no longer exists.

ServiceNow Runs Your IT. PagerDuty Makes Sure It Never Stops.

For most enterprises, ServiceNow has become the backbone of IT operations, the platform where workflows are governed, compliance is maintained, and every incident, change, and request is tracked from start to finish. If you’re running ServiceNow, you’ve made a serious investment in how your IT operates. PagerDuty is built to make that investment work even harder.

Make the most of shift-based schedules

We recently updated our Schedules to better reflect how teams are currently managing their on-call responsibilities. Not everyone is working on weekly shifts or providing 24×7 coverage for all of their services, and that should be easy to schedule in our new tooling. To give you some examples, I’ve gone back through some of the questions we’ve gotten on the PagerDuty Commons over the past couple of years for questions about custom schedules that we weren’t really thinking about.

AI on AI Challenges

Building AI agents is easy until they launch into production and start behaving unpredictably. In this presentation, João Freitas, Chief AI Officer at PagerDuty, dives into the messy reality of scaling non-deterministic systems and shares how PagerDuty manages multi-agent complexities. Speaker: João Freitas, Chief AI Officer, PagerDuty Recorded during GenAI Community x Google Developer Group Lisbon at PagerDuty Portugal offices, July 2026.

Why Faster Recovery Beats Faster Shipping in the AI Era

A year ago, AI coding tools worked alongside developers—suggesting the next line, completing a function, accelerating work that a human was already doing. Today, they’re writing entire modules and services independently, producing code that no human has reviewed line by line, built from components that no single person has fully mapped. And adoption is only accelerating: According to our recent AI Resilience Survey, 84% of organizations are now using AI to write, review, or suggest code.

Why Modern IT Incident Response Needs Social Sentiment Analysis

IT operations teams face an ongoing battle against alert fatigue. Despite running sophisticated telemetry and baseline Application Performance Monitoring, engineers are often bombarded with notifications that lead nowhere. Relying purely on internal dashboards creates a massive visibility gap, and when critical incidents slip through the cracks, the financial damage is swift and severe. To close this gap, DevOps professionals are increasingly looking beyond traditional server metrics and turning to a surprising source for early warning signals: public social sentiment.