Operations | Monitoring | ITSM | DevOps | Cloud

Ticket Deflection Starts With Your Knowledge Base, Not Your Chatbot

Your service desk closed 400 tickets last month. Somewhere between a third and half of them were password resets, VPN questions, licence requests, and "how do I get access to the shared drive." Every one of those had a documented answer. Most of those answers were sitting in a knowledge base that the person raising the ticket either could not find or did not trust.

Why Is GPU Utilization Low During AI Training? 6 Bottlenecks to Check

You bought the GPUs to make AI training faster. So why are they sitting idle? When GPU utilization drops during a training run, the obvious answer is to blame the accelerator. Maybe the workload is too small. Maybe the GPU isn't powerful enough. Maybe it's time to add more hardware. But what if the GPU isn't the problem at all? A training workload is only as fast as the infrastructure feeding it.

Too Much Documentation Is Hurting Your Team

Everyone says "just write it up" — but too much documentation just creates sprawl, not clarity. Docs pile up, nobody maintains them, and a page last touched 8 years ago doesn't count as current. Documentation only works when it's built into the culture, not treated as a default. Watch the full IT Leadership Lab session — linked above.

Hyperping MCP: Run Incidents, Status Pages and Maintenance

Summarize with ChatGPT Claude The Hyperping MCP server now has 49 tools: 28 that read and 21 that write. An agent connected from Claude Code, Cursor, Codex or another MCP client could already manage monitors, publish a status page incident and schedule maintenance. It can now do most of the rest: declare an incident and page on-call, acknowledge and escalate it, correct what was posted on the status page, create and configure status pages, and reschedule, end or cancel maintenance.

From Meeting Rooms to the Contact Center Floor: How Krisp Assists Agents on Every Live Call

Contact centers are where companies do their talking at scale. Thousands of agents, dozens of live calls each per shift, and on the other end a stranger with a problem and a finite amount of patience. A ten-second lookup, repeated across a few million calls a year, is a budget line.

Autonomous IT operations: Scaling business without scaling IT complexity

Autonomous IT operations use AI, operational data, observability, and automation to enable IT environments to detect issues, understand their context, determine the appropriate response, and act with minimal human intervention. As businesses grow, IT environments rarely stay simple. More employees, endpoints, applications, and cloud services generate even more alerts, incidents, and operational work. The traditional model scales linearly: more environment means more manual effort.

Database governance in the AI era: Framework, risks and best practices

AI database governance can get overlooked when teams rush to connect AI tools to production data. IBM’s 2025 research found that 97% of organizations that reported a breach involving an AI model or application lacked proper AI access controls. The risk is easy to see. Give an AI agent too much access and it can expose or alter data in seconds. Feed it poor-quality data and it may produce a confident but incorrect answer.

Grafana Campfire - What's New in Grafana Git Sync - (Grafana Community Call - Sept 2026)

Join us for a follow-up Grafana Campfire to explore what we’ve built in Git Sync since December 2025. We’ll cover new and improved Git provider integrations and authentication options, smoother dashboard editing and pull request workflows, and more control over changes through signed commits, user attribution, and PR and commit conventions. We’ll also explore improvements to folder organization and permissions, READMEs alongside your dashboards, and clearer synchronization status.