Operations | Monitoring | ITSM | DevOps | Cloud

How to Reduce Vendor Sprawl: A Practical Guide for IT Teams

To reduce vendor sprawl, start with inventorying what your tools do, who depends on them, and where their capabilities overlap. Then decide which tools to retain, connect, or replace. When going away from tool sprawl, you need to aim for a smaller toolset that your team can manage while still supporting the work your organization needs. This guide walks through spotting unnecessary fragmentation, assessing whether consolidating makes sense or not, and transitioning to it.

Log Analysis with Machine Learning: An Automated Approach to Analyzing Logs Using ML/AI

AI log analysis helps IT teams turn massive volumes of operational data into actionable insight. By applying statistical methods, machine learning (ML), semantic analysis, and generative AI, organizations can identify unusual behavior, connect related signals, and investigate probable root causes faster. But AI-generated answers should not be mistaken for proof.

How to build a Language Server Protocol (LSP) plugin for Claude Code

Language servers give editors structured, real-time feedback: diagnostics, hover docs, autocomplete, and other guidance that would otherwise surface later. Language servers already exist for many of the languages and tools developers use every day, but Claude Code doesn’t automatically receive their feedback.

Governance is the platform problem worth solving

Based on the LeadDev panel discussion "Governance Is the Platform Problem Worth Solving," hosted in partnership with Harness, August 5, 2026. AI agents are no longer waiting for a human to approve their next move. They open pull requests, adjust configurations, and act on behalf of the people who deployed them — often faster than any review cycle can keep up.

Platform engineering in the age of AI

94% of engineering leaders say their AI metrics are missing. Here's how platform engineering is changing to close that gap. Based on the InfoQ webinar "Platform Engineering in the Age of AI," featuring panelists from Harness, DKB, and Shine, August 18, 2026. 94% of engineering leaders say the AI metrics that matter most to them are missing.

How Businesses Can Improve Efficiency Through Better Financial Management

If you aspire to run a business that's efficient, effective, and ultimately profitable, then finance is something that you (quite literally) can't afford to neglect. During the early phases, many businesses understandable emphasise operational efficiency. But just as important as what you're actually doing is the way that you're allocating your resources. This is something that financial management will get under control for you.

Best AI Infrastructure Providers for Power, Cooling, and Compute

AI infrastructure is becoming a facilities problem as much as a compute problem. Adding accelerators is only useful when the surrounding environment can support them. Power has to reach the rack reliably. Cooling has to remove the heat produced under sustained load. The network fabric has to keep accelerators communicating. Storage has to feed the workload. Orchestration and monitoring then determine whether expensive capacity spends its time doing useful work.

SAP Observability Tools Compared

Comparisons of SAP observability tools often evaluate which platforms can see inside SAP.Today, that’s nearly all of them. Dynatrace, Datadog, New Relic and Splunk can all get SAP telemetry. None of them are likely the best choice for an SAP-centric application, and we will document why. The questions teams evaluating SAP observability solutions should consider: That last one is where most of these platforms stop, and it is the difference between observability and operations.

September 2026 at Bindplane: Bindplane Agent and a new Overview page

Also this month, a new BDOT 1.107.0 release, and you can start a configuration from a Full-Pipeline Blueprint. Here’s what happened in the last month. Prefer to watch? The September Community Call is streamed live on YouTube. Watch " YouTube" on YouTube Watch.

Automate Incident Management with PagerDuty Slack

Most organizations managing major incidents realized that every moment matters. Context-switching between different tools – with multiple web and chat surfaces having to be open to collect information causes friction. The cost of context-switching during an outage is even more painful. These are the problems that PagerDuty’s Slack Transformation just closed.