Operations | Monitoring | ITSM | DevOps | Cloud

The latest News and Information on DevOps, CI/CD, Automation and related technologies.

Introducing Infrastructure Knowledge: Teach Netdata AI What Your Metrics Can't Show

Netdata AI sees everything your infrastructure does: every metric, every anomaly, every alert. It does not see what your infrastructure is: which services matter, which host is supposed to run hot, who owns what, what your team considers normal. Without that context, “CPU at 91%” is just a finding. With it, it might be a machine doing exactly its job.

Monitor test health at a glance in Bitbucket Tests

When a team relies on automated tests in CI/CD, knowing that tests ran is only the beginning. Understanding whether the suite is healthy, which tests need attention, and how a specific test has behaved over time — that’s what drives action. Bitbucket Tests is evolving to make those answers easier to find and give you tools to improve your test health.

How to Build an HR PTO AI Agent with Resolve Agent Lab

See how to build an HR PTO agent with Resolve Agent Lab. In this Resolve Reels demo, we create a purpose-built AI agent by adding automation skills, instructions, conversation starters, and guardrails. The agent can answer PTO questions, check balances, account for calendar conflicts, and submit requests through systems like Workday or ADP. See how Resolve helps teams build AI agents that take action across enterprise systems.

Connect Codex to CircleCI: Fix Failing CI Without Leaving Your Terminal

Connect Codex to CircleCI and give your coding agent direct access to the CI feedback it needs to keep working. In this tutorial, we’ll walk through setting up the CircleCI CLI and CircleCI plugin for Codex, then show how Codex can check pipeline results, validate your CircleCI config, diagnose failed builds, trigger new pipelines, and keep iterating on a fix until CI is green. Instead of bouncing between your terminal and CircleCI to copy logs and errors back to your agent, you can bring the full CI feedback loop directly into your Codex session.

Your existing kit just became more valuable

Hardware costs are rising. But Civo Product Director Russ Smith has a different take: your existing kit just became more valuable. The hyperscalers competing for the same DRAM and compute as you still have to pass that cost on eventually. At high utilisation rates, your resource rental overtakes purchase cost. Typically in under a year.

Continous ORT Testing with Harness

Most Operational Readiness Testing (ORT) programs follow the same ritual. A checklist gets filled out. Someone runs a load test in a war room the week before launch. A failover drill gets scheduled, and everyone hopes it goes cleanly. Then the release is shipped, and testing is done. But with Harness, you can make this process continuous, and your service resilience is protected by the same ORT checklist for every small change in your SDLC.

Monitor Query Costs & Verify Database Changes Automatically

See how Harness Database DevOps and DBmarlin work together to give you full visibility into query performance, automated deployment verification, and AI-assisted database change authoring - all inside your CI/CD pipeline. Most teams deploy database changes blind - they push a schema migration and hope nothing breaks. This demo shows a better way: DBmarlin surfaces the cost and performance of every query before and after a change, while Harness CV uses AI/ML to automatically detect regressions and block bad deployments from reaching production.

What Is RPO and How Can You Reduce It to Minutes?

Learn what RPO means, how it differs from RTO, and how faster backup storage can help reduce data loss after disruption. When designing a backup and disaster recovery strategy, organizations often simply ask how quickly they can recover. But a more important question to ask is how much data they can afford to lose. Backups are only as useful as the point in time they can return you to. If a payment system fails at 2 p.m.

Why Your Help Desk Knowledge Base Isn't Reducing Tickets

The help desk knowledge base has hundreds of articles. Search traffic looks healthy, and employees are reminded to try self-service first. Yet ticket volume barely moves. That gap appears when a knowledge base is measured as content instead of ticket prevention. Article counts, page views, and searches can rise while the queue remains busy. None proves that an employee received a trustworthy answer before filing a ticket.