Operations | Monitoring | ITSM | DevOps | Cloud

Top tips: Small digital habits that save you hours every week

Top tips is a weekly column where we highlight what's trending in the tech world and list practical ways to explore these trends. This week, we're looking at something we rarely think about until the end of the day: the tiny digital habits that quietly eat away at our time. Have you finished a workday feeling busy but strangely unaccomplished? You started with the best intentions.

Synthetic Monitoring Is Broken. Your Production Traffic Can Fix It.

Synthetic monitoring has been a critical part of application reliability for years. It gives engineering and operations teams a way to proactively test applications, APIs, and critical customer journeys before users encounter problems. But there is a fundamental limitation with the traditional approach: Someone has to create the tests. As applications become more distributed and customer journeys become more complex, organizations can end up maintaining hundreds or even thousands of synthetic scripts.

Cavalry or cattle? Let the machine decide

Long before dashboards and decibel-loud alerts, there were watchtowers. Every kingdom worth its salt had them, men perched on hills, lighting fires to signal the moment they spotted something suspicious on the horizon. It was, in its time, a fine system. The trouble was that watchmen, being human, occasionally mistook a herd of cattle for an invading army, or a dust storm for smoke, and lit their fires anyway.

New Alert Policy Templates for Kentik Protect: Security and Compliance Visibility in Minutes

Securing distributed networks and maintaining regulatory compliance is often manual, expertise-heavy work of configuring alert policies. Kentik Protect now includes additional pre-configured, production-ready Alert Policy Templates across three categories: outbound security and DDoS, carpet bombing defense, and geopolitical compliance monitoring.

Migrating from Nagios XI to WhatsUp Gold: A Practical Step-by-Step Guide

Monitoring platforms rarely become complex overnight. In many Nagios XI environments, complexity builds gradually through years of useful customizations, custom plugins, one-off fixes, and undocumented operational knowledge. Each addition may have solved a real problem at the time, but over the years the result can become difficult to maintain, explain, and hand over to new administrators.

Stop Guessing Where the Network Broke

Modern IT teams invest heavily in monitoring infrastructure, applications, servers, and network devices. Yet when users report that a critical cloud service is slow or a branch office loses connectivity, one question often remains difficult to answer: where is the problem actually occurring? Is the issue inside your network? Is it your ISP? Has a routing change introduced excessive latency? Did an upstream provider experience an outage?

Garbage in, garbage out: Splunk's Steve Flanders on why AI can't fix your bad telemetry

Cortex co-founder and CTO Ganesh Datta sits down with Steve Flanders, who leads AI transformation at Splunk and wrote the book on OpenTelemetry, to talk about why AI acceleration without strong observability foundations creates more problems than it solves.

What data sources does agentic ITOps use

Agentic IT operations have arrived. It’s no longer a question of if enterprise IT departments will adopt agentic ITOps, but how quickly. The question we hear most often at BigPanda isn’t “what are agentic ITOps,” it’s “what data do we actually need to get started?” That’s the right question to ask. Agentic AI is only as good as the data and context that feeds it. Real-time observability and telemetry data from machines. Structured ITSM and workflow records.