Operations | Monitoring | ITSM | DevOps | Cloud

KPI cards: build a reliability dashboard that doesn't force tradeoffs

This week's Feature Friday: Principal Product Manager Christine Byun walks through KPI cards, a new way to build custom dashboards in Engineering Intelligence. KPI cards pull key metrics, like change failure rate and rollback frequency, into compact tiles so they stay visible without taking up chart space. That means the metric you're actively working, incidents, in this demo, gets full-size room, without losing sight of the rest of your system.

Data Center: More or Less | SolarWinds TechPod

In this episode, Sean and Crystal explore the complex and rapidly evolving landscape of AI, data center impacts, regulation challenges, and societal implications. They discuss the urgency of establishing standards and the lessons from historical industrial revolutions to navigate AI's future responsibly.

Monitoring Oracle ASM with Custom Metrics | The Tony and Tonie Show Ep 49

Even small Oracle ASM issues can become big database problems. Here's how to spot the warning signs early. Tony and Tonie discuss how Redgate Monitor custom metrics help teams close a common monitoring gap: surfacing Oracle ASM health and performance issues before storage pressure, rebalancing problems, or disk group failures become database incidents.

Creating Escalation and Regular Groups in OnPage

Learn how to create and configure an Escalation Group in OnPage with this step-by-step how-to guide. This video walks through how to create an escalation group, a regular group, configure escalation intervals and factors, enable Round Robin, set failover OPIDs, and add a Fail Report email address. With escalation groups, OnPage can route critical alerts to team members in a predefined order and automatically move to the next responder when needed—helping ensure time-sensitive notifications don’t go unanswered.

How to Create and Import Contacts in the NEW OnPage Web Console

A step-by-step guide for OnPage’s new web management console, including how to create a single contact and how to create multiple contacts at once by importing an Excel spreadsheet. Feel free to comment below with any questions! Whether you’re in IT or healthcare, OnPage helps teams manage critical alerting and communication to ensure urgent messages reach the right people at the right time. If you’re not yet using OnPage or want to see how it works, request a demo or speak with a member of our team to learn more.

Creating an AfterHours OnCall Schedule

Learn how to create and configure an **on-call schedule in OnPage** with this step-by-step how-to guide. In this video, we walk through how to select a group, create a new schedule, assign on-call team members, set escalation priority, and configure coverage for specific days and times. Using OnPage’s on-call scheduling capabilities, teams can ensure the right responders are automatically available for critical alerts, incidents, messages or calls during designated coverage periods, including after-hours, weekends, and other shifts.

What is going wrong with AI coding? Live Laugh Logs ep. 4

Welcome to Episode 4 of Live Laugh Logs, the podcast from the Coralogix Developer Relations team. This week, Chris Cooney joins Annie to share five key DevOps skills that have become even more important in the age of agentic code development, and gives you five key actions you can do today to start levelling up these skills. Subscribe to our channel for more insights into observability and AI.

Grafana Pyroscope: Call Tree, Heat Map, & Adaptive Profiles (August 2026 Community Call)

We will look at some new features: Call Tree, Heat Map, & Adaptive Profiles Can't comment in the chat? You may need to create a channel. Join us live for an introduction to flame graphs. We’ll cover what they are, how to read them, and how to use them to find performance bottlenecks in your applications. Bring your questions! Grafana Cloud is the easiest way to get started with Grafana dashboards, metrics, logs, traces, and profiles. Our forever-free tier includes access to 10k metrics, 50GB logs, 50GB traces and more.