Operations | Monitoring | ITSM | DevOps | Cloud

The latest News and Information on DevOps, CI/CD, Automation and related technologies.

Improve the engineering goals you care about: Sleuth's new Goals & Automations feature

Matt Upton, Director of Software Development at Rewind, joined Sleuth's Don Brown and Daniel de Juan to take our new Goals and Automations feature for a test drive. Hear reactions and input from Matt on how he sees the Goals feature being useful for his teams: ‍Key sections: Give Sleuth a try and see why it's a deploy-based Accelerate / DORA metrics tracker both managers and developers love.

The Platform Engineer Role Explained: Who Is a Platform Engineer?

Poorly designed infrastructure leaves your applications and networks vulnerable to cyberattacks and data breaches. This puts the company at significant risk: the average cost of a data breach reached a record high $4.35 million in 2022. This is where companies bring in platform engineers. A platform engineer is a professional who ensures that security protocols and best practices are in place to protect against potential security threats.

What's new in Ubuntu 23.04

Ubuntu has long been a developer favourite and the preferred platform for Linux for gaming. Ubuntu Desktop 23.04 lays the foundations for a number of key strategic priorities around desktop deployment, identity management and gaming. For IT managers looking to deploy Ubuntu Desktop at scale, the new desktop installer delivers new tools for custom image configuration and network based deployment. Ubuntu 23.04 also extends its identity management integration to support industry shifts towards cloud-based identity providers.

Trust shouldn't start at zero

How often have you heard the phrase “trust is earned” in life? While well-meaning, I think this can actually lead to some strange behaviour at work, especially when you’re on a fast growing team. Startups experience a lot of chaos and unknowns your teams need to navigate, so it’s vital to know you can trust the people around you. As you grow, how you set expectations around trust as people join your team can impact your ability to hire, onboard, ship and ultimately, survive.

Boosting Resilience with Chaos Engineering: Litmus 3.0 & Beyond | Civo TV

Prithvi Raj explores the world of chaos engineering and discusses its security, comparisons between open-source projects, and the latest Litmus 3.0 release. Discover how chaos engineering is not just about inducing failures, but also an essential aspect of building resilient systems across all stages of development.

Understanding Azure Function App Metrics

This article will focus on the metrics side of Azure Functions and features offered by the Azure Portal and then talk about the value of Serverless360. Then about the product that provides beyond the primary feature set in the Azure Portal, which will help you improve the day-to-day operations of your Azure solution. There are many different ways you can manage and operate Azure Functions and features like Application Insights which can also help you with Azure Functions.

Assembly time is where you have the most control of an incident

The FDNY EMS Command responds to more than 4,000 calls per day. They range from car accidents to building fires to cats stuck in trees, and responses vary accordingly. Sometimes they might take hours, sometimes they take just a few minutes. With such unpredictable conditions, the FDNY focuses on improving what they call “response time.” That’s the amount of time between a 911 call being made and emergency responders arriving on the scene. This might sound familiar.

Unlock the Secrets of Kernel Memory Usage

The mem.kernel chart in Netdata provides insight into the memory usage of various kernel subsystems and mechanisms. By understanding these dimensions and their technical details, you can monitor your system's kernel memory usage and identify potential issues or inefficiencies. Monitoring these dimensions can help you ensure that your system is running efficiently and provide valuable insights into the performance of your kernel and memory subsystem.

Monitoring Disks: Understanding Workload, Performance, Utilization, Saturation, and Latency

Netdata provides a comprehensive set of charts that can help you understand the workload, performance, utilization, saturation, latency, responsiveness, and maintenance activities of your disks. In this blog we will focus on monitoring disks as block devices, not as filesystems or mount points. The Disks section in the Overview tab contains all the charts that are mentioned in this blog post.