Operations | Monitoring | ITSM | DevOps | Cloud

The latest News and Information on Monitoring for Websites, Applications, APIs, Infrastructure, and other technologies.

Tips, Tricks, and Shortcuts for Navigating StackState

When it comes to using (desktop) software, especially in tech, there's always that icebreaker you can use to determine whether a prospect is a more "visual" or a more "textual" user. It’s definitely not a black-or-white debate, but it's always interesting to see just how differently we’re wired as individuals at a cognitive level. In this blog post, we'll assume you lean towards being a more "visual" type of person.

Ask the Experts: Distributed Tracing, OpenTelemetry, and Connecting Your Frontend to Your Backend

While baggage isn’t required for distributed tracing, it is required for carrying metadata between services. How will the observability community address that and make it easier over time? Featuring: Winston Hearn, Frontend Observability Expert and Hazel Weakly, Web Developer and SRE.

Ask the Experts: Observability: What Can the Frontend Steal From the Backend?

What is the biggest value of #observability as practiced on the #backend that you are excited to see taken up as more #frontend #developers start practicing observability on their own? Featuring: Winston Hearn, Frontend Observability Expert and Hazel Weakly, Web Developer and #SRE.

How to Scale and Standardize Observability Practices: Hear from Canva and Atlassian | Grafana

This panel discussion, featuring Jenna, Director of Engineering, Reliability Platforms at Canva and Andrew, Head of Engineering at Atlassian, explored the challenges and strategies of implementing standardization in large tech companies. Atlassian, known for its software development and collaboration tools, initially faced resistance to standardization but shifted as inefficiencies and compliance issues emerged. Canva, a graphic design platform, highlighted the balance between flexibility and standardization, using observability tools for accountability.

From "rebooting" to reliable and secure applications: Optimizing the customer experience

Not so long ago in my career, I remember when it was relatively acceptable for infrastructure or development teams to solve a problem by rebooting a server or just “turning things off and on again.” It didn’t matter what caused the problem or how long the reboot would fix things, provided they were fixed for now. Security teams were always held to a different standard.