Operations | Monitoring | ITSM | DevOps | Cloud

The latest News and Information on DevOps, CI/CD, Automation and related technologies.

CloudZero AI Hub: The nexus of autonomous AI cost control

CloudZero originated as a way to make sense of your cloud costs. Costs spread across bills with billions of line items belonging to resources that might or might not have been tagged (or taggable), spun up by engineers working across teams, on different microservices, features, and products, that served a wide range of customers. Kubernetes. Multi-cloud. Check, check, check.

AI ROI: How to measure and provide the return on AI investments in 2026

Every quarter, the same scene plays out in boardrooms across the Fortune 500. The CEO asks: “What is the return on everything the company is spending on AI?” The CTO talks about productivity gains and developer velocity. The CFO points at a cloud bill that doubled but cannot isolate which line items are AI. The board nods politely and tables the discussion until next quarter, when the same question will produce the same non-answer. (If this sounds familiar, you are not alone. Keep reading.)

NVIDIA Earth-2: OSS and Science for AI Weather and Climate | Ubuntu Summit 26.04

Discover how NVIDIA Earth-2 brings open source software and open science to weather and climate forecasting. Niall Robinson (NVIDIA) introduces a new way of making production-ready weather AI fully accessible for organizations to run, fine-tune, and deploy on their own infrastructure: NVIDIA Earth-2.

Level up your Code on Arm and Ubuntu | Ubuntu Summit 26.04

What are the latest developments in Arm tooling on Ubuntu? In this talk, David explores Arm tooling to analyze and optimize workload performance, and how AI-assisted development using agentic AI and static analysis can accelerate porting and tuning applications for the Arm architecture. About David David Haikney is a Technical Product Director at Arm. He is responsible for Arm Performix, a free performance toolkit that helps developers understand and improve real-world performance on Arm architectures.

Improving Digital Employee Experience with Intelligent Automation | Reduce IT Tickets by 70%

Are your employees still waiting hours or days for IT issues to be resolved? Many organizations have invested heavily in ITSM platforms, self-service portals, and chatbots. Yet service desks continue to struggle with growing ticket volumes, rising costs, slow resolution times, and poor digital employee experiences. In this webinar, Resolve and Redington explore why traditional service management approaches often stop at ticket creation instead of ticket resolution, and how intelligent automation is helping enterprises move toward a Zero Ticket IT model.

How IT Teams Can Start Their AI Automation Journey | Agentic AI, ITSM & Zero Ticket IT

How should IT leaders approach automation and AI? Where should they start, and how can they drive measurable results without getting caught up in the hype? In this episode of Agents of IT, Fran Fernandez and Zach Austin sit down with Chris Ellis, Senior Technology Solutions Specialist at RICOne, to discuss practical IT automation strategies, agentic AI, service desk transformation, and the journey toward autonomous operations.

Premium self-hosted runners are generally available

In December, we shared our plans to introduce pricing for self-hosted runners. You told us loud and clear that a free option matters. Today, as Premium Runners become generally available, we are happy to share that we will continue to have a free tier, which includes the use of up to 100 self-hosted runners as part of your plan. If your team needs more scale, dedicated support, or advanced management features, you can upgrade to Premium Runners when you’re ready.

Best APM for Small Teams Without Dedicated DevOps in 2026

You don’t have an SRE. There’s no platform team. Your “monitoring strategy” is someone checking Slack for error alerts. When production breaks, the same two or three senior devs drop everything to debug. Sound familiar? Most APM tools are built for organizations with dedicated operations staff. They assume someone has time to configure dashboards, tune alert thresholds, and learn a complex query language. That person does not exist on your team.

Logs told me something broke. Traffic showed me what.

Here’s a problem I run into constantly: something breaks in production, I can see the 500 errors in my logs, but I can’t reproduce it locally. The trace shows me the dependency graph but not the actual request that failed. This is especially painful in microservices. I was looking at a CNCF example the other day (a simple demo app, like 4 pods) and it already had so many cross-service dependencies that understanding what broke required looking at the whole system at once.

IBM Think 2026 Infrastructure Insights for IT Leaders

IBM Think 2026 made one thing clear: infrastructure leaders are being asked to support more AI, more automation, and faster decision-making without adding unnecessary complexity or risk. Held earlier this month in Boston, IBM Think 2026 focused heavily on enterprise AI, hybrid cloud, automation, governance, and operational transformation.