Operations | Monitoring | ITSM | DevOps | Cloud

The latest News and Information on APIs, Mobile, AI, Machine Learning, IoT, Open Source and more!

AI in IT Operations: How to Build Trust, Automate Smarter & Prepare for Autonomous IT

What does it take to make AI and automation actually work in enterprise IT? In this episode of Agents of IT, Resolve’s Zach Austin sits down with Nick Dimmock, Co-Founder and CEO of TechWorks, to discuss how IT automation is evolving, why trusted data matters, and what organizations need to build before AI can deliver meaningful business outcomes.

AI cost allocation: how to attribute AI spend by team, product, and customer

AI cost allocation is the practice of attributing every dollar of AI spend to the team, product, feature, or customer that generated it. That spend includes API tokens, GPU compute, per-seat tools, and shared infrastructure. It's harder than cloud allocation because AI spend arrives untagged, spans vendors, and pools in shared resources. Four methods cover most cases: tag-based, key-based attribution, proportional split, and usage-telemetry.

This AI agent finds your app's bottlenecks and suggests the fix

Most teams collect the profiles and traffic data that explain a slowdown. Almost nobody has time to read it before users notice. In this Product Highlights conversation, Sylvain Guittard, Senior Director of Product at Upsun who leads the team behind the Upsun console and CLI, breaks down the Upsun Cloud Performance Agent, the first background agent running on Upsun Cloud. His take: "We monitor everything, we feed that into an agent, and the agent will be capable of finding what the bottlenecks are in your application. And on top of it, it gives you a patch, or a way to fix it." We get into.

Watch this AI agent find and fix performance bottlenecks

Anyone can claim an AI agent will fix your performance problems. This demo shows exactly what it looks at and what it hands back. In this Product Highlights conversation, Sylvain Guittard, Senior Director of Product at Upsun who leads the team behind the Upsun console and CLI, runs the Upsun Cloud Performance Agent live on a demo project. His take: "You have a patch that is already available, and a recommendation, so you can see if it fits or not to your context." We get into.

Realtime transaction fraud detection - with an LLM?

Conversational AI with a chatbot is great for drafting emails or debugging code, but it’s less ideal for real-time application middleware. If you’re trying to inspect a financial transaction for potential fraud in the middle of a checkout loop, you don’t need an LLM to write you an essay about why a credit card transaction looks suspicious – you just need a probability score, and you need it as fast as possible.

How Etsy gets its mobile apps ready for peak traffic

Most advice about surviving a traffic spike is about capacity. Scale the fleet, warm the caches, load test the checkout path. All of it assumes you can fix whatever breaks the moment you find it. Mobile apps don’t work that way. I spent an hour on a workshop with Jay Henry, a senior engineering manager at Etsy who owns engineering strategy across three teams covering CI, build, test, release, observe, and SRE. Jay’s take: a web team having a bad day can revert in minutes.

OpenAI outage on September 14, 2026: "some tools are temporarily unavailable" errors hit ChatGPT worldwide

ChatGPT’s tools and workspace features failed for users around the world on September 14, 2026, with the message “Some tools are temporarily unavailable” blocking spreadsheet and document creation, Codex, Projects, file reading, and dictation while core chat largely kept working. StatusGator sent an Early Warning Signal at 14:41 UTC, 1 hour and 17 minutes before OpenAI publicly acknowledged the incident at 15:58 UTC.