Operations | Monitoring | ITSM | DevOps | Cloud

The latest News and Information on APIs, Mobile, AI, Machine Learning, IoT, Open Source and more!

From Telemetry to Shared Understanding: Why Operations Teams Need Better Visual Incident Notes

Modern operations teams are rarely short on data. A production incident can generate thousands of log lines, multiple dashboards, traces across several services, deployment events, alerts, chat messages, and customer reports. The harder problem is turning that data into shared understanding quickly enough for people to act.

Analysing Claude Code telemetry with SquaredUp - diving deeper

In our previous article we looked at the basics of: In this article, we are going to take a deeper dive into some of the complexities of configuration as well as some of the nuances of analysing Claude telemetry. Before we dive into the code, let us just remind ourselves that our telemetry pipeline looks like this: That is, we are emitting Claude Code telemetry to an OpenTelemetry Collector. The telemetry is then exported to an Application Insights endpoint and stored in Log Analytics tables.

Mirror, Cut Over, Move On: A Live Kafka Migration from Confluent to Aiven

Kafka migrations get talked about like they're impossible. They aren't. They're a sequence of decisions, preparations, and good compromises - plus a working playbook. In this session, Dirk runs a real Confluent Cloud-to-Aiven migration live. Top to bottom: target Kafka cluster deployment, MirrorMaker 2 setup, replication flow and offset handling, and a staged cutover that keeps producer and consumer downtime to almost zero.

Introducing Kepler | GitKraken's Agentic Development Environment (ADE)

Kepler is GitKraken's agentic development environment: mission control for running parallel coding agents at scale. Running one agent is easy. Running five of them across three repos is where things break: scattered terminals, no shared view, no idea what's done or stuck. Kepler puts every agent session on one surface so you can plan work, write code, and review what ships without losing track of anything.

Inside the AI Team Weekly: AI Observability workflows and Prometheus exemplars (May 19th, 2026)

The Grafana AI team (Engineers Ivana Huckova and Sonia Aguilar) share what's new in AI Observability this week: a new way to instrument and visualize agent workflows, plus a neat trick for jumping straight from a metric spike to the exact conversation that caused it using Prometheus exemplars. In this episode: We're showing parts of our team meetings to build in public in some small way and give you a sneak preview of what's to come. But not all features we show may make it to production! You've been warned. :)