Operations | Monitoring | ITSM | DevOps | Cloud

UptimeRobot has acquired IsDown.

For 15 years, UptimeRobot has answered one question for millions of people: is my site up? Today we’re taking on a related question that’s become just as important: is everything my site depends on up? We’ve acquired IsDown, which tracks the status of more than 6,000 cloud and SaaS providers in one place. Here are the details on what it is, why we bought it, and what we’re building with it.

Your AI agent can now set up uptime monitoring. No signup required.

We’re happy to introduce you to the first agentic uptime monitoring setup! Now your AI agents can continue their coding, testing, and deploy flows into setting up the monitors, including a free account. The agent submits your site’s URL and your email, you confirm with one click in the email you receive, and you have a working monitor plus a free account. No registration form, API key, or dashboard configuration needed.

Aiven for ClickHouse 26.3 LTS: Full-Text Search, Async Inserts, and Materialized CTEs

Aiven for ClickHouse 26.3 is now available in Early Availability. This upstream Long-Term Support release makes full-text search generally available, enables asynchronous inserts by default, introduces materialized common table expressions, and brings a wide range of JSON and query-performance improvements. The upstream 26.3 release includes 27 new features and 40 performance optimizations.

Who Owns Deployment Governance? Structuring Accountability in the AI Era

In this series, we have talked about how generative AI is shifting the landscape of software creation. In The New Software Creator, we explored how AI expands who can write code. In When Anyone Can Build Software, Deployment Governance Is What Keeps It Safe, we looked at why the deployment pipeline is the ultimate control point. Finally, in Security at Scale: What Changes When Everyone Can Deploy, we dug into the technical realities of patching, container hygiene, and identity management.

AI-Powered Spacecraft Operations with InfluxDB 3

Summary The InfluxDB satellite telemetry demo is a live mission-control application that monitors a simulated fleet of 12 satellites in real-time. It shows how the InfluxDB 3 Processing Engine can detect anomalies as data is written, enrich time series data with third-party data, and power a grounded AI agent using the InfluxDB 3 MCP server to investigate and explain fleet health.

Shipped: Catch a broken regex before it breaks your rules

Let’s say you want a Matches condition that picks up “prod”, “PROD”, and “Prod”, so you write ‘Matches: (?i)prod’ and check it in an online regex tester before publishing. It looks fine. Python accepts it, and so does JavaScript. Dimension Studio used to accept it too, and the rule is published. Then your data stops updating. The new pattern is what’s keeping it from materializing, but nothing tells you that.

Our 3-month AI roadmap - the future of smart dashboards

AI is set to transform our technology landscape. For many of us working in software, it already has — developers are now writing more code, building more features, and deploying more applications, faster. For the teams supporting IT and software services that means more applications to support, across a greater breadth of technologies, and with more complexity (that is probably less well understood by the developers who created it). Your operational tooling needs to keep pace.

A practical guide to React error monitoring

When designing effective error handling for React apps, the troubleshooting information you collect and display is critical. React errors can stem from a variety of causes, including user misconfiguration, backend and network issues, and mismatches in browser environments. Instrumenting your code to log critical context, including feature names, user data, and session activity, enables you to quickly identify where these errors originate.

Reproducing split brain on CloudNativePG

We run Postgres under an operator for automatic failover. That is a promise about what happens during a failure, so the only way to know you have it is to cause the failure and watch. The docs tell you what should happen. A config review tells you which knobs are set. Neither tells you how long an isolated primary keeps accepting writes after its replacement has been promoted, and that number decides whether a failover is clean or leaves you with two versions of your data.