Look For Disruptive Maintenance Windows

The Problem Hosting providers occasionally need to perform disruptive maintenance on your servers. However, most of them won’t send notifications just for your affected servers, but instead for any affected server in a given data center. This isn’t always too helpful since you don’t particularly care if other servers go down, but not your own. […]

Read More →

Recheck a server’s health after an outage

In the example below, we ask Claude to look into a server’s health after an outage. The server has Outgoing PING enabled in its monitoring agent, which gives Claude more information and context about current network health. There are still very light traces of intermittent packet loss, but the system is healthy enough that it […]

Read More →

Investigate the ongoing outage

In this example, we’ll ask Grok to investigate the ongoing outage on one of our development servers. It correctly identified the root cause of the issue and recommended further steps to take in order to debug the issue further.

Read More →

Investigate repeated outages from the same provider

In the example below, we’re asking Claude to investigate yesterday’s outage for one of our servers, since we’ve received an increased number of outage notifications for this particular server recently. We’re immediately presented with a diagnosis of the incident and asked whether we should dig further to see if there’s a pattern with this specific […]

Read More →

Add/edit uptime monitor & investigate outage

In the example below, we’re asking Claude to add a new Uptime Monitor with a given keyword for it to look for, as part of our Keyword Monitoring. Then, once the Uptime Monitor is detected as online/healthy, we’ll ask it to change that monitored keyword to a missing one, thus causing an outage for our […]

Read More →

Manage Server Inventory

The Problem You’re in the middle of upgrading your server fleet, and you’d like to have a quick look at the current status and your overall progress with this task. The manual solution would be to go to your HetrixTools dashboard and manually open the server metrics for all of your uptime monitors, then note […]

Read More →

Investigate Outage

The Problem You receive a downtime notification, and you’d like to know more about what exactly happened. The manual solution would be to open your HetrixTools dashboard, locate the affected uptime monitor, dig through its Location Fail Log and Network Diagnostics, as well as its Server Metrics. This could take quite some time. The MCP […]

Read More →