The Problem Hosting providers occasionally need to perform disruptive maintenance on your servers. However, most of them won’t send notifications just for your affected servers, but instead for any affected server in a given data center. This isn’t always too helpful since you don’t particularly care if other servers go down, but not your own. […]
Articles Tagged: uptime monitor
Recheck a server’s health after an outage
In the example below, we ask Claude to look into a server’s health after an outage. The server has Outgoing PING enabled in its monitoring agent, which gives Claude more information and context about current network health. There are still very light traces of intermittent packet loss, but the system is healthy enough that it […]
Investigate the ongoing outage
In this example, we’ll ask Grok to investigate the ongoing outage on one of our development servers. It correctly identified the root cause of the issue and recommended further steps to take in order to debug the issue further.
Investigate repeated outages from the same provider
In the example below, we’re asking Claude to investigate yesterday’s outage for one of our servers, since we’ve received an increased number of outage notifications for this particular server recently. We’re immediately presented with a diagnosis of the incident and asked whether we should dig further to see if there’s a pattern with this specific […]
Schedule an upcoming maintenance window
In the example below, we’re asking Claude to schedule an upcoming maintenance window based on an email we received from one of our dedicated server providers.
Check current status & investigate discrepancy
In the example below, we’re asking ChatGPT Codex to quickly give us an overview of our Uptime Monitors, then follow up on a detected issue caused by a monitor configuration discrepancy.
Add/edit uptime monitor & investigate outage
In the example below, we’re asking Claude to add a new Uptime Monitor with a given keyword for it to look for, as part of our Keyword Monitoring. Then, once the Uptime Monitor is detected as online/healthy, we’ll ask it to change that monitored keyword to a missing one, thus causing an outage for our […]
Investigate today’s outage
In the example below, we’re asking Claude to investigate a recent outage with one of our Uptime Monitors.
Manage Server Inventory
The Problem You’re in the middle of upgrading your server fleet, and you’d like to have a quick look at the current status and your overall progress with this task. The manual solution would be to go to your HetrixTools dashboard and manually open the server metrics for all of your uptime monitors, then note […]
Investigate Outage
The Problem You receive a downtime notification, and you’d like to know more about what exactly happened. The manual solution would be to open your HetrixTools dashboard, locate the affected uptime monitor, dig through its Location Fail Log and Network Diagnostics, as well as its Server Metrics. This could take quite some time. The MCP […]