In 2009, a national ISP employed a simple Nagios instance to alarm to core network issues. Oncall pages were sent to all staff at all hours, and a nominated engineer was “on call” every week.
As the ISP (and parent, managed-services company) extended our managed services to customers, engineers initially