Nagios-style health monitors over the fleet — checks run on a cron sweep, roll up per-target state, and dispatch digest alerts.