Monitoring an all-in-one installation¶
In an all-in-one installation, a single infra_manager cluster (a server or a cluster of 3 nodes) hosts the database, API, portal, signal, and traefik. The API is not exposed: the portal calls it on the internal Swarm network.
Checks from the monitoring server¶
Check |
Probe |
Condition |
|---|---|---|
Portal health check (API, database, provisioning, relay, signal) |
|
|
Entry point HTTPS certificates (portal, signal, administration portal) |
Always |
|
Signal (HTTP 426 “Upgrade Required”) |
Always |
|
Administration portal health check |
|
|
Workstation (port 8445 and certificate) |
|
|
Appliance portal (port 8444 and certificate) |
|
|
Credential portal (port 8446 and certificate) |
|
|
TURN (actual allocation, port 58200 TCP/UDP) |
|
Checks on infra_manager nodes (NRPE)¶
Check |
Probe |
Nodes |
|---|---|---|
System (CPU, memory, disk |
|
All |
One manager |
||
Swarm services (replicas, updates, one-off tasks) |
One manager |
|
All (only the host node responds) |
||
One |
||
All (if |
Other checks¶
Check |
Probe |
Location |
|---|---|---|
Backup server (if |
||
Ansible administration workstation |
||
|
See also
Summary of Flows to Monitor — Network flows to open and monitor.