You can monitor:
- Utilization of hardware resources.
- ClickHouse server metrics.
ClickHouse does not monitor the state of hardware resources by itself.
It is highly recommended to set up monitoring for:
Load and temperature on processors.
Utilization of storage system, RAM and network.
ClickHouse server has embedded instruments for self-state monitoring.
To track server events use server logs. See the logger section of the configuration file.
- Different metrics of how the server uses computational resources.
- Common statistics on query processing.
You can configure ClickHouse to export metrics to Graphite. See the Graphite section in the ClickHouse server configuration file. Before configuring export of metrics, you should set up Graphite by following their official guide.
You can configure ClickHouse to export metrics to Prometheus. See the Prometheus section in the ClickHouse server configuration file. Before configuring export of metrics, you should set up Prometheus by following their official guide.
Additionally, you can monitor server availability through the HTTP API. Send the
HTTP GET request to
/ping. If the server is available, it responds with
To monitor servers in a cluster configuration, you should set the max_replica_delay_for_distributed_queries parameter and use the HTTP resource
/replicas_status. A request to
200 OK if the replica is available and is not delayed behind the other replicas. If a replica is delayed, it returns
503 HTTP_SERVICE_UNAVAILABLE with information about the gap.