Skip to main content

Overview

Effective monitoring of the object storage cluster ensures early detection of capacity constraints, performance degradation, and data integrity issues. This guide covers the key metrics and commands for ongoing operational visibility.
Administrator Access Required — This operation requires the admin role. Contact your Polystack administrator if you do not have sufficient permissions.

Cluster Capacity

Storage capacity across all nodes
Capacity thresholds:
When any storage node exceeds 85% capacity, the ring rebalancer may be unable to place new replicas, causing 507 Insufficient Storage errors for writes. Plan capacity expansion before reaching 70% utilization.

Proxy Metrics

The proxy-server exposes metrics on the recon middleware endpoint:
Check proxy load
Check proxy memory
Check proxy async pending updates

Replication Health

Replication status across all nodes
Check for quarantined (corrupted) objects
Verify ring file consistency across nodes
Replication health alerts:

Integration with Monitoring

For continuous monitoring, connect the object storage recon endpoint to Monitoring (Polystack Infrastructure Monitoring Platform):
Prometheus scrape config for object storage
Configure alerting rules in Monitoring for the critical thresholds above. Set notification channels for the on-call team to respond to 507 storage errors and quarantine count spikes promptly.

Next Steps

Replication

Deep-dive into replication health and quarantine management

Ring Management

Expand capacity by adding drives and rebalancing rings

Admin Troubleshooting

Respond to monitoring alerts and diagnose failures

Quotas

Set limits to prevent individual projects from consuming all capacity