For the complete documentation index, see llms.txt. This page is also available as Markdown.

Monitor, scale, and maintain the cluster

Manage hardware, scale clusters, and optimize resources to ensure system stability and performance.

Manage routine operations for WEKA clusters running on Kubernetes. Use this page to choose the task area that matches your goal and open the detailed procedure.

The linked topics cover these day-2 workflows:

  • Observability and monitoring: Collect and visualize WEKA health and performance metrics with Kubernetes monitoring tools.

  • Hardware maintenance: Reboot, replace, or remove servers, containers, drives, and nodes while preserving cluster health.

  • Cluster scaling: Expand or shrink backend and S3 resources, and adjust client or backend core allocation.

  • Cluster maintenance: Update WekaCluster and WekaClient settings, rotate pods, manage client token secrets, pause or resume a cluster, and cancel cluster deletion.


Related topics

Observability and monitoring

Hardware maintenance

Cluster scaling

Cluster maintenance

Last updated