OpenShift Day 2 Operations - Part 1

  • 29 Jul, 2025

“Day 2” operations refer to everything that happens after the cluster is installed, which could be a lot or a little, depending on how you plan to use the cluster.

Verifying the Health of Your OpenShift 4 Cluster

Managing an OpenShift 4 cluster effectively involves regular health checks to ensure smooth operation and reliability. An unhealthy cluster can lead to downtime, reduced performance, and compromised workloads.

Node Health

Healthy nodes are crucial for running workloads effectively.

  • Use this command to check the node status: oc get nodes

Verify that all nodes show Ready in the STATUS column.

You can use oc get nodes -o wide to get more details about the cluster

For more details on a specific node:

oc describe node <node-name>

Verify resources allocated to a node:

oc describe node <node name>  | grep -A 10 "Allocated resources"

Get Allocated resources for all nodes:

oc describe nodes | grep -A 10 "Allocated resources"

Check Cluster Operators

Cluster Operators are responsible for managing the lifecycle of key components of an OpenShift cluster. To verify their status:

Run the following command:

oc get clusteroperators

A little addon to the previous command very useful when you are upgrading your cluster:

watch -n5 oc get clusteroperators

Pod Health

Ensuring that pods are running as expected is a key part of cluster health.

Get pods not running nor completed

oc get pods -A -o wide | grep -v -E 'Completed|Running'

Related Posts

OpenShift Day 2 Operations - Part 3

  • 05 Nov, 2025

Retrieving OpenShift Cluster Logs Logs are invaluable for understanding the state of your OpenShift cluster and diagnosing problems. OpenShift provides several ways to access these logs efficiently: Get node logs Display node journal: oc adm node-logs Tail 10 lines from node journal: oc adm node-logs --tail=10 Get kubelet journal logs only: oc adm node-logs -u kubelet.service Grep kernel word on node journal: oc adm node-logs --grep=kernel List /var/log contents: oc adm node-logs --path=/ Get /var/log/audit/audit.log from node: oc adm node-logs --path=audit/audit.log Pod Logs Pod logs provide insights into application behavior. - Retrieve logs for a s

OpenShift Day 2 Operations - Part 3Read More

OpenShift Day 2 Operations - Part 2

  • 19 Aug, 2025

“Day 2” operations refer to everything that happens after the cluster is installed, which could be a lot or a little, depending on how you plan to use the cluster. Effective troubleshooting and monitoring in OpenShift require understanding the right way to retrieve logs and manage issues. While it might be tempting to SSH directly into cluster nodes, OpenShift provides tools and workflows to handle logs more securely and efficiently. Why You Should Avoid SSHing to Nodes SSHing directly into cluster nodes might seem like a quick way to debug issues, but it introduces several risks and challenges: 1. Security Risks - Inconsistent Access Control: Granting SSH access bypasses OpenShift’s

OpenShift Day 2 Operations - Part 2Read More

Exploring Kubernetes Startup Scaling

  • 26 Jun, 2025

Kubernetes has long been a cornerstone for managing containerized workloads, and its continuous evolution keeps it at the forefront of cloud-native technologies. One of the exciting advancements in recent releases is the enhancement of startup scaling capabilities, particularly through features like Kube Startup CPU Boost and dynamic resource scaling. In this blog post, we’ll dive into what startup scaling is, how it works, and why it’s a significant addition for Kubernetes users looking to optimize application performance during startup. What is Startup Scaling in Kubernetes? Startup scaling refers to the ability to dynamically allocate additional resources, such as CPU, to pods during th

Exploring Kubernetes Startup ScalingRead More