Job Summary
Senior Linux & OpenShift Infrastructure, Capacity and Performance Engineer Manage and support large-scale enterprise Linux and OpenShift environments across Production, DR, UAT, and Development environments. • Administer RHEL 5/67/8/9/10 servers and other (Rocky, ubuntu, CentOS Etc) including installation, configuration, troubleshooting, performance tuning, and lifecycle management. • Manage Red Hat OpenShift clusters including cluster administration, upgrades, patching, capacity planning, security, and operational support. • Perform Linux server provisioning, decommissioning, migration, and modernization activities. • Support VMware-to-OpenShift migration initiatives and containerization programs. • Administer Red Hat Satellite infrastructure including content views, lifecycle environments, repositories, activation keys, and patch management. • Plan and execute operating system upgrades, server refreshes, application migrations, and infrastructure modernization projects. • Monitor and manage platform availability, reliability, and operational health across Linux and OpenShift platforms. • Perform advanced troubleshooting of Linux, Kubernetes, OpenShift, storage, networking, and application-related issues. • Manage OpenShift worker nodes, control plane nodes, infrastructure nodes, namespaces, projects, operators, certificates, and cluster resources. • Implement and manage OpenShift RBAC, security policies, authentication, and authorization controls. • Perform capacity planning and forecasting for CPU, memory, storage, network, and cluster resources. • Analyze performance trends and provide recommendations for scaling, optimization, and resource utilization improvements. • Develop and maintain operational dashboards, health reports, and monthly capacity reports. • Manage Linux patching, vulnerability remediation, security compliance, and hardening activities. • Coordinate with application owners, database teams, network teams, storage teams, security teams, and vendors for incident and change activities. • Support Oracle, WebLogic, middleware, and container-based application workloads hosted on Linux and OpenShift platforms. • Administer and troubleshoot storage technologies including SAN, NFS, LVM, XFS, Persistent Volumes, Storage Classes, ODF, and Ceph. • Manage OpenShift backup and recovery solutions including Velero and OADP. • Participate in Disaster Recovery planning, testing, failover validation, and recovery activities. • Develop, maintain, and automate operational processes using Ansible, Shell Scripting, Python, and Terraform. • Create and maintain SOPs, Runbooks, KEDB articles, operational documentation, and technical standards. • Perform root cause analysis (RCA) for major incidents and provide corrective and preventive actions. • Support 24x7 production environments and participate in on-call and major incident bridges when required. • Manage vulnerability remediation activities identified through Qualys, Nessus, Tenable, or similar security tools. • Support GitOps, CI/CD, ArgoCD, Jenkins, Helm, Operators, and container platform automation initiatives. • Monitor platform health using Prometheus, Grafana, Alert Manager, Elasticsearch, Kibana, Azure Monitor, Zabbix, SCOM, or equivalent monitoring solutions. • Administer user access, LDAP/AD integration, SSSD, PAM, SSH, authentication, and privileged access management. • Ensure platform compliance with security, audit, governance, and regulatory requirements. • Review infrastructure growth trends and provide strategic recommendations for platform expansion and modernization. Required Skills • Red Hat Enterprise Linux (RHEL 7/8/9/10) • Red Hat OpenShift 4.x • Kubernetes Administration • Red Hat Satellite • Ansible Automation Platform • Shell Scripting • Python • Terraform • GitOps • ArgoCD • Jenkins • Helm • Podman • CRI-O • LDAP/Active Directory Integration • Storage Administration (SAN, NFS, LVM, XFS) • VMware • Oracle Linux/Oracle Workloads • WebLogic Administration
Key Responsibilities
Senior Linux & OpenShift Infrastructure, Capacity and Performance Engineer Manage and support large-scale enterprise Linux and OpenShift environments across Production, DR, UAT, and Development environments. • Administer RHEL 5/67/8/9/10 servers and other (Rocky, ubuntu, CentOS Etc) including installation, configuration, troubleshooting, performance tuning, and lifecycle management. • Manage Red Hat OpenShift clusters including cluster administration, upgrades, patching, capacity planning, security, and operational support. • Perform Linux server provisioning, decommissioning, migration, and modernization activities. • Support VMware-to-OpenShift migration initiatives and containerization programs. • Administer Red Hat Satellite infrastructure including content views, lifecycle environments, repositories, activation keys, and patch management. • Plan and execute operating system upgrades, server refreshes, application migrations, and infrastructure modernization projects. • Monitor and manage platform availability, reliability, and operational health across Linux and OpenShift platforms. • Perform advanced troubleshooting of Linux, Kubernetes, OpenShift, storage, networking, and application-related issues. • Manage OpenShift worker nodes, control plane nodes, infrastructure nodes, namespaces, projects, operators, certificates, and cluster resources. • Implement and manage OpenShift RBAC, security policies, authentication, and authorization controls. • Perform capacity planning and forecasting for CPU, memory, storage, network, and cluster resources. • Analyze performance trends and provide recommendations for scaling, optimization, and resource utilization improvements. • Develop and maintain operational dashboards, health reports, and monthly capacity reports. • Manage Linux patching, vulnerability remediation, security compliance, and hardening activities. • Coordinate with application owners, database teams, network teams, storage teams, security teams, and vendors for incident and change activities. • Support Oracle, WebLogic, middleware, and container-based application workloads hosted on Linux and OpenShift platforms. • Administer and troubleshoot storage technologies including SAN, NFS, LVM, XFS, Persistent Volumes, Storage Classes, ODF, and Ceph. • Manage OpenShift backup and recovery solutions including Velero and OADP. • Participate in Disaster Recovery planning, testing, failover validation, and recovery activities. • Develop, maintain, and automate operational processes using Ansible, Shell Scripting, Python, and Terraform. • Create and maintain SOPs, Runbooks, KEDB articles, operational documentation, and technical standards. • Perform root cause analysis (RCA) for major incidents and provide corrective and preventive actions. • Support 24x7 production environments and participate in on-call and major incident bridges when required. • Manage vulnerability remediation activities identified through Qualys, Nessus, Tenable, or similar security tools. • Support GitOps, CI/CD, ArgoCD, Jenkins, Helm, Operators, and container platform automation initiatives. • Monitor platform health using Prometheus, Grafana, Alert Manager, Elasticsearch, Kibana, Azure Monitor, Zabbix, SCOM, or equivalent monitoring solutions. • Administer user access, LDAP/AD integration, SSSD, PAM, SSH, authentication, and privileged access management. • Ensure platform compliance with security, audit, governance, and regulatory requirements. • Review infrastructure growth trends and provide strategic recommendations for platform expansion and modernization. Required Skills • Red Hat Enterprise Linux (RHEL 7/8/9/10) • Red Hat OpenShift 4.x • Kubernetes Administration • Red Hat Satellite • Ansible Automation Platform • Shell Scripting • Python • Terraform • GitOps • ArgoCD • Jenkins • Helm • Podman • CRI-O • LDAP/Active Directory Integration • Storage Administration (SAN, NFS, LVM, XFS) • VMware • Oracle Linux/Oracle Workloads • WebLogic Administration
Skill Requirements
Senior Linux & OpenShift Infrastructure, Capacity and Performance Engineer Manage and support large-scale enterprise Linux and OpenShift environments across Production, DR, UAT, and Development environments. • Administer RHEL 5/67/8/9/10 servers and other (Rocky, ubuntu, CentOS Etc) including installation, configuration, troubleshooting, performance tuning, and lifecycle management. • Manage Red Hat OpenShift clusters including cluster administration, upgrades, patching, capacity planning, security, and operational support. • Perform Linux server provisioning, decommissioning, migration, and modernization activities. • Support VMware-to-OpenShift migration initiatives and containerization programs. • Administer Red Hat Satellite infrastructure including content views, lifecycle environments, repositories, activation keys, and patch management. • Plan and execute operating system upgrades, server refreshes, application migrations, and infrastructure modernization projects. • Monitor and manage platform availability, reliability, and operational health across Linux and OpenShift platforms. • Perform advanced troubleshooting of Linux, Kubernetes, OpenShift, storage, networking, and application-related issues. • Manage OpenShift worker nodes, control plane nodes, infrastructure nodes, namespaces, projects, operators, certificates, and cluster resources. • Implement and manage OpenShift RBAC, security policies, authentication, and authorization controls. • Perform capacity planning and forecasting for CPU, memory, storage, network, and cluster resources. • Analyze performance trends and provide recommendations for scaling, optimization, and resource utilization improvements. • Develop and maintain operational dashboards, health reports, and monthly capacity reports. • Manage Linux patching, vulnerability remediation, security compliance, and hardening activities. • Coordinate with application owners, database teams, network teams, storage teams, security teams, and vendors for incident and change activities. • Support Oracle, WebLogic, middleware, and container-based application workloads hosted on Linux and OpenShift platforms. • Administer and troubleshoot storage technologies including SAN, NFS, LVM, XFS, Persistent Volumes, Storage Classes, ODF, and Ceph. • Manage OpenShift backup and recovery solutions including Velero and OADP. • Participate in Disaster Recovery planning, testing, failover validation, and recovery activities. • Develop, maintain, and automate operational processes using Ansible, Shell Scripting, Python, and Terraform. • Create and maintain SOPs, Runbooks, KEDB articles, operational documentation, and technical standards. • Perform root cause analysis (RCA) for major incidents and provide corrective and preventive actions. • Support 24x7 production environments and participate in on-call and major incident bridges when required. • Manage vulnerability remediation activities identified through Qualys, Nessus, Tenable, or similar security tools. • Support GitOps, CI/CD, ArgoCD, Jenkins, Helm, Operators, and container platform automation initiatives. • Monitor platform health using Prometheus, Grafana, Alert Manager, Elasticsearch, Kibana, Azure Monitor, Zabbix, SCOM, or equivalent monitoring solutions. • Administer user access, LDAP/AD integration, SSSD, PAM, SSH, authentication, and privileged access management. • Ensure platform compliance with security, audit, governance, and regulatory requirements. • Review infrastructure growth trends and provide strategic recommendations for platform expansion and modernization. Required Skills • Red Hat Enterprise Linux (RHEL 7/8/9/10) • Red Hat OpenShift 4.x • Kubernetes Administration • Red Hat Satellite • Ansible Automation Platform • Shell Scripting • Python • Terraform • GitOps • ArgoCD • Jenkins • Helm • Podman • CRI-O • LDAP/Active Directory Integration • Storage Administration (SAN, NFS, LVM, XFS) • VMware • Oracle Linux/Oracle Workloads • WebLogic Administration
Other Requirements
Senior Linux & OpenShift Infrastructure, Capacity and Performance Engineer Manage and support large-scale enterprise Linux and OpenShift environments across Production, DR, UAT, and Development environments. • Administer RHEL 5/67/8/9/10 servers and other (Rocky, ubuntu, CentOS Etc) including installation, configuration, troubleshooting, performance tuning, and lifecycle management. • Manage Red Hat OpenShift clusters including cluster administration, upgrades, patching, capacity planning, security, and operational support. • Perform Linux server provisioning, decommissioning, migration, and modernization activities. • Support VMware-to-OpenShift migration initiatives and containerization programs. • Administer Red Hat Satellite infrastructure including content views, lifecycle environments, repositories, activation keys, and patch management. • Plan and execute operating system upgrades, server refreshes, application migrations, and infrastructure modernization projects. • Monitor and manage platform availability, reliability, and operational health across Linux and OpenShift platforms. • Perform advanced troubleshooting of Linux, Kubernetes, OpenShift, storage, networking, and application-related issues. • Manage OpenShift worker nodes, control plane nodes, infrastructure nodes, namespaces, projects, operators, certificates, and cluster resources. • Implement and manage OpenShift RBAC, security policies, authentication, and authorization controls. • Perform capacity planning and forecasting for CPU, memory, storage, network, and cluster resources. • Analyze performance trends and provide recommendations for scaling, optimization, and resource utilization improvements. • Develop and maintain operational dashboards, health reports, and monthly capacity reports. • Manage Linux patching, vulnerability remediation, security compliance, and hardening activities. • Coordinate with application owners, database teams, network teams, storage teams, security teams, and vendors for incident and change activities. • Support Oracle, WebLogic, middleware, and container-based application workloads hosted on Linux and OpenShift platforms. • Administer and troubleshoot storage technologies including SAN, NFS, LVM, XFS, Persistent Volumes, Storage Classes, ODF, and Ceph. • Manage OpenShift backup and recovery solutions including Velero and OADP. • Participate in Disaster Recovery planning, testing, failover validation, and recovery activities. • Develop, maintain, and automate operational processes using Ansible, Shell Scripting, Python, and Terraform. • Create and maintain SOPs, Runbooks, KEDB articles, operational documentation, and technical standards. • Perform root cause analysis (RCA) for major incidents and provide corrective and preventive actions. • Support 24x7 production environments and participate in on-call and major incident bridges when required. • Manage vulnerability remediation activities identified through Qualys, Nessus, Tenable, or similar security tools. • Support GitOps, CI/CD, ArgoCD, Jenkins, Helm, Operators, and container platform automation initiatives. • Monitor platform health using Prometheus, Grafana, Alert Manager, Elasticsearch, Kibana, Azure Monitor, Zabbix, SCOM, or equivalent monitoring solutions. • Administer user access, LDAP/AD integration, SSSD, PAM, SSH, authentication, and privileged access management. • Ensure platform compliance with security, audit, governance, and regulatory requirements. • Review infrastructure growth trends and provide strategic recommendations for platform expansion and modernization. Required Skills • Red Hat Enterprise Linux (RHEL 7/8/9/10) • Red Hat OpenShift 4.x • Kubernetes Administration • Red Hat Satellite • Ansible Automation Platform • Shell Scripting • Python • Terraform • GitOps • ArgoCD • Jenkins • Helm • Podman • CRI-O • LDAP/Active Directory Integration • Storage Administration (SAN, NFS, LVM, XFS) • VMware • Oracle Linux/Oracle Workloads • WebLogic Administration