Job Summary
We are seeking an experienced Unix L2/L3 Support Engineer with strong expertise in Linux (RedHat), Solaris administration to provide 24x7 operational support for enterprise datacenter and cloud environments. The role involves incident management, problem determination, system administration, automation, infrastructure support, performance optimization, patch management, and ensuring high availability of critical business services.
The candidate should possess strong troubleshooting skills across Unix platforms, cloud technologies, automation tools, clustering solutions, and infrastructure management frameworks.
Key Responsibilities
Unix/Linux/Solaris
- Provide L2 administration and operational support for Linux, Solaris, servers across Data Center and Cloud environments.
- Perform installation, configuration, administration, patching, upgrades, migration, and lifecycle management of Unix/Linux platforms.
- Manage virtualization technologies including:
- Solaris ZFS and LDOMs
- Perform OS hardening, capacity management, performance tuning, and system optimization.
- Troubleshoot server, operating system, hardware, storage, and application-related issues.
- Handle break-fix activities and service restoration within agreed SLAs.
- Create and maintain environments for batch processing and associated reporting systems.
Infrastructure & Platform Management
- Manage Red Hat Satellite for patching, provisioning, and lifecycle management.
- Administer Centrify, GPFS, Puppet, and Ansible environments.
- Support Docker and Kubernetes platforms.
- Manage AWS and Azure virtual machine deployments and administration.
- Support storage integration, clustering, partitioning, and virtualization technologies.
- Provide support for HA and Cluster technologies.
Automation & DevOps
- Hands-on experience in Ansible and/or Terraform for infrastructure provisioning, configuration management, and automation.
- Strong scripting skills using Shell Scripting, Python to automate operational and administrative tasks.
- Develop, maintain, and enhance automation workflows, reusable playbooks, and Infrastructure-as-Code (IaC) solutions.
- Drive continuous improvement initiatives to reduce manual effort and improve operational efficiency.
- Automate patching, monitoring, reporting, server provisioning, and compliance-related activities.
Incident, Problem & Change Management
- Monitor infrastructure health and proactively address issues.
- Participate in incident, problem, and change management processes.
- Perform root cause analysis (RCA) and provide corrective and preventive actions.
- Collaborate with application, network, cloud, storage, and security teams during major incidents.
- Support system integration and production rollout activities.
24x7 Operational Support
- Provide 24x7 operational support, monitoring, incident management, and troubleshooting to ensure service availability and business continuity.
- Participate in shift-based support, on-call rotations, weekend support, and major incident management.
- Ensure compliance with SLA, OLA, and operational standards.
- Support planned maintenance activities and emergency changes whenever required.
Skill Requirements
Operating Systems
- Red Hat Enterprise Linux (RHEL)
- Sun Solaris
Unix Technologies
- Solaris LDOMs
- GPFS
- Centrify
Configuration Management & Automation
- Ansible
- Terraform
- Puppet
- Shell Scripting
- Python
- Infrastructure as Code (IaC)
Cloud Platforms
- AWS:
- EC2
- S3
- ELB
- Auto Scaling
- Microsoft Azure
- GCP
- Openshift
- VM Deployment and Administration
Containerization & Orchestration
- Docker
- Kubernetes
Monitoring & Operations
- Performance Monitoring
- Capacity Management
- Patch Management
- Incident Management
- Problem Management
- Change Management
- Root Cause Analysis
High Availability & Clustering
- High Availability (HA) Solutions
- OS Clustering Technologies
Required Qualifications
- Bachelor's Degree in Engineering, Computer Science, Information Technology, or related field.
- Minimum 6+ years of experience in Unix/Linux system administration and support.
- Strong troubleshooting experience in Linux, Solaris, and AIX environments.
- Experience supporting enterprise-scale infrastructure in data center and cloud environments.
- Excellent verbal and written communication skills.
- Strong analytical, problem-solving, and customer management skills.
Other Requirements
Preferred Certifications
One or more of the following certifications preferred:
- RHCSA – Red Hat Certified System Administrator
- RHCE – Red Hat Certified Engineer
- Solaris Certified Administrator
- Docker Certified Associate
- AWS SysOps Administrator
- AWS Solutions Architect
- Microsoft Azure Administrator
- Microsoft Azure Solutions Architect
Desired Competencies
- Strong customer focus and stakeholder management skills.
- Ability to work independently in a 24x7 support environment.
- Experience in large enterprise infrastructure operations.
- Knowledge of ITIL processes and service management practices.
- Continuous improvement mindset with focus on automation and operational excellence.