Administrator (Tools & Automation)
India
Job Description
Administrator (Tools & Automation)
Chennai, Tamil Nadu

Job Summary

We are seeking an experienced Mainframe z/OS System Programmer / Administrator responsible for maintaining and supporting IBM z/OS environments, ensuring maximum system availability, performance, and operational stability. The role involves troubleshooting complex technical issues, managing system operations, supporting 24x7 production environments, performing root cause analysis, and ensuring adherence to customer SLAs and operational processes. The ideal candidate will possess strong expertise in z/OS infrastructure, JES, Sysplex technologies, system performance analysis, security administration, and mainframe operations support.

Key Responsibilities

Production Support & Incident Management Provide technical support for z/OS production environments and ensure maximum uptime in accordance with SLAs. Troubleshoot and resolve critical system, application, and infrastructure issues. Perform problem determination, root cause analysis (RCA), and implement preventive actions. Manage and resolve incidents using enterprise incident and change management processes and tools. Support 24x7 operational environments and participate in on-call support activities. Mainframe Systems Administration Administer and maintain IBM z/OS operating systems and related system software. Manage JES2/JES3 environments and job processing activities. Support Parallel Sysplex, GDPS, catalog management, and system recovery functions. Perform console monitoring, system IPLs, and system health monitoring activities. Manage USS (Unix System Services) environments and related administration tasks. Provide support for system upgrades, maintenance activities, and implementation projects. Performance Monitoring & Optimization Monitor system performance and identify optimization opportunities. Perform performance analysis using: RMF (Resource Measurement Facility) SMF (System Management Facility) Subsystem performance analysis tools Recommend and implement performance improvements to ensure system stability and efficiency. Analyze capacity trends and support infrastructure planning activities. Security & Compliance Support RACF security administration and troubleshooting. Investigate security violations, access issues, and related error codes. Ensure compliance with organizational security and operational standards. Assist with audit activities and system reviews. Backup, Recovery & Business Continuity Develop and maintain backup and restoration strategies supporting 24x7 operations. Participate in disaster recovery planning, testing, and execution. Support system recovery activities during planned and unplanned outages. Validate backup and recovery procedures regularly. Process Governance & Service Delivery Define, improve, and maintain operational processes and procedures. Ensure process adherence across support teams. Monitor service delivery metrics and SLA compliance. Support change implementation activities while minimizing operational risk. Prepare technical documentation, operational runbooks, and knowledge articles.

Skill Requirements

Required Technical Skills Mandatory Skills Strong understanding of IBM z/OS architecture and services. Expertise in: JES2/JES3 Parallel Sysplex GDPS Catalog Management Strong knowledge of: Console Monitoring IPL Procedures System Startup and Shutdown Activities Knowledge of: z/OS Hardware Architecture LPAR Configuration HMC Administration Working knowledge of: JCL REXX CLIST Experience supporting and troubleshooting: CICS IMS IBM MQ Knowledge of RACF administration and security troubleshooting. Experience with workload scheduling tools such as Control-M. Knowledge of USS (Unix System Services) administration. Proficiency in performance analysis using RMF, SMF, and subsystem monitoring tools. Preferred Skills Experience working in large enterprise or financial services environments. Knowledge of automation and scripting for operational efficiency. Familiarity with disaster recovery and business continuity planning. Exposure to capacity planning and performance engineering. Experience supporting highly available multi-LPAR environments.

Other Requirements

 Key Competencies Mainframe Infrastructure Administration Incident & Problem Management Performance Monitoring & Capacity Planning System Recovery & Disaster Recovery RACF Security Administration Service Level Management Process Governance & Compliance Customer Relationship Management Continuous

Information at a Glance

Why HCLTech?

At HCLTech, you'll supercharge your potential. You'll find your career. And you'll find your spark. All at a place that knows that helping its customers stay on top starts by putting its people first.

HCLTech is a global technology company, home to more than 223,000 people across 60 countries, delivering industry-leading capabilities centered around digital, engineering, cloud and AI, powered by a broad portfolio of technology services and products. We work with clients across all major verticals, providing industry solutions for Financial Services, Manufacturing, Life Sciences and Healthcare, Technology and Services, Telecom and Media, Retail and CPG, and Public Services. Consolidated revenues as of 12 months ending June 2026 totaled $14.8 billion.