Job Summary
Seeking a Cloud Operations / Systems Engineer (5–8 yrs) to manage enterprise cloud infrastructure, production systems, and mission-critical hybrid environments. Role focuses on platform stability, operational continuity, infrastructure servicing, environment health, and incident remediation.
Key Responsibilities
Job Responsibilities : 1. Cloud Infrastructure Operations & Environment Sustainment • Manage LH/SH/Internal Production environments ensuring availability and stability • Maintain OEM baselines, security compliance, and servicing alignment • Execute patching, firmware updates, and configuration management • Monitor compute, storage, virtualization, and infrastructure health ________________________________________ 2. Patch Management & Servicing Operations • Execute patch deployments, firmware upgrades, and servicing cycles • Validate patches, hotfixes, and servicing milestones • Perform rollback validation and post-servicing verification ________________________________________ 3. Environment Monitoring & Systems Reliability • Monitor infrastructure telemetry, performance, and stability • Conduct health assessments across compute, storage, and hardware layers • Identify anomalies, recurring failures, and utilization risks • Improve monitoring to reduce downtime and enhance resilience ________________________________________ 4. Incident Management & Root Cause Analysis • Investigate failures, servicing issues, and instability scenarios • Manage IcM incidents and ADO defects end-to-end • Analyze logs, telemetry, and hardware alerts ________________________________________ 5. Hardware Monitoring & Infrastructure Stability • Monitor OEM and lab hardware health • Perform hardware validation and firmware checks • Support diagnostics, remediation, and maintenance tracking • Ensure infrastructure stability and operational continuity
Skill Requirements
Skill Requirement : Core Competencies • Cloud Infrastructure Operations • Windows/Linux Systems Administration • Azure DevOps (ADO) Operations & Incident Management (ICM) • Enterprise Patch & Release Servicing • Virtualization & Hybrid Environment Support
Other Requirements
Job Description : 1. Cloud Infrastructure Operations & Environment Sustainment\\\\r\\\\n• Manage LH/SH/Internal Production environments ensuring availability and stability\\\\r\\\\n• Maintain OEM baselines, security compliance, and servicing alignment\\\\r\\\\n• Execute patching, firmware updates, and configuration management\\\\r\\\\n• Monitor compute, storage, virtualization, and infrastructure health\\\\r\\\\n________________________________________\\\\r\\\\n2. Patch Management & Servicing Operations\\\\r\\\\n• Execute patch deployments, firmware upgrades, and servicing cycles\\\\r\\\\n• Validate patches, hotfixes, and servicing milestones\\\\r\\\\n• Perform rollback validation and post-servicing verification\\\\r\\\\n________________________________________\\\\r\\\\n3. Environment Monitoring & Systems Reliability\\\\r\\\\n• Monitor infrastructure telemetry, performance, and stability\\\\r\\\\n• Conduct health assessments across compute, storage, and hardware layers\\\\r\\\\n• Identify anomalies, recurring failures, and utilization risks\\\\r\\\\n• Improve monitoring to reduce downtime and enhance resilience\\\\r\\\\n________________________________________\\\\r\\\\n4. Incident Management & Root Cause Analysis\\\\r\\\\n• Investigate failures, servicing issues, and instability scenarios\\\\r\\\\n• Manage IcM incidents and ADO defects end-to-end\\\\r\\\\n• Analyze logs, telemetry, and hardware alerts\\\\r\\\\n________________________________________\\\\r\\\\n5. Hardware Monitoring & Infrastructure Stability\\\\r\\\\n• Monitor OEM and lab hardware health\\\\r\\\\n• Perform hardware validation and firmware checks\\\\r\\\\n• Support diagnostics, remediation, and maintenance tracking\\\\r\\\\n• Ensure infrastructure stability and operational continuity\\\\r\\\\n