Senior Technical Lead
India
Job Description
Senior Technical Lead
Hyderabad, Telangana

Job Summary

Observability

 

Job Description — Observability Engineer / Platform Engineer

Role Overview

We are looking for a mid-level Observability Engineer / Platform Engineer to support our transition from a self-managed observability stack to a more standardized and fully managed/SaaS model.

The role focuses on improving observability standards, simplifying operations, and assisting with the migration of monitoring, logging, and tracing into a modern, cloud-based environment.

This position is ideal for someone with solid hands-on skills in observability tools and cloud-native environments, who is ready to contribute to a structured migration initiative.


Key Responsibilities

  • Support the improvement and streamlining of our observability platform across logging, metrics, and tracing.
  • Assist in the migration from a self-hosted stack to a managed/SaaS observability platform.
  • Work with engineering teams to establish consistent dashboards, alerts, and basic service monitoring practices.
  • Help define and roll out foundational observability standards (dashboards, alerting rules, instrumentation basics).
  • Ensure data collection pipelines (agents/collectors) are configured, stable, and efficient.
  • Contribute to ongoing health, reliability, and performance of the observability tooling.
  • Collaborate with platform, application, and security teams on integration and onboarding tasks.

Required Skills

Candidates should have 3–6 years’ experience with most of the following:

Technical Skills

  • Experience with observability tools such as Grafana, Prometheus, Loki, or similar.
  • Good understanding of Kubernetes (EKS or equivalent) and cloud-native environments.
  • Familiarity with OAC(Observability as a code), metrics, logs, and traces fundamentals.
  • Experience deploying/configuring agents or collectors (Grafana Alloy, Promtail, OTEL Collector, etc.).
  • Strong understanding of dashboards, alerting, and common monitoring patterns.
  • Basic Infrastructure-as-Code experience (Terraform, Helm, GitOps workflows).

General Skills

  • Comfortable working across multiple teams and supporting migration activities.
  • Strong troubleshooting skills—able to investigate issues across applications and infrastructure.
  • Good communication skills: able to explain monitoring requirements to engineers and stakeholders.

Nice-to-Have Skills

  • Exposure to Grafana Cloud or other SaaS observability services.
  • Knowledge of SLIs/SLOs and reliability best practices.
  • Experience with multi-cloud or hybrid observability data sources.
  • Understanding of network fundamentals (VPCs, connectivity concepts).

Key Responsibilities

Observability

 

Job Description — Observability Engineer / Platform Engineer

Role Overview

We are looking for a mid-level Observability Engineer / Platform Engineer to support our transition from a self-managed observability stack to a more standardized and fully managed/SaaS model.

The role focuses on improving observability standards, simplifying operations, and assisting with the migration of monitoring, logging, and tracing into a modern, cloud-based environment.

This position is ideal for someone with solid hands-on skills in observability tools and cloud-native environments, who is ready to contribute to a structured migration initiative.


Key Responsibilities

  • Support the improvement and streamlining of our observability platform across logging, metrics, and tracing.
  • Assist in the migration from a self-hosted stack to a managed/SaaS observability platform.
  • Work with engineering teams to establish consistent dashboards, alerts, and basic service monitoring practices.
  • Help define and roll out foundational observability standards (dashboards, alerting rules, instrumentation basics).
  • Ensure data collection pipelines (agents/collectors) are configured, stable, and efficient.
  • Contribute to ongoing health, reliability, and performance of the observability tooling.
  • Collaborate with platform, application, and security teams on integration and onboarding tasks.

Required Skills

Candidates should have 3–6 years’ experience with most of the following:

Technical Skills

  • Experience with observability tools such as Grafana, Prometheus, Loki, or similar.
  • Good understanding of Kubernetes (EKS or equivalent) and cloud-native environments.
  • Familiarity with OAC(Observability as a code), metrics, logs, and traces fundamentals.
  • Experience deploying/configuring agents or collectors (Grafana Alloy, Promtail, OTEL Collector, etc.).
  • Strong understanding of dashboards, alerting, and common monitoring patterns.
  • Basic Infrastructure-as-Code experience (Terraform, Helm, GitOps workflows).

General Skills

  • Comfortable working across multiple teams and supporting migration activities.
  • Strong troubleshooting skills—able to investigate issues across applications and infrastructure.
  • Good communication skills: able to explain monitoring requirements to engineers and stakeholders.

Nice-to-Have Skills

  • Exposure to Grafana Cloud or other SaaS observability services.
  • Knowledge of SLIs/SLOs and reliability best practices.
  • Experience with multi-cloud or hybrid observability data sources.
  • Understanding of network fundamentals (VPCs, connectivity concepts).

Skill Requirements

null

Other Requirements

null
Information at a Glance

Why HCLTech?

At HCLTech, you'll supercharge your potential. You'll find your career. And you'll find your spark. All at a place that knows that helping its customers stay on top starts by putting its people first.

HCLTech is a global technology company, home to more than 223,000 people across 60 countries, delivering industry-leading capabilities centered around digital, engineering, cloud and AI, powered by a broad portfolio of technology services and products. We work with clients across all major verticals, providing industry solutions for Financial Services, Manufacturing, Life Sciences and Healthcare, Technology and Services, Telecom and Media, Retail and CPG, and Public Services. Consolidated revenues as of 12 months ending June 2026 totaled $14.8 billion.