Job Summary
Observability
Job Description — Observability Engineer / Platform Engineer
Role Overview
We are looking for a mid-level Observability Engineer / Platform Engineer to support our transition from a self-managed observability stack to a more standardized and fully managed/SaaS model.
The role focuses on improving observability standards, simplifying operations, and assisting with the migration of monitoring, logging, and tracing into a modern, cloud-based environment.
This position is ideal for someone with solid hands-on skills in observability tools and cloud-native environments, who is ready to contribute to a structured migration initiative.
Key Responsibilities
- Support the improvement and streamlining of our observability platform across logging, metrics, and tracing.
- Assist in the migration from a self-hosted stack to a managed/SaaS observability platform.
- Work with engineering teams to establish consistent dashboards, alerts, and basic service monitoring practices.
- Help define and roll out foundational observability standards (dashboards, alerting rules, instrumentation basics).
- Ensure data collection pipelines (agents/collectors) are configured, stable, and efficient.
- Contribute to ongoing health, reliability, and performance of the observability tooling.
- Collaborate with platform, application, and security teams on integration and onboarding tasks.
Required Skills
Candidates should have 3–6 years’ experience with most of the following:
Technical Skills
- Experience with observability tools such as Grafana, Prometheus, Loki, or similar.
- Good understanding of Kubernetes (EKS or equivalent) and cloud-native environments.
- Familiarity with OAC(Observability as a code), metrics, logs, and traces fundamentals.
- Experience deploying/configuring agents or collectors (Grafana Alloy, Promtail, OTEL Collector, etc.).
- Strong understanding of dashboards, alerting, and common monitoring patterns.
- Basic Infrastructure-as-Code experience (Terraform, Helm, GitOps workflows).
General Skills
- Comfortable working across multiple teams and supporting migration activities.
- Strong troubleshooting skills—able to investigate issues across applications and infrastructure.
- Good communication skills: able to explain monitoring requirements to engineers and stakeholders.
Nice-to-Have Skills
- Exposure to Grafana Cloud or other SaaS observability services.
- Knowledge of SLIs/SLOs and reliability best practices.
- Experience with multi-cloud or hybrid observability data sources.
- Understanding of network fundamentals (VPCs, connectivity concepts).
Key Responsibilities
Observability
Job Description — Observability Engineer / Platform Engineer
Role Overview
We are looking for a mid-level Observability Engineer / Platform Engineer to support our transition from a self-managed observability stack to a more standardized and fully managed/SaaS model.
The role focuses on improving observability standards, simplifying operations, and assisting with the migration of monitoring, logging, and tracing into a modern, cloud-based environment.
This position is ideal for someone with solid hands-on skills in observability tools and cloud-native environments, who is ready to contribute to a structured migration initiative.
Key Responsibilities
- Support the improvement and streamlining of our observability platform across logging, metrics, and tracing.
- Assist in the migration from a self-hosted stack to a managed/SaaS observability platform.
- Work with engineering teams to establish consistent dashboards, alerts, and basic service monitoring practices.
- Help define and roll out foundational observability standards (dashboards, alerting rules, instrumentation basics).
- Ensure data collection pipelines (agents/collectors) are configured, stable, and efficient.
- Contribute to ongoing health, reliability, and performance of the observability tooling.
- Collaborate with platform, application, and security teams on integration and onboarding tasks.
Required Skills
Candidates should have 3–6 years’ experience with most of the following:
Technical Skills
- Experience with observability tools such as Grafana, Prometheus, Loki, or similar.
- Good understanding of Kubernetes (EKS or equivalent) and cloud-native environments.
- Familiarity with OAC(Observability as a code), metrics, logs, and traces fundamentals.
- Experience deploying/configuring agents or collectors (Grafana Alloy, Promtail, OTEL Collector, etc.).
- Strong understanding of dashboards, alerting, and common monitoring patterns.
- Basic Infrastructure-as-Code experience (Terraform, Helm, GitOps workflows).
General Skills
- Comfortable working across multiple teams and supporting migration activities.
- Strong troubleshooting skills—able to investigate issues across applications and infrastructure.
- Good communication skills: able to explain monitoring requirements to engineers and stakeholders.
Nice-to-Have Skills
- Exposure to Grafana Cloud or other SaaS observability services.
- Knowledge of SLIs/SLOs and reliability best practices.
- Experience with multi-cloud or hybrid observability data sources.
- Understanding of network fundamentals (VPCs, connectivity concepts).