Similar Jobs

See all

Accountabilities:

  • Define, implement, and monitor SLIs, SLOs, SLAs, MTTR, and MTTD.
  • Implement and evolve observability solutions including monitoring, alerting, dashboards, and APM.
  • Prevent, investigate, and resolve incidents, and conduct root-cause analyses.

Requirements:

  • Professional experience as a Site Reliability Engineer or equivalent reliability role.
  • Practical experience with cloud environments (GCP, AWS, Azure).
  • Solid knowledge of Kubernetes and Docker.

Preferred Qualifications:

  • Experience with GKE, EKS, or AKS.
  • Familiarity with Dynatrace, Datadog, Grafana, Prometheus, ELK.
  • Knowledge of Terraform and Ansible.

Benefits:

  • Meal allowance and food allowance.
  • Medical, dental, and life insurance.
  • Home office allowance and wellness benefit.

Unknown

The company is a technology organization focused on building and maintaining reliable digital environments. It fosters a culture of engineering excellence, collaboration, and data-driven decision-making.

Apply for This Position