Similar Jobs

See all

What You'll Do:

  • Ensure the reliability, scalability, and resilience of critical digital environments.
  • Work across cloud infrastructure, Kubernetes, and observability to maintain system stability.
  • Combine proactive engineering with hands-on troubleshooting to prevent production issues.

Key Responsibilities:

  • Define and monitor SLIs, SLOs, SLAs, MTTR, and MTTD.
  • Implement observability, monitoring, and alerting solutions across applications and infrastructure.
  • Automate operational activities and reduce manual tasks through scripting and Infrastructure as Code.

Requirements:

  • Proven experience as an SRE or equivalent role.
  • Practical experience with cloud environments (GCP, AWS, Azure) and Kubernetes/Docker.
  • Strong understanding of SRE concepts and metrics.

Benefits:

  • Meal and food allowance, home office allowance.
  • Medical, dental, and life insurance.
  • Birthday Day Off and wellness support.

Not Disclosed

Our partner is a technology company focused on building and maintaining reliable, scalable digital environments. They promote a culture of continuous improvement, collaboration, and proactive engineering.

Apply for This Position