Similar Jobs
See allSenior Site Reliability Engineer
NIQ
Global
AWS
Kubernetes
Terraform
Site Reliability Engineer
Glia
Canada
AWS
Kubernetes
Python
Analista de SRE Pleno
Experian
Global
AWS
Kubernetes
Terraform
Senior DevOps Engineer - Embedded Payments
Xplor Technologies
Global
AWS
Terraform
DevOps
Senior DevOps Engineer
Waymark
US
AWS
Terraform
Ansible
Observability & Monitoring:
- Design, implement, and maintain monitoring and alerting systems for production and development environments.
- Leverage tools like Prometheus, Grafana, DataDog, Elastic Stack to track system performance and application health.
- Proactively detect and troubleshoot performance bottlenecks, infrastructure issues, and failures.
Reliability Engineering:
- Optimize performance and reliability of highly available systems supporting healthcare payment applications.
- Lead incident response efforts, including root cause analysis, and implement measures to prevent recurrence.
- Develop and maintain SLOs and SLIs to measure system reliability and availability.
Automation & CI/CD Pipelines:
- Create and maintain automated deployment pipelines (CI/CD) to reduce release cycle times.
- Automate infrastructure provisioning and management through tools such as Terraform, Helm, Ansible, or CloudFormation.
- Improve operational efficiency of development and deployment processes.
LMI
LMI is a digital solutions provider accelerating government impact with innovation and speed, bringing commercial-grade platforms and mission-ready AI to federal agencies. Headquartered in Tysons, Virginia, LMI serves the defense, space, healthcare, and energy sectors, focusing on agility and collaboration to drive impactful results.