Similar Jobs
See allSr/Staff Site Reliability Engineer, Consumer Apps
Attain
US
AWS
GCP
Kubernetes
Site Reliability Engineer
Blitzy
US
Kubernetes
Terraform
Python
Staff Platform Engineer
Postscript
Global
Kubernetes
AWS
Terraform
Director, DevOps Technology
Inizio Evoke
US
AWS
Terraform
AWS CDK
Site Reliability Engineer
Boson AI
Global
Kubernetes
Linux
Networking
Technical Delivery & Implementation:
- Design, build, and operate Kubernetes infrastructure for AI workloads using Terraform and GitOps.
- Build CI/CD pipelines for AI models, agents, and prompt updates with automated evaluation.
- Define SLOs, run incident response, and implement AI-specific observability.
Client Delivery & Stakeholder Management:
- Work with client workstreams to ensure operability from day one.
- Guide architecture and AWS service choices, building cost-benefit cases for infrastructure decisions.
- Communicate reliability and cost trade-offs to technical and non-technical audiences.
Documentation & Handover:
- Leave runbooks, decision records, and Terraform that client engineers can maintain.
- Bring client engineers along in SRE practice through knowledge transfer.
- Treat clean handover as part of definition of done.
CreateFuture
We are an AI-native consulting partner working with clients like PayPal, adidas, and NatWest to build digital products and services. Our team of over 600 has scaled quickly, earning Great Place to Work-Certified status multiple years in a row.