Similar Jobs
See allStaff Site Reliability Engineer
Unknown
US
Kubernetes
Python
Go
Principal Site Reliability Engineer
Experian
Brazil
Kubernetes
AWS
Terraform
Site Reliability Engineer
Runpod
Global
Linux
Python
Go
Senior Site Reliability Engineer
Valtech
Portugal
Site Reliability Engineering
DevOps
Cloud Engineering
Senior Director, Site Reliability Engineering
Ping Identity
US
Site Reliability Engineering
Kubernetes
Infrastructure As Code
Role Summary:
- As a Staff Site Reliability Engineer, you are the senior technical authority on the SRE team and a strategic partner to engineering leadership.
- You define the technical standard for how Filevine runs in production and bridge the gap between business goals and technical execution.
- You own the roadmap across Observability & Alerting and Platform Infrastructure and ensure reliability problems are solved permanently.
Who You Are:
- You are a master of the craft with deep expertise in distributed systems, cloud infrastructure, and reliability engineering.
- You are a technical leader and mentor, passionate about investing in the growth of engineers around you.
- You are forward-thinking and AI/ML fluent, driving the use of AI in observability and incident response.
What You Will Do:
- Define and execute the technical strategy for Observability & Alerting, Platform Infrastructure, and operational excellence.
- Lead the evolution of reliable, scalable, secure, and efficient cloud platforms and distributed systems.
- Champion SLIs, SLOs, error budgets, capacity planning, and automation across the service lifecycle.
Filevine
Filevine is a Legal AI company delivering a unified platform for legal work, powered by LOIS (Legal Operating Intelligence System). The company is rapidly growing, recognized by Deloitte and Inc. as one of the most innovative and fastest-growing technology companies.