Own and operate the AWS and Kubernetes platform, taking responsibility for reliability, cost, and performance.
Lead GitOps deployments and manage CI/CD end-to-end with GitLab pipelines.
Establish platform observability and maintain security baselines.
Veeva Systems is a mission-driven organization and pioneer in industry cloud, helping life sciences companies bring therapies to patients faster. It is one of the fastest-growing SaaS companies in history, surpassing $3B in revenue in its last fiscal year, with a culture focused on values like Do the Right Thing, Customer Success, Employee Success, and Speed.
Handle infrastructure work across a growing scope — from hand-picked, low-urgency familiarization tasks toward independently managing significant components.
Troubleshoot and resolve infrastructure issues with autonomy, seeking guidance on necessary changes and architecture decisions.
Engage in infrastructure-as-code projects and play an active role in CI/CD pipeline management.
Planning Center exists to help churches build stronger, more connected communities. Since 2006, over 80,000 churches have trusted our tools, and we are an independent, profitable company with no outside investors, operating with a fully remote team.
Own the technical direction and architecture of critical infrastructure domains, establishing scalable patterns and standards.
Lead complex, multi-team infrastructure initiatives from design through implementation and production operation.
Design and evolve AWS and Kubernetes infrastructure to enable teams to build and deploy systems reliably at scale.
We provide innovative identity and risk solutions, empowering institutions and individuals to transact with confidence. Our company is backed by world-class investors including Craft Ventures and Andreessen Horowitz, with offices across the US and India, and we are growing extremely quickly.
Build and run monitoring, tracing, and alerting infrastructure to ensure platform reliability and security.
Lead incident response and recovery, including root cause analysis, and improve deployment processes for fast, safe code changes.
Collaborate with engineering teams to deliver a stable, scalable platform and handle load for resource-intensive applications.
WellSaid Labs is the leading AI voiceover studio for enterprise and professional use, providing ultra-realistic voices that the world’s biggest brands trust. We are a fully distributed team across the U.S. with a focus on responsible AI and an inclusive culture.
Architect the end-to-end reliability, performance, and resilience of cloud environments, including the SLO framework for critical services.
Lead incident response, on-call rotation, root cause analysis, and build a culture of corrective actions.
Build observability platforms to detect issues proactively and mentor engineers on reliability standards.
Garner is on a mission to transform the U.S. healthcare system by partnering with employers to steer members to better-performing doctors, resulting in better care and lower costs. With 550+ proprietary clinical metrics, they have helped over 2.5 million people and saved $1B in healthcare costs, recently raising a Series E and doubling five years running.
Lead and grow a team of platform engineers, coaching them on infrastructure and cloud challenges.
Drive the platform roadmap, balancing reliability, cost, security, and developer experience with AWS and Kubernetes.
Partner cross-functionally to align platform priorities with business goals and ensure system reliability.
PerfectServe is a leading provider of clinical communication and physician scheduling solutions in the health IT space. The company has 400+ employees and 30,000+ customers, with over $100 million in annual revenue, and has received multiple Best in KLAS awards.
Design and implement infrastructure using Terraform, Python, and Kubernetes on AWS.
Collaborate with engineering and data science teams to improve cloud infrastructure.
Automate CI/CD pipelines and enforce security governance and compliance.
Lyra Health is a mental health care provider serving 20 million people through employer and health plan partnerships. The company has delivered 15 million sessions and published 35 peer-reviewed studies, with a culture focused on clinical effectiveness.
Design and operate scalable cloud infrastructure across AWS and GCP.
Build and improve Kubernetes, Linux, and cloud networking environments.
Strengthen security, disaster recovery, and platform resilience.
Hubstaff provides workforce analytics and time tracking for remote teams, serving over 200,000 global users. The company is a product-led organization with a winning culture and a fully remote team of experienced engineers.
Design, build, and maintain robust, scalable, and secure infrastructure systems supporting Laurel's AI-driven platform.
Manage and optimize cloud infrastructure (AWS and Azure), Kubernetes orchestration, and CI/CD pipelines to increase deployment frequency and reliability.
Implement comprehensive observability, monitoring, and alerting to maintain system health and partner with engineering teams to optimize performance and cost-efficiency.
Laurel is an AI Time platform for professional services firms, automating work time capture and connecting time data to business outcomes for clients like EY and Crowell & Moring. The company comprises top AI, product, and engineering talent, is VC-backed by Google Ventures and IVP, and fosters an inclusive, ambitious culture.
Lead CI/CD pipeline optimization and container orchestration to scale infrastructure across teams.
Architect secure infrastructure with secrets management, IAM, and vulnerability scanning as default.
Drive platform reliability through SLOs, observability, and incident response.
Evolve aims to make vacation rental easy for everyone. The team is high-performing, customer-obsessed, and runs on curiosity, communication, and accountability.
Design, build, and operate shared cloud infrastructure using AWS, Kubernetes, Terraform, Databricks, and Cloudflare.
Deliver SRE and DevOps initiatives to improve reliability, scalability, observability, and deployment safety.
Build reusable infrastructure modules, automation, and self-service workflows to reduce manual work and improve developer experience.
YipitData is the leading market research and analytics firm for the disruptive economy, recently raising up to $475M from The Carlyle Group at a valuation over $1B. We analyze billions of alternative data points daily and have been recognized as one of Inc’s Best Workplaces, cultivating a people-centric culture focused on mastery, ownership, and transparency.
Support and optimize AWS infrastructure for Amazon Connect and enterprise cloud platforms.
Manage Infrastructure as Code using Terraform and automate tasks with Python and AWS CLI.
Develop dashboards, alerts, and monitoring solutions using CloudWatch, Dynatrace, Splunk, Grafana, and OpenTelemetry.
Miratech is a global IT services and consulting company that helps visionaries change the world by bringing together enterprise and start-up innovation. The company retains nearly 1000 full-time professionals, operates in over 25 countries, and has a culture of Relentless Performance with a 99% project success rate since 1989.
Design and build scalable, secure cloud infrastructure and deployment pipelines.
Own infrastructure end-to-end, including architecture, provisioning, deployment, and operation.
Lead technical design discussions and contribute to infrastructure and platform architecture decisions.
VulnCheck is the Exploit Intelligence Company, delivering structured exploit intelligence for cybersecurity. Founded in 2021, the company has a transparent, collaborative, and supportive culture with a team of experts.
Design and implement scalable cloud infrastructure to support growth.
Develop monitoring, alerting, and incident response for system reliability.
Automate deployment pipelines and ensure high availability and security.
Tekmetric is the all-in-one, cloud-based software helping auto repair shops run smarter, grow faster, and serve customers better. Founded in Houston in 2017, we've grown into an industry-leading team of builders who value transparency, integrity, and a service-first mindset.
Improve system availability, scalability, and resilience across Flowcode's platforms.
Manage and scale core AWS infrastructure through Infrastructure as Code (Terraform) and enhance disaster recovery.
Oversee monitoring, logging, and alerting infrastructure, and develop high-signal metrics and dashboards.
Flowcode is a technology company specializing in QR code and smart link solutions for offline-to-online engagement. The company is a growth-stage startup seeking high-performing individuals who thrive in a fast-paced, demanding environment.
Design and build resilient AWS and Kubernetes platforms to improve reliability, scalability, and security.
Define SLOs, build observability, automate operational work, and lead incident response and post-incident reviews.
Partner with engineering, platform, security, and QA teams to establish reliability standards and optimize cost.
Electric Power Engineers (EPE) provides consulting expertise and energy intelligence software solutions for power and energy clients, focusing on renewable energy and grid modernization. With over half a century in the industry, the company fosters innovation and collaboration, working with industry leaders to build a secure and resilient grid.
Design, build, and optimize cloud infrastructure (AWS/Kubernetes/EKS) and CI/CD pipelines across multiple teams.
Troubleshoot and resolve production incidents of varying scope, ensuring reliability and performance.
Drive infrastructure projects end-to-end, mentor engineers, and establish standards that improve developer productivity.
Pacvue is a leading Commerce Media OS powering over $12B in advertising spend across 100+ global retail media networks. It enables over 70,000 brands and agencies with an inclusive global community that fosters innovation and career growth.
Create and test reliable cloud infrastructure services supporting Webflow's product range.
Lead initiatives to reduce triage load, increase reliability, and handle growing customer scale.
Collaborate with product engineering teams to deliver new solutions and improve existing services.
Webflow is an agentic web marketing platform that helps modern marketing teams build, manage, and optimize high-performing web experiences. The company values grit, speed, and craft, fostering a culture of ownership and continuous improvement.
Empower engineers on other teams by maintaining monitoring tooling and collaborating on observability best practices.
Enhance reliability of Kubernetes applications through resource optimization, streamlined upgrades, and scalability.
Participate in on-call and incident response processes, occasionally diving into application code to debug production issues.
Webflow is the agentic web marketing platform for modern marketing teams, helping organizations build, manage, and optimize high-performing web experiences. It serves over 2 million users worldwide across 190 countries, with tens of thousands of projects launched each month, and fosters a culture of grit, speed, and craft.
Own infrastructure as code and build self-service paths for product teams.
Make reliability a property of the delivery path with SLOs and alerting.
Shift Left security and manage cloud costs as an engineering responsibility.
what3words is a global addressing company that assigns unique three-word addresses to every 3m square on Earth, making locations precise and easy to share. Their technology is used by emergency services, delivery companies, and automakers across 193 countries, with a growing user base and a microservices architecture on AWS.