Lead and grow a team of platform engineers, coaching them on infrastructure and cloud challenges.
Drive the platform roadmap, balancing reliability, cost, security, and developer experience with AWS and Kubernetes.
Partner cross-functionally to align platform priorities with business goals and ensure system reliability.
PerfectServe is a leading provider of clinical communication and physician scheduling solutions in the health IT space. The company has 400+ employees and 30,000+ customers, with over $100 million in annual revenue, and has received multiple Best in KLAS awards.
Own and evolve Quansight's cloud infrastructure across AWS, Azure, and GCP.
Lead infrastructure engagements for clients from scoping through delivery.
Contribute to open-source projects and participate in upstream communities.
Quansight is rooted in the Python data science community and helps companies build sustainable solutions on open-source software. The team is a small, collaborative, fully distributed group of open-source maintainers and engineers.
Own the infrastructure end-to-end for ScaleOps' self-hosted and SaaS platforms.
Manage cloud infrastructure across AWS, GCP, and Azure, including networking, security, and compute.
Collaborate with customers and internal teams to ensure rapid feature delivery without compromising reliability.
ScaleOps is redefining autonomous cloud and AI infrastructure, freeing DevOps from manual resource management. Backed by $210M+ in funding, they are trusted by leading enterprises and Fortune 100 companies, with a fast-paced, innovative culture.
Lead a distributed team of engineers, mentoring them in technical and professional growth.
Manage technical execution, prioritization, and cross-team collaboration on infrastructure projects.
Foster an inclusive, high-performance team environment with continuous feedback and career development.
Pilot provides small businesses with dedicated finance experts and custom software for accurate bookkeeping and financial management. With over 3,000 customers and $170 million in funding, the company values high trust, ownership, and continuous iteration.
Design, build, and maintain robust, scalable, and secure infrastructure systems supporting Laurel's AI-driven platform.
Manage and optimize cloud infrastructure (AWS and Azure), Kubernetes orchestration, and CI/CD pipelines to increase deployment frequency and reliability.
Implement comprehensive observability, monitoring, and alerting to maintain system health and partner with engineering teams to optimize performance and cost-efficiency.
Laurel is an AI Time platform for professional services firms, automating work time capture and connecting time data to business outcomes for clients like EY and Crowell & Moring. The company comprises top AI, product, and engineering talent, is VC-backed by Google Ventures and IVP, and fosters an inclusive, ambitious culture.
Lead the discovery, design, and delivery of complex reliability and infrastructure initiatives, translating ambiguous problems into robust technical solutions.
Define and operate reliability practices including SLOs, SLIs, error budgets, alerting strategies, and observability standards.
Manage and scale production Kubernetes environments, build cloud infrastructure on AWS, and mentor less-senior engineers.
Jobgether is an AI-powered recruitment platform that connects candidates with partner companies through automated matching. They are a globally distributed organization focused on fair, efficient hiring processes.
Design and implement cloud infrastructure using AWS, Azure, and Terraform.
Manage Kubernetes clusters for container orchestration and deployment.
Collaborate with cross-functional teams to ensure scalability and reliability.
BSC Analytics is a leader in advanced data analytics for highly regulated enterprises, providing technical strategy and teams of exclusively senior talent. The company fosters a culture of senior expertise, tackling the toughest data challenges.
Design, build, and operate services and automations to manage Kubernetes clusters at scale, partnering with product management and technical leadership.
Drive rigorous code reviews and maintain high testing standards across the platform.
Manage cloud configurations across AWS and Azure using Terraform, ensuring deep observability and reliability.
Twilio is shaping the future of communications by delivering innovative solutions to hundreds of thousands of businesses and empowering millions of developers. They are a remote-first company with a strong culture of connection and global inclusion, employing a vibrant and diverse team.
Design and implement enterprise-scale Azure cloud platforms and infrastructure automation frameworks.
Lead CI/CD pipelines, Infrastructure as Code, and DevOps practices across the software delivery lifecycle.
Collaborate with cross-functional teams to translate business goals into technical roadmaps and scalable solutions.
Our partner is a technology organization focused on cloud transformation and infrastructure engineering. They foster a collaborative, innovative environment that values ownership and continuous learning.
Lead cloud infrastructure strategy for resilient, secure, and cost-efficient multi-account cloud environments across AWS, Azure, and GCP.
Drive Kubernetes excellence as a technical authority for production clusters including EKS and AKS.
Advance AI-enabled operations by introducing LLM-based tooling and agentic workflows to improve infrastructure development and operational efficiency.
Jobgether is a platform that uses AI-powered matching to connect candidates with hiring companies. They are a technology company focused on improving the hiring process through automation and data analysis.
Design, provision, and maintain Azure cloud infrastructure using infrastructure-as-code.
Build and improve CI/CD pipelines to enable fast, reliable software delivery.
Implement observability practices and support incident response to ensure system reliability.
INNERGY transforms the woodworking industry with cloud-based ERP software for custom manufacturers. Founded in 2016, we are a globally distributed team of 200+ professionals united by deep expertise and a passion for solving real-world problems.
Build and operate backend services at scale, working with Kubernetes, Terraform, and multi-cloud infrastructure across AWS and Azure.
Participate in code reviews, follow testing standards, and maintain observability coverage including metrics, alerts, and distributed tracing.
Collaborate with senior engineers, seek mentorship, and contribute to sprint planning while growing technical skills.
Twilio is shaping the future of communications, delivering innovative solutions to hundreds of thousands of businesses and empowering millions of developers worldwide. The company is remote-first with a strong culture of connection and global inclusion, employing a diverse team that makes a global impact.
Define and lead platform engineering strategy across complex, multi-environment cloud systems.
Architect scalable Kubernetes platforms, own IaC standards, and drive DevSecOps implementation.
Mentor engineers, partner with leadership on infrastructure direction, and lead complex migrations.
Robots & Pencils is an applied AI engineering firm that designs and ships AI co-workers for enterprise operations. Founded in 2009, with delivery centers in Canada, the US, Eastern Europe, and Latin America, we are a nimble team of senior engineers averaging 15+ years of experience.
Own the technical direction and architecture of critical infrastructure domains, establishing scalable patterns and standards.
Lead complex, multi-team infrastructure initiatives from design through implementation and production operation.
Design and evolve AWS and Kubernetes infrastructure to enable teams to build and deploy systems reliably at scale.
We provide innovative identity and risk solutions, empowering institutions and individuals to transact with confidence. Our company is backed by world-class investors including Craft Ventures and Andreessen Horowitz, with offices across the US and India, and we are growing extremely quickly.
Design, build, and operate Kubernetes infrastructure for AI workloads using Terraform and GitOps.
Define SLOs, run incident response, and create runbooks for reliable AI platform operations.
Drive AI-specific observability, FinOps, and security practices across the platform.
We are an AI-native consulting partner working with clients like PayPal, adidas, and NatWest to build digital products and services. Our team of over 600 has scaled quickly, earning Great Place to Work-Certified status multiple years in a row.
Own Primer's internal developer platform end to end, including CI/CD pipelines, deployment workflows, and self-service tooling.
Build the human-AI development loop, creating tooling and automation for coding agent workflows.
Treat developer productivity as a measurable system, using frameworks like DORA to identify and fix delivery bottlenecks.
Primer provides a unified infrastructure for global payments, enabling finance and payments teams to reduce complexity and capture revenue. Backed by top investors like Sofina and Accel, they operate as a remote-first, async culture with high autonomy and low bureaucracy.
Lead and mentor a distributed engineering team across US and EU, fostering growth and collaboration.
Drive proactive ownership and engineering excellence for a cloud platform at exabyte scale.
Architect global infrastructure expansions across multi-cloud and compliance environments.
New Relic is an intelligent observability platform that helps companies gain insight into their complex systems and thrive in an AI-first world. As a global team of innovators, we foster a diverse, welcoming, and inclusive environment where employees can be their authentic selves.
Lead Cloud Platform and SRE teams, driving infrastructure strategy and ownership including Kubernetes, GCP, and Terraform.
Champion SRE culture, define SLOs, SLAs, and enhance observability and incident management.
Own the developer-facing platform as a product, ensuring self-service infrastructure and security compliance (SOC2, ISO-27001).
Prolific builds human data infrastructure for AI development, connecting researchers with a global participant pool. They foster a culture of high performance, ownership, and cross-functional collaboration.
Design, maintain, and improve cloud infrastructure using Azure and Kubernetes.
Automate CI/CD pipelines and operational processes with PowerShell and Python.
Ensure security, reliability, and cost optimization across production environments.
The company provides critical healthcare applications and services through cloud infrastructure. It fosters a collaborative engineering culture focused on learning, knowledge sharing, and continuous improvement.
Empower engineers on other teams by maintaining monitoring tooling and collaborating on observability best practices.
Enhance reliability of Kubernetes applications through resource optimization, streamlined upgrades, and scalability.
Participate in on-call and incident response processes, occasionally diving into application code to debug production issues.
Webflow is the agentic web marketing platform for modern marketing teams, helping organizations build, manage, and optimize high-performing web experiences. It serves over 2 million users worldwide across 190 countries, with tens of thousands of projects launched each month, and fosters a culture of grit, speed, and craft.