Own critical infrastructure across compute, networking, CI/CD, Kubernetes, and observability.
Manage Kubernetes environments and infrastructure-as-code with Terraform, improving developer experience and reducing operational friction.
Lead production incident response, influence architecture, and integrate AI-powered tools to boost engineering efficiency.
Jobgether is an AI-powered recruitment platform that connects candidates with global hiring companies. This role is with a partner company, a globally distributed technology organization offering a collaborative, informal culture and long-term opportunities.
Own and drive key infrastructure modernization initiatives toward container-orchestrated infrastructure.
Design and maintain infrastructure as code across multiple cloud providers.
Provide technical leadership and mentorship across the Systems Engineering team.
Intellum is the leader in corporate education technology, powering large learning programs for brands like Google, Meta, and Amazon. We are a remote-first company with a culture that values curiosity, creativity, perseverance, and kindness, and we invest in our people through personal development budgets and annual retreats.
Drive automation and modernize cloud infrastructure using AWS, Kubernetes, and infrastructure as code.
Design and implement enterprise-grade service mesh architectures and CI/CD pipelines.
Collaborate with cross-functional teams to align technical solutions with business and regulatory requirements.
Our partner is a technology company driving digital transformation through modern cloud infrastructure. They are a growing organization with a focus on secure, scalable solutions and engineering excellence.
Develop and evolve foundational software and services enabling product and development teams.\n- Architect, design, and implement Infrastructure as Code using Terraform.\n- Deploy, manage, and optimize Kubernetes clusters on GCP (GKE) and AWS (EKS).
DoiT is a global technology company that helps cloud-driven organizations leverage cloud for business growth and innovation through data, technology, and human expertise. They work with over 4,000 customers worldwide and foster a remote-first, entrepreneurial culture.
Build and operate the Kubernetes platform supporting AI test and evaluation frameworks.
Design infrastructure-as-code, GitOps workflows, and automated deployment pipelines.
Own platform reliability, observability, capacity planning, and operational readiness.
OpenTeams helps enterprises and governments build AI they control, govern, and evolve themselves. Founded by the creator of NumPy and SciPy, the company is built by people with deep roots across the open-source ecosystem and maintains a remote-first culture.
Lead cloud infrastructure strategy for resilient, secure, and cost-efficient multi-account cloud environments across AWS, Azure, and GCP.
Drive Kubernetes excellence as a technical authority for production clusters including EKS and AKS.
Advance AI-enabled operations by introducing LLM-based tooling and agentic workflows to improve infrastructure development and operational efficiency.
Jobgether is a platform that uses AI-powered matching to connect candidates with hiring companies. They are a technology company focused on improving the hiring process through automation and data analysis.
Design and implement infrastructure using Terraform, Python, and Kubernetes on AWS.
Collaborate with engineering and data science teams to improve cloud infrastructure.
Automate CI/CD pipelines and enforce security governance and compliance.
Lyra Health is a mental health care provider serving 20 million people through employer and health plan partnerships. The company has delivered 15 million sessions and published 35 peer-reviewed studies, with a culture focused on clinical effectiveness.
Drive the performance, stability, security, and reliability of production environments with a focus on automation and proactive improvements.
Design and maintain infrastructure using Infrastructure as Code tools like Terraform, and manage Kubernetes and cloud environments.
Lead vulnerability management, incident response, and secure CI/CD practices to ensure resilience and operational excellence.
Jobgether is a platform that uses AI-powered matching to connect candidates with hiring companies. It processes applications and shares shortlists with employers, offering a remote-first and inclusive work environment.
Design and evolve highly available cloud architecture on Google Cloud using Terraform and GitOps.
Build and maintain secure CI/CD pipelines for Infrastructure-as-Code and develop self-service developer platforms.
Strengthen platform observability and apply SRE principles to improve reliability and operational maturity.
They are a financial technology company that provides production-critical infrastructure. They have a globally distributed team and a culture of autonomy and async-first collaboration.
Own core platform infrastructure including Terraform migration, CI/CD pipelines, and environment provisioning as part of a growing team.
Build and maintain tooling and abstractions that let product engineers ship and run code without solving infrastructure problems from scratch.
Advance observability foundation, standardize monitoring and alerting, and support GCP infrastructure scaling.
Astra builds mission-critical infrastructure for moving money at scale, processing billions in annual transaction volume with 99.9%+ uptime. We are a remote-first company hiring within the U.S., with a small team focused on thoughtful collaboration and clarity.
Design and advance core infrastructure for multi-cloud Kubernetes clusters and developer toolchains.
Automate operations and engineering tasks to improve productivity and reliability.
Build machine learning infrastructure to enable AI teams to train and deploy large-scale models.
Cresta provides an AI platform that transforms customer conversations into competitive advantages by combining conversational AI, real-time agent augmentation, and conversation intelligence. The company has raised over $270 million from top investors like a16z, Greylock, and Sequoia, and is led by AI industry veterans.
Drive complex infrastructure migrations and build platform tooling and automation across multiple production environments.
Support development teams by consulting on infrastructure needs and improving observability and incident response.
Provide operational support and maintain platform reliability through structured debugging and on-call rotations.
PENN Entertainment is North America's leading provider of integrated entertainment, sports content, and casino gaming experiences. We operate across numerous locations in North America and foster a culture that cares about career growth and skill expansion.
Support the architecture, deployment, and testing of bare-metal infrastructure including GPU compute, storage, and networking.
Deploy, configure, maintain, and troubleshoot UDS and Kubernetes-based platform services.
Develop and manage infrastructure as code (IaC) to enable consistent, repeatable deployments.
Defense Unicorns delivers mission value by streamlining software delivery for mission-focused customers. The team is composed of innovators, software engineers, and veterans with decades of experience.
Design, build, and maintain robust, scalable, and secure infrastructure systems supporting Laurel's AI-driven platform.
Manage and optimize cloud infrastructure (AWS and Azure), Kubernetes orchestration, and CI/CD pipelines to increase deployment frequency and reliability.
Implement comprehensive observability, monitoring, and alerting to maintain system health and partner with engineering teams to optimize performance and cost-efficiency.
Laurel is an AI Time platform for professional services firms, automating work time capture and connecting time data to business outcomes for clients like EY and Crowell & Moring. The company comprises top AI, product, and engineering talent, is VC-backed by Google Ventures and IVP, and fosters an inclusive, ambitious culture.
Architect, deploy, and manage highly available, fault-tolerant cloud infrastructure across Google Cloud Platform (GCP) and Google Kubernetes Engine (GKE).
Maintain and scale declarative infrastructure using Terraform across a multi-hundred-file estate, enforcing GitOps workflows with Atlantis.
Build, maintain, and optimize robust automated pipelines for continuous integration and delivery using GitHub Actions, Jenkins, and ArgoCD.
Point Wild helps customers monitor, manage, and protect against the risks associated with their identities and personal information in a digital world. Backed by WndrCo, Warburg Pincus and General Catalyst, Point Wild is a scrappy, nimble organization dedicated to creating the world’s most comprehensive portfolio of industry-leading cybersecurity solutions.
Manage and troubleshoot complex distributed large-scale software systems
Build scalable, secure and reliable container-based infrastructure
Automate software delivery processes with CI/CD pipelines
Coinspaid Dev is the engineering brand behind the technology, infrastructure, and R&D expertise built within Coinspaid, focusing on advancing blockchain infrastructure engineering. With over 120 engineers and more than 11 years of industry experience, they bring together teams building distributed systems and blockchain infrastructure across 20+ blockchain networks.
Maintain and extend foundational technologies to provide highly available, distributed services to customers.
Design, implement, and operate large scale, multi-cloud Kubernetes infrastructure driven by GitOps methodology.
Embrace an empathetic, supportive, and communicative environment, pulling from one another's strengths.
InfluxData is the creator of InfluxDB, the leading time series platform for collecting, storing, and analyzing time series data at any scale. They are a remote-first company with a globally distributed workforce.
Own the technical direction of the AWS platform, building cost visibility tooling and managing Kubernetes on EKS.
Design and maintain Terraform modules for self-service infrastructure provisioning and standardize CI/CD across services.
Harden the platform alongside security, lead incident response, and mentor senior engineers through design and code review.
PlayOn powers high school sports ticketing, streaming, and fundraising through platforms like GoFan, NFHS Network, and MaxPreps. Backed by KKR, the company is a growth-stage leader focused on making high school sports more accessible and connected.
Lead the discovery, design, and delivery of complex reliability and infrastructure initiatives, translating ambiguous problems into robust technical solutions.
Define and operate reliability practices including SLOs, SLIs, error budgets, alerting strategies, and observability standards.
Manage and scale production Kubernetes environments, build cloud infrastructure on AWS, and mentor less-senior engineers.
Jobgether is an AI-powered recruitment platform that connects candidates with partner companies through automated matching. They are a globally distributed organization focused on fair, efficient hiring processes.
Design, build, and maintain cloud infrastructure on GCP and AWS using Terraform, optimizing CI/CD pipelines for rapid deployments.
Implement comprehensive observability including monitoring, logging, alerting, and distributed tracing to ensure platform health.
Establish and enforce security best practices, support AI/ML infrastructure, and build developer experience tooling.
Re:Build operates an advanced, end-to-end manufacturing platform that partners with industrial companies to bring products from concept to full-scale production. The company is guided by The Re:Build Way principles and aims to revitalize America's manufacturing base, creating meaningful jobs across the country.