Build and maintain production-grade automation using Ansible, Terraform, and Go.
Engage deeply with Kubernetes internals, including scheduler, kubelet, controllers, and CRDs.
Harden platform infrastructure through security best practices like vulnerability scanning, container image signing, and admission controllers.
Vultr makes high-performance cloud infrastructure easy to use, affordable, and locally accessible for enterprises and AI innovators. It is the world's largest privately-held cloud infrastructure company, trusted by hundreds of thousands of customers across 185 countries.
Design, build, and optimize reliable infrastructure for healthcare technology.
Improve scalability, reliability, and performance across distributed systems.
Collaborate with engineers and data professionals to shape modern infrastructure practices.
This company provides innovative healthcare technology solutions. It fosters a remote-first culture with a focus on engineering excellence and collaboration.
Design, build, and maintain highly available Kubernetes infrastructure at scale.
Lead design for components and features, and contribute to architecture decisions for container orchestration.
Mentor engineers on Kubernetes best practices and drive initiatives to improve system reliability.
Marqeta provides a card issuing platform for companies to issue cards, authorize transactions, and manage payment operations in real time. They are a publicly-traded company with a Flex First culture that values remote work and employee growth.
Design, write and deliver software to implement and support large web-scale, highly-performant, highly-available infrastructure on GCP/AWS.
Monitor infrastructure, respond to incidents, correct and improve systems to prevent incidents, and plan capacity.
Tune large-scale clusters for optimal performance and efficiency and support system deployments and product releases.
OpenX develops digital advertising marketplaces and technologies to optimize ad delivery for publishers and advertisers. The company operates a large-scale cloud infrastructure in Poland and values teamwork, customer centricity, and continuous learning.
Manage and maintain cloud infrastructure environments across development, staging, and production to ensure high availability and operational stability.
Administer Kubernetes clusters and containerized workloads, including deployment, scaling, and troubleshooting.
Develop and maintain automation tools and infrastructure-as-code frameworks to improve efficiency and scalability.
The company is a partner organization focused on building scalable, cloud-native infrastructure. The team is growing and operates in a remote-first, collaborative environment with a focus on continuous learning.
Architect and improve cloud foundations on Google Cloud Platform to support scalable, secure, and well-governed workloads.
Design and build platform capabilities across GCP, Kubernetes, CI/CD, GitOps, and developer tooling.
Mentor engineers and raise the technical bar through code review, architecture guidance, and direct implementation.
Wpromote is a digital marketing agency focused on performance marketing and technology. The company fosters a diverse, inclusive culture with a remote-friendly environment and office hubs in Los Angeles, Chicago, and New York.
Build systems for declarative application and infrastructure lifecycle management, including CI/CD, Kubernetes, and service inventory.
Prioritize and troubleshoot infrastructure issues to minimize downtime and respond to alerts efficiently.
Contribute to setting the SRE team's direction and streamline automation of infrastructure processes.
Counterpart Health develops Counterpart Assistant, an AI-enabled primary care tool that supports physicians in chronic disease management. It is a subsidiary of Clover Health, with a remote-first culture and a focus on value-based care through technology.
Design and scale highly reliable platform systems supporting complex cloud-native workloads across multiple deployment environments.
Build and enhance core platform services while contributing to distributed systems, event-driven architectures, and cloud-native infrastructure.
Optimize cloud resources, networking, storage, compute, and observability to improve system performance, scalability, reliability, and maintainability.
Jobgether uses an AI-powered matching process to connect candidates with hiring companies. They operate as a job platform, processing applications and sharing top candidates with employers.
Develop internal tools and automate infrastructure using AWS, Kubernetes, and programming languages.
Research and design solutions to increase website robustness, availability, and cost efficiency.
Collaborate on documentation, code reviews, and rollout of new processes.
Angi powers the future of the home services industry, connecting homeowners with skilled pros. With 9 brands in 8 countries and employees worldwide, Angi has helped homeowners with over 300 million home projects.
Lead deployment and operation of product infrastructure in federal environments within AWS.
Build and maintain scalable, secure cloud-native platforms using Kubernetes, Terraform, and GitLab CI.
Improve development and deployment processes, create tooling for telemetry, and foster documentation culture.
Horizon3.ai is a fast-growing, remote cybersecurity company that helps organizations proactively find and fix exploitable attack vectors. We are a team of former special ops cyber operators and engineers committed to a culture of respect, collaboration, ownership, and results.
Own core compute infrastructure across multiple cloud providers and regions.
Design capabilities for greater performance and flexibility in service deployment.
Investigate and resolve challenging cloud and compute issues across the stack.
Render is a cloud platform for developers building AI-native, full-stack, multi-service applications. Trusted by over 6 million developers, the company has raised $257M in funding and values craft, velocity, and user experience.
Partner with engineering teams to improve reliability, scalability, and operational health of production systems.
Investigate and resolve complex production incidents, designing sustainable long-term solutions.
Design, build, and maintain automation tools and infrastructure to enhance developer productivity.
The company builds and maintains highly reliable, scalable production systems supporting millions of users worldwide. It fosters a remote-first culture that values innovation, collaboration, and engineering excellence.
Manage Kubernetes clusters and maintain infrastructure in the cloud.
Administer Linux servers and implement configuration management using Puppet or Ansible.
Troubleshoot and ensure observability of systems with CI/CD integration.
Xsolla is a global commerce company providing tools and services to help developers solve challenges in the video game industry. They employ over 1,500 developers and cultivate a supportive, collaborative culture focused on creativity and professional growth.
Lead a high-impact infrastructure team, evolving internal platforms and CI/CD systems to support large-scale engineering operations.
Drive automation initiatives and AI-driven practices to reduce operational complexity and improve developer experience.
Define and execute strategies for scalable infrastructure, cloud environments, and platform engineering.
The partner company is a technology organization focused on building infrastructure platforms that enable engineering teams to deliver software faster. It is a remote-first company with a collaborative culture and a focus on innovation and scalability.
Lead infrastructure strategy and cloud architecture for scalability and reliability.
Define enterprise CI/CD strategy and champion Infrastructure as Code (IaC) practices.
Mentor engineering teams and drive SRE principles for operational excellence.
Anaplan optimizes business decision-making through its AI-infused scenario planning and analysis platform. With over 2,400 global customers including Fortune 50 companies, it fosters a culture of innovation, diversity, and winning.
Keep user-facing services and production systems reliable, scalable, and efficient with automation and infrastructure-as-code.
Operate and troubleshoot production systems on Kubernetes, and contribute to observability with metrics, logs, and SLOs.
Participate in on-call, incident response, and post-incident reviews to drive improvements in automation and processes.
GitLab is the intelligent orchestration platform for DevSecOps, enabling organizations to increase developer productivity and improve operational efficiency. With more than 50 million registered users and over 50% of the Fortune 100 as customers, GitLab fosters a high-performance, all-remote culture driven by values and continuous knowledge exchange.
Design, implement, and operate core services that power Docker’s Cloud Sandboxes platform.
Build scalable systems for microVM orchestration, workload scheduling, and lifecycle management.
Ensure system reliability, observability, and performance across Docker’s Cloud Sandbox infrastructure.
Docker is a globally distributed, remote-first company that builds tools for developers to build, share, and run applications. Trusted by over 20 million monthly users and 20 billion container image pulls, it has a collaborative culture focused on innovation and reliability.
Design and implement AI inference and training cloud products optimized for Kubernetes, including autoscaling and distributed jobs across GPU fleets.
Write clean, efficient Go code for Kubernetes controllers, operators, and custom resources supporting AI workloads.
Build APIs, CLIs, and developer tools to simplify deployment, lifecycle management, and monitoring of AI applications.
Gcore is a global provider of infrastructure and software solutions for AI, cloud, network, and security, powering digital experiences worldwide. With 550+ professionals and 210+ edge locations, the company collaborates with partners like Intel, NVIDIA, and Equinix to build the foundation for an AI-driven world.
Perform operational deployments, implementations, and maintenance for production systems.
Implement and maintain monitoring, reporting, and alerting systems for Core Speech products.
Be part of an on-call rotation and work collaboratively to improve system performance and architecture.
Solventum is a new healthcare company with a long legacy of solving big challenges to improve lives and enable healthcare professionals to perform at their best. They are a large company that values empathy, insight, and clinical intelligence, collaborating with top minds in healthcare.