Design and deliver significant components and core subsystems of our Kubernetes platform, such as secrets management, workload identity, storage, or cluster networking, from design through production operation.
Contribute to the architecture of distributed workloads, working with dependent teams to get runtime and isolation models right, while spending most time hands-on in code.
Own operability of built systems including SLOs, failure modes, upgrades, migrations, and on-call, and mentor earlier-career engineers.
Design, build, and operate components of the Kubernetes platform and its core subsystems end to end.
Write and review Go code for controllers, operators, platform services, and automation.
Help operate the platform, including on-call, incident investigation, and follow-up work to prevent recurrence.
ServiceNow provides an AI platform for business reinvention, helping 85% of the Fortune 500 work smarter. The company fosters an AI-native culture where technology and talent collaborate.
Own the design and implementation of platform subsystems end to end.
Anchor meaningful projects, balancing reliability, scalability, and time to market.
Mentor SWE and SWE II engineers and contribute to a strong code review culture.
Stacklok is building the control plane for enterprise AI agents, enabling organizations to run, govern, and secure them on their own infrastructure. The company is led by Kubernetes co-creators and is a startup with a collaborative, AI-maximalist culture.
Design, build, and operate services and automations to manage Kubernetes clusters at scale, partnering with product management and technical leadership.
Drive rigorous code reviews and maintain high testing standards across the platform.
Manage cloud configurations across AWS and Azure using Terraform, ensuring deep observability and reliability.
Twilio is shaping the future of communications by delivering innovative solutions to hundreds of thousands of businesses and empowering millions of developers. They are a remote-first company with a strong culture of connection and global inclusion, employing a vibrant and diverse team.
Own the infrastructure end-to-end for ScaleOps' self-hosted and SaaS platforms.
Manage cloud infrastructure across AWS, GCP, and Azure, including networking, security, and compute.
Collaborate with customers and internal teams to ensure rapid feature delivery without compromising reliability.
ScaleOps is redefining autonomous cloud and AI infrastructure, freeing DevOps from manual resource management. Backed by $210M+ in funding, they are trusted by leading enterprises and Fortune 100 companies, with a fast-paced, innovative culture.
Lead the architecture and implementation of managed Kubernetes infrastructure across AWS, Azure, and GCP.
Own the systems that provision and manage cloud accounts and subscriptions across providers.
Design and implement the networking layer routing traffic into customer environments.
Ditto builds the world's leading edge sync platform, enabling applications to share data peer-to-peer with or without internet connectivity. With over $145 million in funding and trusted by major organizations, Ditto is a globally distributed, fast-growing startup committed to diversity and inclusion.
Build and maintain the Shadeform GPU platform and automated AI infrastructure services.
Work on novel solutions to GPU market challenges including provisioning, orchestration, and virtualization.
Own the systems that turn fragmented GPU capacity into a reliable, production-ready platform.
Shadeform provides a unified platform for deploying and managing GPU infrastructure across cloud providers, neoclouds, and data centers. They are a remote-first startup focused on innovation and working with the latest AI technologies.
Design and deliver solutions for cloud hosted production infrastructure.
Shape how mission-critical enterprise software solutions are developed and deployed using optimized CI/CD pipelines.
Design, build and support infrastructure and security technologies within the cloud.
Ping Identity provides an intelligent cloud identity platform that enables secure and seamless digital experiences. They serve over half of the Fortune 100 companies and have a global team that values diversity and individuality.
Keep user-facing services and production systems reliable, scalable, and efficient with automation and infrastructure-as-code.
Operate and troubleshoot production systems on Kubernetes, and contribute to observability with metrics, logs, and SLOs.
Participate in on-call, incident response, and post-incident reviews to drive improvements in automation and processes.
GitLab is the intelligent orchestration platform for DevSecOps, enabling organizations to increase developer productivity and improve operational efficiency. With more than 50 million registered users and over 50% of the Fortune 100 as customers, GitLab fosters a high-performance, all-remote culture driven by values and continuous knowledge exchange.
Lead projects with complete autonomy, planning ahead and thinking globally for the company's benefit.
Contribute to operation and deployment of all Dashboard Platform features using Kubernetes, Terraform, and AWS.
Share knowledge on software engineering, testing, and deployment best practices to enhance developer experience.
Algolia is a pioneer and market leader in AI Search, empowering 17,000+ businesses to deliver fast, predictive search and browse experiences. We have raised $150 million in Series D funding, value at $2.25 billion, and foster a high-trust, flexible workplace culture.
End-to-end ownership of internal orchestration platform built on event-driven architecture with Redpanda, including code, architecture, and roadmap.
Own infrastructure-as-code using Terraform Cloud, manage Kubernetes workloads with Helm, and provide self-service tooling for engineering teams.
Set SLOs, handle production on-call, lead incident response, author design docs, and operate AI-natively using tools like Cursor and Notion AI.
Velora unifies Aplos, Raisely, and Keela into one company with a shared mission to help nonprofit organizations thrive by offering fundraising, donor management, financial tracking, and communications tools. We are a financially solid company with a combined team dedicated to making nonprofit work easier, more impactful, and more sustainable.
Build and operate backend services at scale, working with Kubernetes, Terraform, and multi-cloud infrastructure across AWS and Azure.
Participate in code reviews, follow testing standards, and maintain observability coverage including metrics, alerts, and distributed tracing.
Collaborate with senior engineers, seek mentorship, and contribute to sprint planning while growing technical skills.
Twilio is shaping the future of communications, delivering innovative solutions to hundreds of thousands of businesses and empowering millions of developers worldwide. The company is remote-first with a strong culture of connection and global inclusion, employing a diverse team that makes a global impact.
Deploy and scale MCP-based AI agents on Kubernetes for enterprise customers across the US-East and EMEA regions.
Lead complex technical engagements, build reusable deployment patterns, and mentor engineers on the team.
Shape product roadmap by feeding back field insights from regulated industries and defining regional engagement standards.
Stacklok builds the control plane for enterprise AI agents, enabling organizations to run, govern, and secure them on Kubernetes and private cloud. Founded by two Kubernetes creators, the company is already adopted by leading tech and regulated industries, fostering a collaborative, AI-maximalist culture with deep open-source roots.
Design and operate the infrastructure for a high-throughput messaging platform operating at 500K+ events/sec.
Build guardrails, runbooks, and validation gates that enable AI agents to safely execute deployments and operations.
Lead incident response and encode every fix as a new runbook and regression test.
Postscript is an AI messaging platform trusted by 20,000+ Shopify brands to drive revenue through SMS. The company is fully remote, backed by Greylock and Y Combinator, and has a culture of ownership and innovation.
Build and maintain production-grade automation using Ansible, Terraform, and Go.
Engage deeply with Kubernetes internals, including scheduler, kubelet, controllers, and CRDs.
Harden platform infrastructure through security best practices like vulnerability scanning, container image signing, and admission controllers.
Vultr makes high-performance cloud infrastructure easy to use, affordable, and locally accessible for enterprises and AI innovators. It is the world's largest privately-held cloud infrastructure company, trusted by hundreds of thousands of customers across 185 countries.
Architect, secure, and operate production Kubernetes platforms across AWS and hybrid environments.
Design reusable Infrastructure as Code using Terraform, Ansible, or similar tools.
Lead development of secure CI/CD and GitOps workflows using GitLab, Argo CD, Flux, or equivalent technologies.
Raft is a customer-obsessed non-traditional defense tech company empowering U.S. military and government agencies with cutting-edge AI/ML and data solutions. The company's culture values collaboration, innovation, and diversity, with a team focused on building impactful digital solutions.
Build and operate internal platform services and APIs in Go.
Codify infrastructure with Terraform and GitOps practices.
Operate and scale multi-tenant EKS clusters and traffic systems.
Docker builds tools for developers to build, share, and run applications, trusted by over 20 million monthly users. They are a globally distributed, remote-first team with offices in Seattle and Paris, focused on innovation and inclusion.
Define architecture and best practices for the platform and infrastructure layer the product is built on.
Own the deploy pipeline and lead the move to a GitOps model (Argo) for fast, safe releases.
Design and harden multi-tenant isolation and blast-radius protection for top-tier customers, including dedicated deployments.
We are the Engineering Operations Platform - mission control for the AI software factory, providing visibility, governance, and golden paths. We are a group of 80 passionate individuals, backed by $60M Series C from Sequoia, IVP, and others, with a fully remote culture.
Design and build control-plane services, drivers, and tooling for high-performance storage integration with Kubernetes.
Write production-quality Go software with strong testing and operational rigor for automation.
Deliver storage integration for Kubernetes via Cluster API and K0rdent in hybrid and air-gapped environments.
Mirantis is a Kubernetes-native AI infrastructure company enabling organizations to build scalable infrastructure for AI and data-intensive applications. A Silicon Valley leader with a young, collaborative culture.
Design, develop, and maintain CI/CD pipelines using Golang, Python, and Terraform to automate engineering processes.
Build and operate cloud-native platforms with Kubernetes and Docker, ensuring high reliability and scalability.
Collaborate with teams to improve automation, observability, security, and production support across services.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. It uses an automated system to review applications and share top-fitting candidates with employers, focusing on efficiency and fairness in the hiring process.
Design, build, and run distributed cloud architectures and large-scale production systems.
Ensure reliability, observability, performance, and cost efficiency of the platform.
Collaborate with product and backend teams to design system architecture and optimize resource use.
Tinybird helps developers and data teams unlock the power of real-time data, enabling them to build data pipelines and innovative data products quickly. They are a remote-first company with a culture of ownership, transparency, and clear communication.