Design and deliver significant components and core subsystems of our Kubernetes platform, such as secrets management, workload identity, storage, or cluster networking, from design through production operation.
Contribute to the architecture of distributed workloads, working with dependent teams to get runtime and isolation models right, while spending most time hands-on in code.
Own operability of built systems including SLOs, failure modes, upgrades, migrations, and on-call, and mentor earlier-career engineers.
ServiceNow is the AI control tower for business reinvention, bringing together any AI, any data, and any workflow to help 85% of the Fortune 500 work smarter, faster, and better. The company fosters an AI-native culture where technology and talent are unstoppable together, with a focus on freeing people from busywork.
Own and evolve Kubernetes and cloud infrastructure on AWS for scalability, reliability, and usability.
Design and improve CI/CD pipelines and developer workflows to enable fast, safe, repeatable deployments.
Work cross-functionally with product engineers to understand needs and enable them through tooling and best practices.
Artsy is an online platform that connects collectors, artists, and gallerists to make the art world more accessible. The company values an inclusive culture and a diverse workforce, with a team that operates with open-source principles and a focus on impact.
Design, build, and operate Kubernetes infrastructure for AI workloads using Terraform and GitOps.
Define SLOs, run incident response, and create runbooks for reliable AI platform operations.
Drive AI-specific observability, FinOps, and security practices across the platform.
We are an AI-native consulting partner working with clients like PayPal, adidas, and NatWest to build digital products and services. Our team of over 600 has scaled quickly, earning Great Place to Work-Certified status multiple years in a row.
Design, build, and operate components of the Kubernetes platform and its core subsystems end to end.
Write and review Go code for controllers, operators, platform services, and automation.
Help operate the platform, including on-call, incident investigation, and follow-up work to prevent recurrence.
ServiceNow provides an AI platform for business reinvention, helping 85% of the Fortune 500 work smarter. The company fosters an AI-native culture where technology and talent collaborate.
Build and maintain the Shadeform GPU platform and automated AI infrastructure services.
Work on novel solutions to GPU market challenges including provisioning, orchestration, and virtualization.
Own the systems that turn fragmented GPU capacity into a reliable, production-ready platform.
Shadeform provides a unified platform for deploying and managing GPU infrastructure across cloud providers, neoclouds, and data centers. They are a remote-first startup focused on innovation and working with the latest AI technologies.
Own the infrastructure end-to-end for ScaleOps' self-hosted and SaaS platforms.
Manage cloud infrastructure across AWS, GCP, and Azure, including networking, security, and compute.
Collaborate with customers and internal teams to ensure rapid feature delivery without compromising reliability.
ScaleOps is redefining autonomous cloud and AI infrastructure, freeing DevOps from manual resource management. Backed by $210M+ in funding, they are trusted by leading enterprises and Fortune 100 companies, with a fast-paced, innovative culture.
Lead projects with complete autonomy, planning ahead and thinking globally for the company's benefit.
Contribute to operation and deployment of all Dashboard Platform features using Kubernetes, Terraform, and AWS.
Share knowledge on software engineering, testing, and deployment best practices to enhance developer experience.
Algolia is a pioneer and market leader in AI Search, empowering 17,000+ businesses to deliver fast, predictive search and browse experiences. We have raised $150 million in Series D funding, value at $2.25 billion, and foster a high-trust, flexible workplace culture.
Lead the architecture and implementation of managed Kubernetes infrastructure across AWS, Azure, and GCP.
Own the systems that provision and manage cloud accounts and subscriptions across providers.
Design and implement the networking layer routing traffic into customer environments.
Ditto builds the world's leading edge sync platform, enabling applications to share data peer-to-peer with or without internet connectivity. With over $145 million in funding and trusted by major organizations, Ditto is a globally distributed, fast-growing startup committed to diversity and inclusion.
Design, build, and operate services and automations to manage Kubernetes clusters at scale, partnering with product management and technical leadership.
Drive rigorous code reviews and maintain high testing standards across the platform.
Manage cloud configurations across AWS and Azure using Terraform, ensuring deep observability and reliability.
Twilio is shaping the future of communications by delivering innovative solutions to hundreds of thousands of businesses and empowering millions of developers. They are a remote-first company with a strong culture of connection and global inclusion, employing a vibrant and diverse team.
Design and operate the infrastructure for a high-throughput messaging platform operating at 500K+ events/sec.
Build guardrails, runbooks, and validation gates that enable AI agents to safely execute deployments and operations.
Lead incident response and encode every fix as a new runbook and regression test.
Postscript is an AI messaging platform trusted by 20,000+ Shopify brands to drive revenue through SMS. The company is fully remote, backed by Greylock and Y Combinator, and has a culture of ownership and innovation.
Design, build, and operate core cloud infrastructure on AWS, including compute, networking, and container orchestration.
Own the CI/CD platform used across engineering teams, including build pipelines, environment promotion, and progressive rollout.
Build and maintain the observability stack across the organization, including logging, metrics, distributed tracing, and alerting.
RxSense is a healthcare technology company that provides platforms and solutions to improve the management and access of cost-effective pharmacy benefits. As a leader in SaaS technology for healthcare, the company is an Equal Opportunity and Affirmative Action employer committed to diversity and collaboration.
Design, build, and scale reliable infrastructure for Klover's fintech platform using modern technologies like Kubernetes, Terraform, and Istio.
Use AI agents as force multipliers to automate manual processes and improve developer experience.
Collaborate with engineering teams to ensure system reliability, performance, and security across production systems.
Attain powers Klover, a fast-growing fintech platform serving over one million active users monthly, processing over $1.5 billion annually. The company emphasizes collaboration, reliability, and innovation, with a culture of automation and AI-driven development.
Lead deployment and operation of product infrastructure in federal environments within AWS.
Build and maintain scalable, secure cloud-native platforms using Kubernetes, Terraform, and GitLab CI.
Improve development and deployment processes, create tooling for telemetry, and foster documentation culture.
Horizon3.ai is a fast-growing, remote cybersecurity company that helps organizations proactively find and fix exploitable attack vectors. We are a team of former special ops cyber operators and engineers committed to a culture of respect, collaboration, ownership, and results.
Architect, secure, and operate production Kubernetes platforms across AWS and hybrid environments.
Design reusable Infrastructure as Code using Terraform, Ansible, or similar tools.
Lead development of secure CI/CD and GitOps workflows using GitLab, Argo CD, Flux, or equivalent technologies.
Raft is a customer-obsessed non-traditional defense tech company empowering U.S. military and government agencies with cutting-edge AI/ML and data solutions. The company's culture values collaboration, innovation, and diversity, with a team focused on building impactful digital solutions.
Dive deep into Kubernetes and related technologies like Istio, Helm, and Prometheus.
Tackle daily engineering challenges and drive innovation in Kubernetes optimization.
Work with cutting-edge tools and frameworks such as container orchestration, microservices architecture, and cloud-native applications.
We are redefining autonomous cloud and AI infrastructure, freeing DevOps teams from manual resource management to maximize performance and reduce cloud costs by up to 80%. Backed by over $210M from leading VCs and trusted by Fortune 100 companies, we are building a world-class platform with a culture of innovation.
Design and operate scalable cloud infrastructure across AWS and GCP.
Build and improve Kubernetes, Linux, and cloud networking environments.
Strengthen security, disaster recovery, and platform resilience.
Hubstaff provides workforce analytics and time tracking for remote teams, serving over 200,000 global users. The company is a product-led organization with a winning culture and a fully remote team of experienced engineers.
Define architecture and best practices for the platform and infrastructure layer the product is built on.
Own the deploy pipeline and lead the move to a GitOps model (Argo) for fast, safe releases.
Design and harden multi-tenant isolation and blast-radius protection for top-tier customers, including dedicated deployments.
We are the Engineering Operations Platform - mission control for the AI software factory, providing visibility, governance, and golden paths. We are a group of 80 passionate individuals, backed by $60M Series C from Sequoia, IVP, and others, with a fully remote culture.
Build and operate backend services at scale, working with Kubernetes, Terraform, and multi-cloud infrastructure across AWS and Azure.
Participate in code reviews, follow testing standards, and maintain observability coverage including metrics, alerts, and distributed tracing.
Collaborate with senior engineers, seek mentorship, and contribute to sprint planning while growing technical skills.
Twilio is shaping the future of communications, delivering innovative solutions to hundreds of thousands of businesses and empowering millions of developers worldwide. The company is remote-first with a strong culture of connection and global inclusion, employing a diverse team that makes a global impact.
Lead the design and implementation of reliable, scalable, and secure production platforms and services.
Provide technical leadership and mentorship, promoting strong engineering standards across the organization.
Drive standardization, automation, and documentation to improve consistency and reduce operational overhead.
Cision is a global leader in PR, marketing, and social media management technology, helping brands connect with customers and stakeholders. With offices in 24 countries and a network of over 1.1 billion influencers, they foster an inclusive culture of innovation and collaboration.
Design, build, and operate reliable infrastructure supporting AI-powered products.
Own and improve Kubernetes environments and cloud infrastructure.
Enhance production reliability through observability, automation, and incident response.
The company builds advanced AI-driven products and services. It values engineering excellence, autonomy, and individual contribution, with a global team of skilled engineers.