Own observability end to end and define how we measure reliability.
Own CI/CD pipelines and make shipping fast and safe.
Footprint builds Percy, an AI agent that runs financial crime investigations end to end. The company is backed by QED, Index, and other investors, and its small, senior team ships fast and grew revenue 5x in the past year.
Build enterprise-scale infrastructure using infrastructure-as-code and Kubernetes-native systems.
Sustain platform health and performance by owning critical systems in production.
Enable teams and customers to move faster with abstractions and tooling for AI/ML workloads.
Cake makes cutting-edge AI accessible to enterprise teams by removing infrastructure barriers, enabling 10x faster and cheaper AI/ML platform deployment. Backed by top investors, they have a small senior team focused on ownership and operational excellence.
Build and maintain cloud infrastructure across GCP, Kubernetes, and Terraform.
Own CI/CD pipelines and deploy fully automated, locked-down systems.
Strengthen security, access control, and observability for a growing platform.
Gauntlet builds the financial systems of the future, operating across the entire stack to offer best-in-class vault products. The team serves over $1.5B in client TVL and brings together traditional finance and crypto-native expertise.
Design, build, and scale reliable infrastructure for Klover's fintech platform using modern technologies like Kubernetes, Terraform, and Istio.
Use AI agents as force multipliers to automate manual processes and improve developer experience.
Collaborate with engineering teams to ensure system reliability, performance, and security across production systems.
Attain powers Klover, a fast-growing fintech platform serving over one million active users monthly, processing over $1.5 billion annually. The company emphasizes collaboration, reliability, and innovation, with a culture of automation and AI-driven development.
Own cloud infrastructure across AWS and GCP, including Kubernetes, networking, databases, and CI/CD pipelines.
Scale single-tenant deployments and build observability, incident response, and compliance practices.
Manage infrastructure cost, improve developer experience, and contribute to backend systems at the infrastructure-application intersection.
Elicit is an AI research assistant that uses language models to help researchers with literature review and evidence synthesis. The company is a ~30-person Public Benefit Corporation with a high-agency, low-bureaucracy culture.
Design, build, and maintain platform infrastructure using IaC principles with tools like Terraform.
Develop memory services, vector storage patterns, and semantic search capabilities.
Operate a Kubernetes-based platform for the Data & AI department, enabling deployment and scaling via GitOps.
Redcare Pharmacy is Europe's No.1 e-pharmacy, striving to improve global health through innovation and collaboration. The company fosters a healthy, inclusive work environment where employees feel valued and inspired.
Design, operate, and improve reliable infrastructure for AI training and inference workloads.
Build monitoring, alerting, runbooks, and incident-response practices for easier operations.
Partner with ML, research, and platform teams to translate workload needs into infrastructure improvements.
Boson AI builds production-grade AI systems that make communication with AI more natural, capable, and useful. The team is focused on infrastructure reliability, operating GPU clusters and networks for AI workloads.
Design and operate the infrastructure for a high-throughput messaging platform operating at 500K+ events/sec.
Build guardrails, runbooks, and validation gates that enable AI agents to safely execute deployments and operations.
Lead incident response and encode every fix as a new runbook and regression test.
Postscript is an AI messaging platform trusted by 20,000+ Shopify brands to drive revenue through SMS. The company is fully remote, backed by Greylock and Y Combinator, and has a culture of ownership and innovation.
Define and implement SLIs/SLOs for critical services, lead incident response, and conduct blameless postmortems to drive systemic improvements.
Design and improve monitoring and alerting with Prometheus and Grafana, build internal tooling, and automate operational workflows to reduce toil.
Partner with cross-functional engineering teams to improve system resilience, contribute to architectural discussions, and strengthen production readiness standards.
Runpod provides a cloud platform for AI development, used by over one million developers for training, fine-tuning, and deploying AI models. The company is a small, remote-first team that closed a $100M Series A in June 2026, emphasizing ownership, speed, and impact at scale.
Lead design and evolution of secure cloud infrastructure and deployment systems for critical decentralized applications.
Drive improvements across CI/CD pipelines, deployment workflows, and engineering productivity practices.
Collaborate with developers, security specialists, product leaders, and infrastructure teams in a remote-first environment.
Our partner is building and scaling secure, high-performance infrastructure powering one of the most widely used decentralized technology platforms in the world. They operate as a fully remote, globally distributed team with a focus on DevOps, security, and blockchain technology.
Build and operate production-grade model serving infrastructure using vLLM, TGI, or Triton frameworks.
Design and implement auto-scaling, multi-model architectures, and intelligent request routing for ML inference.
Optimize GPU utilization, memory efficiency, and observability to ensure low-latency, cost-effective systems.
They are a distributed cloud infrastructure startup building AI-native cloud services with GPU-powered compute. The company is well-funded, fast-scaling, and operates in a remote-first environment with a focus on sustainability and decentralization.
Lead a high-leverage remote team of four infrastructure engineers, driving the evolution toward a scalable zero-toil platform.
Guide the team through an AI-driven engineering approach to reduce manual work and achieve zero-touch, scalable infrastructure.
Prepare and execute the strategy for CI/CD and artifact distribution systems to scale during a quality surge without increasing engineering toil.
Camunda is the enterprise platform for agentic orchestration, enabling organizations to coordinate AI agents, people, and systems across complex business processes. Trusted by over 700 organizations worldwide, including 9 of top 10 US banks, Camunda is a fully remote and global company with 150+ engineers across 20+ teams, and is transforming into an AI-first organization.
End-to-end ownership of internal orchestration platform built on event-driven architecture with Redpanda, including code, architecture, and roadmap.
Own infrastructure-as-code using Terraform Cloud, manage Kubernetes workloads with Helm, and provide self-service tooling for engineering teams.
Set SLOs, handle production on-call, lead incident response, author design docs, and operate AI-natively using tools like Cursor and Notion AI.
Velora unifies Aplos, Raisely, and Keela into one company with a shared mission to help nonprofit organizations thrive by offering fundraising, donor management, financial tracking, and communications tools. We are a financially solid company with a combined team dedicated to making nonprofit work easier, more impactful, and more sustainable.
Design, build, and operate reliable infrastructure supporting AI-powered products.
Own and improve Kubernetes environments and cloud infrastructure.
Enhance production reliability through observability, automation, and incident response.
The company builds advanced AI-driven products and services. It values engineering excellence, autonomy, and individual contribution, with a global team of skilled engineers.
Architect and scale high-availability backend systems across AWS.
Build CI/CD, developer environments, and AI workloads infrastructure.
Develop advanced autoscaling, caching, and storage strategies.
Source.dev builds software tools to simplify device software development. The company has a high-trust, passionate engineering culture with a focus on end-to-end ownership.
Own core compute infrastructure across multiple cloud providers and regions.
Design capabilities for greater performance and flexibility in service deployment.
Investigate and resolve challenging cloud and compute issues across the stack.
Render is a cloud platform for developers building AI-native, full-stack, multi-service applications. Trusted by over 6 million developers, the company has raised $257M in funding and values craft, velocity, and user experience.
Work with a team of DevOps and DBA professionals to improve infrastructure and streamline deployments across countries.
Continuously improve Kubernetes platform stability, efficiency, and GitOps-first environment provisioning.
Monitor cloud infrastructure, own on-call operations, and define SLIs/SLOs for reliability improvements.
Sporty Group is a remote-first company focused on sustainability in the sports and gaming industry. They maintain a competitive, performance-driven culture with a distributed team across EMEA.
Work on core Radar products using TypeScript, Rust, Python, and Scala.
Build highly available systems with 99.99%+ monthly uptime.
Own the internal development platform and integrate safe AI-enabled development practices.
Radar is the global leader in geolocation, providing geofencing SDKs, maps APIs, and AI-enabled solutions for marketing, fraud, and operations teams. They have raised $85.5M from top investors, process over 1 billion API calls per day, and have a high-performance culture with ambitious teammates.
Harden, simplify, and operationalize the production platform on Google Cloud Platform for enterprise customers.
Own and evolve core infrastructure, CI/CD pipelines, and infrastructure as code practices using Terraform or Pulumi.
Drive observability, developer productivity, and engineering culture to raise the bar across the team.
GC AI is the fastest-growing legal AI platform for in-house legal teams, building the future of legal work. With over 1,700 companies using the platform, including 150+ public companies and 25+ unicorns, the team has 10x'd revenue in 12 months and raised a $60 million Series B.
Design and build infrastructure primitives that define how CI/CD, build systems, and developer environments scale across the engineering org.
Build and operate the Kubernetes-based control plane behind CI/CD, including GitHub Actions runners, GitOps workflows, and ephemeral environments.
Develop core infrastructure components like Kubernetes Operators and scaling automation that product teams use directly, reducing bespoke per-team tooling.
Chainlink is the industry-standard oracle platform that brings capital markets onchain and powers the majority of decentralized finance. The company has enabled tens of trillions in transaction value and is adopted by major financial institutions and top protocols.