Own the platform including GCP, Kubernetes, Temporal, GPU fleet, and deploy/rollback machinery.
Contribute to AI enablement substrate: GPU capacity, training/inference pipelines, and cost optimization.
Strengthen team practices through tooling, standards, tests, observability, and release processes.
Descript is building a simple, intuitive, fully-powered editing tool for video and audio — an editing tool built for the age of AI. They are a team of 150 backed by top investors like OpenAI and Andreessen Horowitz, with a culture that values collaboration and serendipitous discovery.
Evolve an Internal Developer Platform enabling development teams to deploy and operate applications securely with high self-service.
Play a crucial role in infrastructure architecture in a multi-cloud environment with a focus on GCP.
Build and maintain the platform used by over 50 internal clients, then support their concrete use by teams.
Lifen believes medical data can transform healthcare by reducing administrative burden, improving care coordination, and accelerating scientific discovery. Since 2015, the company has connected 800 hospitals and 150,000 healthcare professionals, with over 150 employees working remotely and from offices to unlock the potential of health data.
Design, build, and operate infrastructure for real-time systems handling millions of concurrent connections and billions of monthly API requests.
Drive Kubernetes end to end: cluster architecture, workload design, and migration of existing services from AWS to GCP.
Own cloud cost and efficiency optimization, measuring impact against real spend and utilization data.
Stream powers real-time chat, video, activity feeds, and AI moderation for billions of end-users across thousands of apps. We are a Series B company with around 145 employees from over 35 countries, offering a fast-paced startup culture with real ownership.
Lead end-to-end technical engagements: Partner directly with engineering teams to diagnose, unblock, and resolve complex infrastructure challenges.
Execute critical migrations: Develop reference implementations, tooling, and guidance to transition teams off deprecated systems seamlessly.
Accelerate platform adoption: Act as primary technical contact for new teams onboarding to Planet's core infrastructure.
Planet designs, builds, and operates the largest constellation of imaging satellites in history, delivering unprecedented dataset via a cloud-based platform for commercial, environmental, and humanitarian sectors. A global company with offices in the US, Europe, and Slovenia, Planet values a people-centric culture and community.
Lead the architecture and implementation of complex cloud solutions across AWS and GCP.
Drive cloud automation and optimization initiatives to improve scalability and reliability.
Provide technical leadership and mentorship to engineers while collaborating with global teams.
The company focuses on cloud infrastructure and platform engineering. They operate with global teams and emphasize automation, security, and reliability.
Develop and evolve foundational software and services enabling product and development teams.\n- Architect, design, and implement Infrastructure as Code using Terraform.\n- Deploy, manage, and optimize Kubernetes clusters on GCP (GKE) and AWS (EKS).
DoiT is a global technology company that helps cloud-driven organizations leverage cloud for business growth and innovation through data, technology, and human expertise. They work with over 4,000 customers worldwide and foster a remote-first, entrepreneurial culture.
Lead Cloud Platform and SRE teams to scale securely and efficiently.
Drive infrastructure strategy, including Kubernetes (GKE) clusters and developer platform.
Champion SRE culture with SLOs, error budgets, and observability.
Prolific builds human data infrastructure for AI development. The company is a fast-growing, mission-driven organization with cross-functional teams and a strong ownership culture.
Own the reliability, performance, and scalability of Runlayer's infrastructure across AWS and GCP.
Manage Kubernetes clusters, database reliability, and CI/CD pipelines for rapid deployments.
Lead incident response and partner with product engineers to design resilient systems for enterprise customers.
Runlayer builds a unified platform for MCPs, Skills, and AI Agents, providing enterprises with security, governance, and observability to deploy AI safely and at scale. Founded by engineers who built AI Actions for OpenAI and Zapier Agents, the team has raised $42M from Felicis and Khosla Ventures, serving companies like Gusto, Instacart, and Opendoor.
Define DevOps strategy and lead infrastructure architecture across multi-environment, multi-region cloud systems.
Architect and own scalable Kubernetes platforms, infrastructure as code, and DevSecOps implementation.
Drive platform reliability, performance SLAs, cost optimization, and lead complex migrations and AI/ML platform infrastructure.
Robots & Pencils is an applied AI engineering firm that designs and ships AI co-workers for enterprise operations. Founded in 2009, the company has delivery centers across Canada, the US, Eastern Europe, and Latin America, with teams averaging over 15 years of experience.
Design and manage high-availability platforms using Kubernetes, Terraform, and Ansible with native-AI capabilities.
Develop and operate the observability stack: Grafana, Mimir, Loki, Tempo, and Prometheus on Kubernetes via GitLab CI/CD.
Build automation scripts in Python, maintain GitOps pipelines, and mentor mid-level engineers.
Flexential builds and operates critical IT platforms including observability, DevOps, and ITSM technologies. The company fosters a collaborative engineering culture and values diversity.
Own and evolve production infrastructure, leading the migration from Docker Swarm to Kubernetes on premises.
Drive observability, enforce IaC practices, and ensure CI/CD reliability across ~50 services.
Participate in on-call rotation, resolve incidents, and build platform tooling to reduce infrastructure toil.
Webshare is an enterprise-grade proxy platform providing access to over 80 million global IPs across 195 countries. With 99.97% uptime, it serves tens of thousands of businesses, and the team emphasizes mentorship, knowledge-sharing, and team events.
Own the technical roadmap for Platform Engineering, architecting scalable backend services and infrastructure.
Build and operate Kubernetes infrastructure, CI/CD pipelines, and developer tooling to boost engineering productivity.
Design and maintain enterprise integrations, billing platforms, and AI infrastructure including LLM orchestration and observability.
EasyLlama is an AI-powered Human Risk Management platform replacing outdated compliance training with bite-sized learning. Trusted by 6,000+ organizations, it's bootstrapped, profitable, and Inc. 5000-recognized with a fast-paced, ownership-driven culture.
Support management and maintenance of Google Cloud Platform infrastructure.
Assist in ensuring reliability, scalability, and performance of cloud services.
Contribute to GitLab CI/CD pipelines, work with Kubernetes, and write automation scripts.
Miratech is a global IT services and consulting company that helps visionaries change the world through digital transformation. With nearly 1,000 professionals across 25 countries, the company maintains a culture of Relentless Performance with a 99% project success rate and over 30% year-over-year growth.
Design, deploy, and maintain GCP infrastructure, including Compute Engine, GKE, Cloud Storage, IAM, and Cloud Interconnect, following well-architected principles.
Automate provisioning using Terraform, implement IAM best practices, and partner on security audits and hardening.
Monitor performance, availability, and cost, and support cross-cloud and on-premises networking as needed.
Dragos is the global leader in xOT cybersecurity, protecting critical infrastructure systems that deliver water, power, and healthcare. The remote-first team spans North America, Europe, the Middle East, and APAC, built on authenticity, transparency, and trust.
Automate and build runtime production environments, serving as the link between application development and platform teams.
Validate solutions and implementations to ensure alignment with business requirements and maintain platform integrity.
Independently solve business problems and develop small components to address challenges without explicit architecture diagrams.
Defense Unicorns delivers mission value by streamlining software delivery so our customers can focus on the most important challenges. Our team is composed of innovators, software engineers, and veterans with decades of experience delivering technology programs across the federal market.
Architect and build a robust, scalable, and highly available distributed infrastructure.
Build a cutting-edge cloud-native platform on top of the public cloud and automate cloud resource management.
Work closely with core database development and security teams to produce the SaaS offering.
ClickHouse is a real-time analytics and data warehousing company recognized on the Forbes Cloud 100 list. With over 4,000 customers and rapid growth, the company is a leader in its space.
Operate and improve Linux infrastructure and Kubernetes clusters across bare-metal, virtualized, and on-premise environments.
Design and maintain complex networking architectures and automation using Ansible, Bash, Python, and GitOps.
Lead incident response, define SLOs, and build observability platforms with Prometheus, Grafana, and ELK.
Jobgether is a platform that connects job seekers with opportunities through an AI-powered matching process. The company fosters a remote-first culture and emphasizes autonomy and ownership for engineers.
Set the technical direction and roadmap for Platform engineering, owning the reliability, performance, and cost of Traild's GCP infrastructure.
Improve developer experience by enhancing CI/CD pipelines, build tooling, observability, and reducing friction for other engineers.
Lead a small team of platform engineers, staying hands-on while growing and supporting your team members.
Traild is a high-growth SaaS company redefining how finance teams operate by combining AI, automation, and payments infrastructure for B2B finance. They are a rapidly growing global team with a strong culture, as evidenced by an eNPS score of 78.
Define the strategy, roadmap, and feature priorities for k0rdent AI Kubernetes services, empowering Neocloud operators to launch managed Kubernetes on their own GPU infrastructure.
Translate requirements from NeoClouds, GPU clouds, telcos, and enterprise platform teams into clear product direction and partner with engineering to ship secure, scalable cluster lifecycle capabilities.
Manage the Kubernetes backlog, define positioning and competitive differentiation, and create field-facing assets to support strategic accounts.
Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI. They have a world-class, distributed team committed to openness and technical excellence.