Lead the architecture and delivery of a large-scale GPU infrastructure platform, evolving from managed Kubernetes to bare-metal with Slurm and inference support.
Manage a distributed engineering team across backend, frontend, DevOps, QA, and documentation, setting technical standards and overseeing implementation.
Own GPU infrastructure operations, including Slurm, Kubernetes, NVIDIA hardware, observability, and incident response, while acting as the primary technical interface with partners.
Jobgether is an AI-powered job matching platform that connects candidates with relevant roles, ensuring a fair and objective review process. It operates with a distributed team and partners with companies globally, focusing on efficient and transparent recruitment.
Own network solutioning for k0rdent AI, designing GPU cluster interconnects and multi-tenant isolation.
Publish reference architectures and run proofs of concept on real hardware with customers.
Research emerging AI networking technologies and help shape the product roadmap.
Mirantis is a Kubernetes-native AI infrastructure company helping organizations build and operate scalable, secure infrastructure for modern AI workloads. It is committed to open standards, technical excellence, and a distributed, collaborative team culture.
Operate and expand Telnyx's own B300 GPU fleet to maximize inference throughput per GPU-dollar.
Design and implement serverless inference for open-weight models and dedicated enterprise deployments.
Work upstream in open-source technologies like vLLM, SGLang, and Kubernetes.
Telnyx is an industry leader building the future of global connectivity through a private, multi-cloud IP network and edge APIs. The company is financially stable and profitable, with a global team and a focus on innovation and continuous learning.
Act as senior technical resource and final escalation point for strategic and VIP customers, owning complex issues across Kubernetes, GPU, and enterprise stack.
Train and mentor Technical Support Engineers in advanced Linux troubleshooting and customer architectures.
Author advanced troubleshooting documentation and drive incident resolution through root cause analysis.
Vultr makes high-performance cloud infrastructure easy to use and affordable for enterprises and AI innovators worldwide. It is the world's largest privately-held cloud infrastructure company with 33 data centers and hundreds of thousands of active customers, committed to growth and employee investment.
Own the technical relationship and be the trusted advisor for CTOs and platform teams.
Architect real solutions that translate customer constraints into deployable AI infrastructure.
Lead proofs of concept under real conditions to demonstrate operational fit.
Mirantis is a Kubernetes-native AI infrastructure company that helps organizations build scalable and secure infrastructure for modern AI workloads. The company has an installed base of 1,500 enterprise customers and values open source innovation, collaboration, and continuous growth.
Design, build, and ship production services, APIs, and user-facing interfaces.
Build and operate production AI systems including RAG, fine-tuning, and inference optimization.
Architect AWS/GCP environments with Kubernetes and Terraform and control cloud/AI costs.
Motive empowers people who run physical operations with tools to make their work safer, more productive, and more profitable. Serving nearly 100,000 customers across industries, the company values a diverse and inclusive workplace.
Own end-to-end execution of high-stakes customer projects, working independently and with minimal supervision.
Deploy and support containerized software on edge devices and in air-gapped, on-premise environments.
Troubleshoot networking and systems issues, manage configuration, load balancing, container orchestration, and upgrades.
The company delivers software solutions for tactical edge environments. It operates as a small, customer-facing engineering team focused on high-stakes deployments.
Act as a trusted technical advisor and architect AI infrastructure solutions using Mirantis k0rdent for enterprise customers.
Lead technical discovery, demos, and proof-of-concept engagements to validate and accelerate adoption.
Collaborate with internal teams to influence roadmap and create reusable technical assets.
Mirantis is the Kubernetes-native AI infrastructure company, helping organizations build and operate scalable, secure, and sovereign infrastructure for modern AI and data-intensive applications. They are a Silicon Valley leader with a collaborative, high-energy culture, serving Fortune 500 and Global 2000 customers.
Own and operate production infrastructure across Kubernetes, Linux, networking, and virtualization.
Lead incident response and implement observability to improve availability and performance.
Define SLOs and automate infrastructure with Ansible, Bash, Python, and GitOps.
Jobgether is an AI-powered job matching platform that connects candidates with hiring companies through objective, data-driven processes. They foster a collaborative, international, and fully remote work environment, emphasizing autonomy and ownership for their small to mid-sized team.
Build and improve the inference layer of the Gcore Inference platform, integrating frameworks like vLLM and TensorRT-LLM.
Bring new language and multimodal models into production, optimizing latency, throughput, and cost efficiency.
Debug performance issues across model code, GPU execution, and Kubernetes, collaborating with cross-functional teams.
Gcore is a global provider of AI, cloud, network, and security infrastructure and software. They are a team of 550+ professionals with a collaborative culture and partnerships with Intel, NVIDIA, Dell, and Equinix.
Own ScaleOps' infrastructure end-to-end, including self-hosted product, SaaS platform, and AI infrastructure.
Manage cloud infrastructure across AWS, GCP, and Azure, covering networking, security, SSO, and compute.
Collaborate with customers and internal teams to ensure reliable feature delivery and eliminate operational toil.
ScaleOps is redefining autonomous cloud and AI infrastructure, freeing DevOps and platform engineers from manual resource management. We are the category leader backed by over $210M in funding, trusted by leading enterprises including Adobe, Coinbase, and Fortune 100 companies.
Manage and optimize Kubernetes clusters in GKE through Terraform.
Design and implement golden paths and automation strategies that empower developers to self-serve.
Serve as the technical point-of-contact for GCP and Kubernetes queries, supporting compliance and internal teams.
Prolific is the architect of human data infrastructure, connecting researchers and companies with a global pool of participants to collect high-quality, ethically sourced behavioral data for AI development. They are a mission-driven, remote-first company at the forefront of AI innovation.
Design and scale cloud-native infrastructure for a fast-growing bandwidth marketplace.
Lead platform reliability with deep expertise in Kubernetes, GCP, and container orchestration.
Drive infrastructure decisions across the full stack, from virtualization to production APIs.
Share is a venture-backed company building the marketplace for bandwidth, connecting internet capacity suppliers to businesses and individuals. The company is investor-backed with a high-ownership environment and a steep learning curve.
Operate and improve Linux infrastructure and Kubernetes clusters across bare-metal, virtualized, and on-premise environments.
Design and maintain complex networking architectures and automation using Ansible, Bash, Python, and GitOps.
Lead incident response, define SLOs, and build observability platforms with Prometheus, Grafana, and ELK.
Jobgether is a platform that connects job seekers with opportunities through an AI-powered matching process. The company fosters a remote-first culture and emphasizes autonomy and ownership for engineers.