Design scalable storage architectures for edge and core GPU deployments using platforms like StorPool, NVMe, VAST Data, and Weka.
Optimize storage for AI workloads including distributed training, fine-tuning, and inference with GPU Direct Storage and RDMA.
Act as primary storage design authority, influencing platform architecture and mentoring engineers across infrastructure domains.
Radian Arc builds scalable storage architectures powering AI and GPU infrastructure across edge and core environments. They are a growing company with a remote-friendly work model and an inclusive environment focused on next-generation AI infrastructure.
Deploy and integrate high-performance NFS-based storage into Kubernetes clusters via CSI for GPU-accelerated workloads.
Automate storage provisioning and monitoring using infrastructure-as-code tools like Terraform and GitOps pipelines.
Tune Linux and network settings to optimize throughput and latency for demanding AI and machine learning applications.
Mirantis is a Kubernetes-native AI infrastructure company that helps organizations build scalable, secure infrastructure for AI and data-intensive workloads. With deep expertise in open source and Kubernetes orchestration, they enable platform engineering teams across on-premises, cloud, edge, and sovereign environments.
Build and deliver high-quality solutions that power Honeycomb's query and data storage infrastructure.
Scope and deliver projects independently, breaking down complex storage problems into achievable steps.
Support our services in production, participating in on-call rotations and reducing toil.
Honeycomb is a service defining observability and raising expectations for developer tools, working with companies like HelloFresh and Slack. We've scaled past 200 people, closed Series D funding, and were named to Forbes' America's Best Startups in 2022 and 2023.
Work directly on petabyte-scale storage infrastructure, and the networking and performance challenges that come with it.
Collaborate daily with researchers and engineers who are some of the best in the world at what they do.
Build and maintain the high-performance data layer that Modeling teams rely on for training and evaluation jobs.
Cohere is a security-first enterprise AI company that builds cutting-edge foundation AI models and end-to-end products. They are a global team of researchers, engineers, and designers passionate about their craft, with offices in Toronto, London, New York City, San Francisco, Montreal, Paris, Berlin, and Seoul.
Build large-scale real-time services and applications leveraging massive datasets.
Develop and maintain data pipelines, messaging systems, databases, and cloud services.
Work with Machine Learning Engineers and Security Researchers on security solutions.
Censys provides real-time Internet intelligence and threat insights to global governments and Fortune 500 companies. It is a growing company with a focus on comprehensive internet mapping and security solutions.
Design, build, and operate GitLab Orbit backend services, primarily in Rust, within a distributed, cloud-native environment.
Improve deployment, monitoring, and operations using Kubernetes, Helm, Terraform, and cloud services from AWS or GCP.
Automate operational work, strengthen observability, and manage production issues to reduce single points of failure.
GitLab is the intelligent orchestration platform for DevSecOps, helping organizations increase developer productivity, improve operational efficiency, and accelerate digital transformation. Trusted by more than 50 million registered users and over 50% of the Fortune 100, GitLab fosters a high-performance culture driven by shared values and continuous knowledge exchange.
Design and build automation systems for package creation, test generation, and image building.
Develop AI-powered tooling using LLMs for manifest generation and validation.
Write production Go code and create quality tools to improve customer reliability.
Chainguard delivers hardened, secure, and production-ready builds of open source software. Backed by leading investors, they serve Fortune 500 enterprises and global industry leaders.
Design, implement, and manage scalable cloud infrastructure using Kubernetes and Pub/Sub.
Refactor systems for scalability, including transforming stateful components to stateless ones.
Drive platform reliability initiatives like alerting, health checking, and incident management.
Syllo is a unified litigation platform that empowers lawyers and paralegals to safely use language models and agentic AI throughout the litigation lifecycle. The company has gained enterprise customers including major law firms and corporations, and is quickly expanding with a focus on reducing litigation costs and improving access to justice.
Own the platform including GCP, Kubernetes, Temporal, GPU fleet, and deploy/rollback machinery.
Contribute to AI enablement substrate: GPU capacity, training/inference pipelines, and cost optimization.
Strengthen team practices through tooling, standards, tests, observability, and release processes.
Descript is building a simple, intuitive, fully-powered editing tool for video and audio — an editing tool built for the age of AI. They are a team of 150 backed by top investors like OpenAI and Andreessen Horowitz, with a culture that values collaboration and serendipitous discovery.
Architect and build a robust, scalable, and highly available distributed infrastructure.
Build a cutting-edge cloud-native platform on top of the public cloud and automate cloud resource management.
Work closely with core database development and security teams to produce the SaaS offering.
ClickHouse is a real-time analytics and data warehousing company recognized on the Forbes Cloud 100 list. With over 4,000 customers and rapid growth, the company is a leader in its space.