Lead the architecture and delivery of a large-scale GPU infrastructure platform, evolving from managed Kubernetes to bare-metal with Slurm and inference support.
Manage a distributed engineering team across backend, frontend, DevOps, QA, and documentation, setting technical standards and overseeing implementation.
Own GPU infrastructure operations, including Slurm, Kubernetes, NVIDIA hardware, observability, and incident response, while acting as the primary technical interface with partners.
Jobgether is an AI-powered job matching platform that connects candidates with relevant roles, ensuring a fair and objective review process. It operates with a distributed team and partners with companies globally, focusing on efficient and transparent recruitment.
Define and maintain end-to-end system architecture for secure AI evaluation platforms.
Evaluate open-source and commercial technologies to guide build-versus-buy decisions.
Guide technical solutions and mentor engineering teams across distributed environments.
OpenTeams helps enterprises and governments build AI they can fully control and govern. Built by pioneers from NumPy, SciPy, and PyTorch, the company fosters a culture of ownership and innovation with a global, distributed team.
Evaluate GPU kernel tasks for technical accuracy, realism, solvability, reproducibility, and robust testing criteria.
Rigorously test and troubleshoot complex GPU programming scenarios to identify memory allocation bugs, execution bottlenecks, and parallel computing logic errors.
Review CUDA, Triton, and other GPU kernel implementations and provide clear, actionable technical feedback.
Our partner is a company specializing in AI training and evaluation, seeking experienced GPU kernel specialists to audit AI training tasks. The project is globally distributed and offers fully remote freelance work.
Guide architectural decisions across products, reviewing solutions from internal teams and vendors.
Establish shared engineering standards and modernise legacy systems for reliability and cost-effectiveness.
Define guardrails for production LLM systems and collaborate with teams on Kubernetes and cloud infrastructure.
INFUSE is a technology company that develops software products, including AI and LLM-based solutions, across a large engineering landscape. Its software department spans 500+ repositories and fosters a remote-first culture with access to cutting-edge AI tools and human-centric values.
Operate and expand Telnyx's own B300 GPU fleet to maximize inference throughput per GPU-dollar.
Design and implement serverless inference for open-weight models and dedicated enterprise deployments.
Work upstream in open-source technologies like vLLM, SGLang, and Kubernetes.
Telnyx is an industry leader building the future of global connectivity through a private, multi-cloud IP network and edge APIs. The company is financially stable and profitable, with a global team and a focus on innovation and continuous learning.
Build and operate the Kubernetes platform supporting AI test and evaluation frameworks.
Design infrastructure-as-code, GitOps workflows, and automated deployment pipelines.
Own platform reliability, observability, capacity planning, and operational readiness.
OpenTeams helps enterprises and governments build AI they control, govern, and evolve themselves. Founded by the creator of NumPy and SciPy, the company is built by people with deep roots across the open-source ecosystem and maintains a remote-first culture.
Lead and mentor the software architecture team, setting technical direction and standards.
Architect scalable, secure cloud and edge solutions, focusing on AI-enabled capabilities.
Drive platform modernization and evaluate build, buy, and partner opportunities.
Daktronics designs, engineers, manufactures, and supports digital LED display technology and audio systems for sports, businesses, and transportation. With a global presence, the company offers competitive compensation, meaningful benefits, and a collaborative culture focused on innovation and impact.
Lead cross-domain architectural strategy and long-term technical vision for Cisco's policy and security platforms.
Drive industry-leading innovation in cloud and client infrastructure, influencing senior leadership and external standards.
Mentor principal engineers and represent Cisco externally, translating emerging trends into scalable business outcomes.
Cisco transforms how data and infrastructure connect and protect organizations in the AI era, delivering security, visibility, and insights across digital footprints. With a global network of over 75,000 employees, Cisco fosters a collaborative, empathetic culture that encourages experimentation and meaningful impact at scale.
Build the foundations of the EdgeRunner Research organization, including data pipelines, model evaluation, and efficient parallelization.
Own codebases in areas like distributed training, quantization, compression, or compute cluster management.
Collaborate with a team of self-starters who operate independently and translate business objectives into technical solutions.
EdgeRunner AI builds state-of-the-art AI models for the tactical edge, enabling warfighters to make faster decisions and interact with robotic platforms. As a Series-A startup, we move quickly, fostering a culture of initiative, ownership, and comfort with ambiguity.
Develop deep understanding of customer problems and map them to Cohere solutions as a trusted technical advisor.
Lead design and deployment of customer pilots, owning the technical narrative from discovery to business case.
Collaborate with Product and Engineering teams and partners to scale solutions and gather customer feedback.
Cohere is a leading security-first enterprise AI company building cutting-edge foundation AI models and products for real-world business problems. It is a global team of researchers, engineers, and designers headquartered in Toronto with offices worldwide, committed to driving AI adoption.
Act as a subject matter expert for compute, supporting sales and customer success teams across your region.
Develop and nurture customer relationships within your assigned territory, including attending industry events and conferences.
Define and execute the go-to-market compute sales strategy, collaborating with Solution Architects and Customer Success Managers to drive growth.
Megaport is the global leader in Network as a Service (NaaS), transforming how businesses connect to cloud, data centers, and each other. Headquartered in Brisbane with over 600 employees across Asia-Pacific, Europe, and the Americas, they foster a collaborative, supportive, and fun culture that values curiosity and teamwork.
Build and scale massive distributed compute and storage systems for frontier model training.
Architect multi-cluster orchestration layers to optimize workload placement across diverse hardware and regions.
Design future-proof storage and metadata systems to handle exabyte-scale growth.
Mistral provides full-stack AI solutions from frontier models to developer tools, applications, and compute. We are a dynamic, collaborative team passionate about AI, with a diverse workforce distributed across Europe, North America, Asia, and the Middle East.
Lead in-depth technical discovery with engineering teams and customer stakeholders to understand AI inference requirements.
Translate customer objectives into production-ready architectures and define PoC success criteria.
Identify recurring workload patterns and communicate insights to Product and Engineering for platform evolution.
They are a partner company focused on AI infrastructure and performance-sensitive AI inference workloads. They have an international, engineering-led team solving complex challenges at the forefront of AI.
Analyze and optimize CPU cluster performance including cache hierarchies and interconnects.
Build and use performance models and simulation environments to evaluate architectural concepts.
Lead architectural tradeoff studies across performance, scalability, and power to influence design decisions.
Tenstorrent is leading the industry on cutting-edge AI technology. Their diverse team of technologists is passionate about AI and building the best AI platform, valuing collaboration and curiosity.
Build and improve the inference layer of the Gcore Inference platform, integrating frameworks like vLLM and TensorRT-LLM.
Bring new language and multimodal models into production, optimizing latency, throughput, and cost efficiency.
Debug performance issues across model code, GPU execution, and Kubernetes, collaborating with cross-functional teams.
Gcore is a global provider of AI, cloud, network, and security infrastructure and software. They are a team of 550+ professionals with a collaborative culture and partnerships with Intel, NVIDIA, Dell, and Equinix.
Work directly on petabyte-scale storage infrastructure, and the networking and performance challenges that come with it.
Collaborate daily with researchers and engineers who are some of the best in the world at what they do.
Build and maintain the high-performance data layer that Modeling teams rely on for training and evaluation jobs.
Cohere is a security-first enterprise AI company that builds cutting-edge foundation AI models and end-to-end products. They are a global team of researchers, engineers, and designers passionate about their craft, with offices in Toronto, London, New York City, San Francisco, Montreal, Paris, Berlin, and Seoul.
Lead the effort to make Runpod the fastest and most cost-efficient place for LLM inference, owning performance end to end.
Profile and diagnose performance bottlenecks across the serving stack, from scheduling to kernels, and implement fixes.
Work closely with product and infrastructure teams to shape how inference is offered, turning improvements into production-ready defaults.
Runpod is the AI Developer Cloud, providing a platform for over one million developers to experiment, train, fine-tune, deploy, and scale AI. We're a small, remote-first team that takes ownership seriously, moves fast, and has processed more than 20 billion inference requests.
Oversee engineering teams and ensure successful project delivery while maintaining client satisfaction.
Perform pre-sales engineering tasks, design complex multi-vendor solutions, and provide technical oversight.
Lead and mentor team members, fostering growth and collaboration, and contribute to strategic initiatives.
ePlus is a technology solutions provider that believes technology is a people business, delivering innovative IT solutions to clients. Their team of skilled professionals fosters a culture of collaboration, innovation, and respect, with a focus on work-life balance and community support.
Bring newly racked CPU and GPU servers from hardware handoff through full provisioning and configuration to production readiness.
Configure out-of-band management, perform PXE booting and OS provisioning, and troubleshoot common imaging failures.
Maintain inventory accuracy, coordinate with upstream teams to resolve blockers, and follow established runbooks for process improvement.
Vultr makes high-performance cloud infrastructure easy to use, affordable, and locally accessible for enterprises and AI innovators worldwide. As the world's largest privately-held cloud infrastructure company with a $3.5 billion valuation, we serve hundreds of thousands of customers across 185 countries.