Source Job

12 jobs similar to Senior Solutions Architect – AI Infrastructure Networking

Jobs ranked by similarity.

$97,200–$113,400/yr
Global

  • Lead end-to-end technical deployments for GPU neocloud and AI Factory customers.
  • Configure and troubleshoot bare metal GPU infrastructure including CNI, GPU Operator, and storage.
  • Build playbooks and transfer knowledge to customer teams for self-sufficiency.

vCluster Labs provides a platform for AI infrastructure, helping GPU cloud providers and enterprises build Kubernetes-based clusters. They are a venture-backed startup with 40+ engineers and a remote-first culture, powering over 100,000 GPUs.

Europe

  • Own the technical relationship and be the trusted advisor for CTOs and platform teams.
  • Architect real solutions that translate customer constraints into deployable AI infrastructure.
  • Lead proofs of concept under real conditions to demonstrate operational fit.

Mirantis is a Kubernetes-native AI infrastructure company that helps organizations build scalable and secure infrastructure for modern AI workloads. The company has an installed base of 1,500 enterprise customers and values open source innovation, collaboration, and continuous growth.

Europe

  • Act as a trusted technical advisor and architect AI infrastructure solutions using Mirantis k0rdent for enterprise customers.
  • Lead technical discovery, demos, and proof-of-concept engagements to validate and accelerate adoption.
  • Collaborate with internal teams to influence roadmap and create reusable technical assets.

Mirantis is the Kubernetes-native AI infrastructure company, helping organizations build and operate scalable, secure, and sovereign infrastructure for modern AI and data-intensive applications. They are a Silicon Valley leader with a collaborative, high-energy culture, serving Fortune 500 and Global 2000 customers.

China

  • Operate and expand Telnyx's own B300 GPU fleet to maximize inference throughput per GPU-dollar.
  • Design and implement serverless inference for open-weight models and dedicated enterprise deployments.
  • Work upstream in open-source technologies like vLLM, SGLang, and Kubernetes.

Telnyx is an industry leader building the future of global connectivity through a private, multi-cloud IP network and edge APIs. The company is financially stable and profitable, with a global team and a focus on innovation and continuous learning.

$184,000–$318,000/yr
US

  • Architect GPU cluster topologies spanning compute, networking, storage, and control planes.
  • Model AI workloads like LLM training and inference to guide performance tradeoffs.
  • Collaborate with cross-functional teams to deploy and scale reliable AI infrastructure.

This role is with a partner company focused on designing next-generation AI infrastructure at massive scale. The company fosters a fast-moving, engineering-led culture with a collaborative, international team.

Canada

  • Lead the architecture and delivery of a large-scale GPU infrastructure platform, evolving from managed Kubernetes to bare-metal with Slurm and inference support.
  • Manage a distributed engineering team across backend, frontend, DevOps, QA, and documentation, setting technical standards and overseeing implementation.
  • Own GPU infrastructure operations, including Slurm, Kubernetes, NVIDIA hardware, observability, and incident response, while acting as the primary technical interface with partners.

Jobgether is an AI-powered job matching platform that connects candidates with relevant roles, ensuring a fair and objective review process. It operates with a distributed team and partners with companies globally, focusing on efficient and transparent recruitment.

Europe

  • Operate and improve Linux infrastructure and Kubernetes clusters across bare-metal, virtualized, and on-premise environments.
  • Design and maintain complex networking architectures and automation using Ansible, Bash, Python, and GitOps.
  • Lead incident response, define SLOs, and build observability platforms with Prometheus, Grafana, and ELK.

Jobgether is a platform that connects job seekers with opportunities through an AI-powered matching process. The company fosters a remote-first culture and emphasizes autonomy and ownership for engineers.

$135,000–$200,000/yr
US

  • Own end-to-end execution of high-stakes customer projects, working independently and with minimal supervision.
  • Deploy and support containerized software on edge devices and in air-gapped, on-premise environments.
  • Troubleshoot networking and systems issues, manage configuration, load balancing, container orchestration, and upgrades.

The company delivers software solutions for tactical edge environments. It operates as a small, customer-facing engineering team focused on high-stakes deployments.

US

  • Architect AIStor deployments for federal customers, spanning on-prem, air-gapped, sovereign cloud, and tactical edge environments.
  • Serve as technical authority on FIPS 140-3, FedRAMP, NIST 800-53, and Zero Trust, translating compliance into deployable architectures.
  • Lead PoVs and technical demonstrations for DoW, IC, and civilian agencies, working with partners and ecosystem teams.

MinIO is the data and memory foundation for enterprise AI, delivering AIStor and MemKV to unify data across core, edge, and cloud. Trusted by 77% of the Fortune 100, the company is redefining how AI factories and intelligent applications secure and unlock data value.

Europe

  • Build and scale massive distributed compute and storage systems for frontier model training.
  • Architect multi-cluster orchestration layers to optimize workload placement across diverse hardware and regions.
  • Design future-proof storage and metadata systems to handle exabyte-scale growth.

Mistral provides full-stack AI solutions from frontier models to developer tools, applications, and compute. We are a dynamic, collaborative team passionate about AI, with a diverse workforce distributed across Europe, North America, Asia, and the Middle East.

US

  • Own the technical architecture and roadmap for a multi-RAT RAN digital twin covering LTE and 5G NR.
  • Integrate production MAC and scheduler software into deterministic closed-loop simulations with stable interfaces.
  • Develop AI/ML-based RAN capabilities including neural channel estimation and learned scheduling policies.

Parallel Wireless is a U.S.-based pioneer in Open RAN innovation, transforming how mobile networks are built, optimized, and powered. The company culture emphasizes software-centric, hardware-agnostic approaches to reduce complexity and total cost of ownership.

EU

  • Own and operate production infrastructure across Kubernetes, Linux, networking, and virtualization.
  • Lead incident response and implement observability to improve availability and performance.
  • Define SLOs and automate infrastructure with Ansible, Bash, Python, and GitOps.

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies through objective, data-driven processes. They foster a collaborative, international, and fully remote work environment, emphasizing autonomy and ownership for their small to mid-sized team.