Similar Jobs
See allLLM Inference Engineer
NEAR AI
US
PyTorch
CUDA
AI Infrastructure Engineer
Pragmatike
EMEA
Python
Kubernetes
VLLM
Principal ML Ops Engineer
Jobgether
Switzerland
MLOps
VLLM
Python
Senior / Staff AI Platform Engineer
Clear Street
US
Rust
Postgres
TypeScript
Staff MLOps Engineer
Sequen
US
Python
PyTorch
Docker
Engineering:
- Define the architecture and build systems for model serving, evaluating emerging technologies.
- Spend majority of time designing, building, and optimizing production systems while shaping long-term AI infrastructure strategy.
- Mentor engineers as the team grows and help establish engineering best practices.
Qualifications:
- Significant experience designing and operating production AI inference systems.
- Experience with modern inference runtimes like vLLM, SGLang, TensorRT-LLM, or comparable.
- Strong background in distributed systems, backend infrastructure, or high-performance platform engineering.
Compensation:
- Salary range $190,000 – $230,000 per year.
- Plus health insurance and equity.
Syllo
Syllo is on a mission to transform litigation with a unified platform that enables lawyers to safely harness AI. Since going to market, they have gained diverse enterprise customers including big law firms and corporations, and are quickly expanding.