AI/ML Infra / Systems Engineer (Inference)
Sapience
- Salary
- $204k - $216k
- Experience
- 5+ yrs
- Location
- Seattle or Washington
- Job type
- Full-time
Required skills
- vLLM
- TensorRT
- Triton
- CUDA
- Kubernetes
- Docker
- AWS
- GCP
- Azure
- Python
- Go
- Rust
- C++
About the role
5+ years in infrastructure, systems, or ML infrastructure engineering; hands-on production ML/LLM serving; deep understanding of latency, throughput, and performance optimization; GPU/accelerator experience; strong systems programming and distributed systems fundamentals; observability and operational excellence.
About Sapience
Private AI software company helping membership organizations and professional communities make institutional knowledge searchable and actionable.
This page is fully interactive when JavaScript is enabled. Please enable JavaScript to apply or browse related roles.