Staff Engineer, Distributed Storage and HPC & AI Infrastructure at Together AI
- Company: Together AI
- Location: San Francisco
- Employment type: Full-time
- Posted: 2026-06-04
- Technologies: Kubernetes
About the role
About the Role . In this role, you will operate, scale, and optimize multi-petabyte storage systems purpose-built for the world’s largest AI training and inference workloads. You’ll manage and scale high-performance parallel filesystems and object stores, evaluate and integrate cutting-edge technologies such as Vast, Weka, Ceph, and Lustre, and solve the complex engineering challenges of operating at extreme throughput, low-latency data paths, and massive cluster-scale storage operations. You will also build Kubernetes-native storage operators and self-service platforms that provide automated provisioning, strict multi-tenancy, performance isolation, and quota enforcement at cluster scale.…
Apply on Together AI's official careers page: https://job-boards.greenhouse.io/togetherai/jobs/5155722007