Staff Engineer, Distributed Storage and HPC & AI Infrastructure

Together AI · San Francisco

About the Role In this role, you will operate, scale, and optimize multi-petabyte storage systems purpose-built for the world’s largest AI training and inference workloads. You’ll manage and scale high-performance parallel filesystems and object stores, evaluate and integrate cutting-edge technologies such as Vast, Weka, Ceph, and Lustre, and solve the complex engineering challenges of operating at extreme throughput, low-latency data paths, and massive cluster-scale storage operations. You will also build Kubernetes-native storage operators and self-service platforms that provide automated pr

Sign in to apply — one profile, every role on PreferHired.

Sign in to apply
Staff Engineer, Distributed Storage and HPC & AI Infrastructure at Together AI — PreferHired