Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Staff Engineer, Distributed Storage and HPC & AI Infrastructure”. A match may be a passing mention rather than the job itself. Titles only.
5 roles
…and scalability in AI/ML and HPC workloads. AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In…
…and track storage SLOs/SLIs, and ensure reliable deployment and maintenance of distributed storage infrastructure. - Innovate: Stay current with AI and HPC storage research…
…running distributed AI/ML/HPC workloads across thousands of GPUs, leveraging technologies like RoCE and Infiniband. As a Consulting Member of Technical Staff, you…
…infrastructure for AI/ML or HPC workloads. - Hands-on experience with distributed training and/or inference workloads at scale, including parallelism strategies and performance…
…and implement high-speed networking infrastructure (100G/200G/400G Ethernet) connecting compute nodes, storage, and upstream peering, in coordination with network and platform engineering…