Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Engineer, GPU Infrastructure (HPC)”. A match may be a passing mention rather than the job itself. Titles only.
338 roles across 389 listings · show every listing · page 2 of 14
…GPU or HPC cluster networking: GPUDirect, RoCE, or Spectrum-X fabrics. Experience taking a project into open source, including defining community and contribution strategy…
…Develop thermal validation tools, processes, automation, and infrastructure for a global team of system engineers Create software and simulation tools to assess thermal feasibility…
…native, AI, and GPU workloads. We are looking for a senior IC5 software engineer with deep Kubernetes expertise, required cloud infrastructure experience, and a…
…About the Role We are seeking a Senior Software Engineer to join our Managed Kubernetes (Mk8s) team. You will play a crucial role in…
…Bachelor’s or Master’s degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent experience. 12+ years of software engineering experience, including the…
…NVIDIA is looking for a Senior AI/HPC Engineer to join its infrastructure Specialist team. Academic and commercial groups around the world are using…
…GPUs, while optimizing software to improve system robustness, performance, and security Participating in testing new and existing firmware, and developing tools and infrastructure to…
…Computer Architecture (CPUs, GPUs, FPGAs or other accelerators), GPU Programming Models, Performance-Oriented Parallel Programming, Optimizing for High-Performance Computing (HPC), Algorithms, Numerical Methods…
…QEMU) - 10+ years of experience in systems software engineering working on low-level systems or infrastructure teams. Benefits: - Competitive compensation and equity packages - Restricted…
…As an AI/HPC System Performance Engineer on the Network Infrastructure Engineering team, you will drive end-to-end performance characterization, bottleneck analysis, and…
…The MTIA Software team is part of the **AI & Compute Foundation (ACF)** organization within Meta Infrastructure. Because the hardware is ours, the software is…
…AI/ML systems, distributed/HPC systems, software tooling and infrastructure or model optimization with GPUs/accelerators Demonstrated ability to engage deep in technical work…
…infrastructure. What you'll be doing: : Hybrid Quantum–HPC Platform Engineering Build, deploy, and operate a hybrid computing platform combining large-scale NVIDIA GPU…
…The Lambda Infrastructure Engineering organization forges the foundation of high-performance AI clusters by welding together the latest in AI storage, networking, GPU and…
…Primary responsibilities will include building AI/HPC infrastructure for new and existing customers. Support operational and reliability aspects of large-scale AI clusters, focusing…
…Strong understanding of AI/HPC platforms, GPU servers, high-density compute environments, and data center cooling infrastructure, including facility-level cooling integration. Reliability, Failure…
…performance computing (HPC), AI infrastructure, or GPU-based systems; knowledge of coolant chemistry, reliability, and contamination control Familiarity with data center infrastructure, including chilled…
Senior Network Development engineer to Support Infiniband, Roce and Roma network in OCI. We need people to support the over expansion of the GPUs…
…What you bring Have done capacity, demand, or supply planning for large-scale technical infrastructure (cloud, HPC, hyperscale, or a large internal platform) and…
…Industry-leading acceleration technologies such as TPUDirect, TPUDirect Storage (TDS), and GPUDirect Storage (GDS). As a Senior Staff Software Engineer, you will lead the…
Engineering organization within Oracle Cloud Infrastructure (OCI) sits at the center of the fastest-growing and most technically demanding part of the cloud: building…
…Lead operational and software engineering efforts that improve the reliability, availability, observability, and performance of OCI AI/HPC networking fabrics. Apply deep networking knowledge…
…and tooling. - Partner with Storage Engineers, Fleet Orchestration, and Release Engineering to automate the deployment and configuration of software-defined storage across new and…
…At NVIDIA, our Solution Architects are drawn from elite developers and scientists who enjoy working with the latest GPU hardware and software. We need…
…AI and HPC GPU infrastructure. Are you keen to join a team that brings GenAI, AI, and ML hardware and software technologies into real…