Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Staff Software Engineer, GPU Inference”. A match may be a passing mention rather than the job itself. Titles only.
39 roles across 43 listings · show every listing · page 1 of 2
…Provide Staff-Level Technical Leadership: Own the long-term technical roadmap for foundation model development. Mentor senior engineers, lead rigorous design reviews, and establish…
…10+ years of relevant professional software engineering experience, with at least 3+ years in Staff, or Lead Architect role. BS, MS, or PhD in…
…12+ years of relevant professional software engineering experience, with at least 3+ years in Staff, or Lead Architect role. BS, MS, or PhD in…
…15 + years of relevant professional software engineering experience, with at least 3+ years in Staff, or Lead Architect role. BS, MS, or PhD in…
…Engineering Group, Engineering Group > PPT Systems Engineering General Summary: Qualcomm's Chip Architecture multi-site team of systems engineers, hardware and software architects are…
…ROLE OVERVIEW As a Staff Software Engineer on the Model Infrastructure team, you'll lead the design and development of the systems that power…
…launching software products. Experience with machine learning infrastructure, C++, performance, GPU programming, mobile GPU. Preferred qualifications: Master’s degree or PhD in Engineering, Computer…
…needs of AI training and inference at scale. - Cross-functional leadership: Partner with Infrastructure Engineering, Cloud Software Engineering, and SRE to define technical specifications…
…throughput, cost, and reliability of large-scale inference and scoring workloads Partner closely with researchers and engineers across Safeguards to understand their workflows, anticipate…
…technologies. - A strong command of software engineering fundamentals, with a record of building and shipping AI/ML inference systems. - Experience with Docker and Kubernetes…
…an Engineering Manager to lead the team behind our telemetry agent, the software that collects metrics and logs from hosts and GPUs across the…
…THE ROLE We are seeking an experienced Software Engineer with a strong background in Platform Engineering to join our AI Applications team. You will…
…We are looking for an Embedded Software Development engineer to build and own the server related firmware. As an embedded software development engineer in…
…full compute bandwidth of clustered GPUs. About the Role: We are seeking a seasoned Staff Storage Software Engineer with deep experience designing and deploying…
…real-time and batch inference, powering model inference at enterprise scale. We are looking to hire high-agency engineers who bridge the gap between…
…design and implement the software that models the full lifecycle of a physical host from discovery, inference bring-up to GPU driver/CUDA stack…
…algorithms - PhD, or Master's degree - Knowledge of engineering practices and patterns for the full software/hardware/networks development life cycle, including coding standards…
…Bachelor’s degree or higher in Computer Science or related technical field 3+ years of experience in software engineering or ML engineering Experience with…
…You’ll collaborate closely with research engineering, infrastructure, inference, and finance teams. The work requires someone who can move between data engineering, systems engineering…
…Master's degree or PhD in Computer Science, AI/ML, Electrical Engineering, or related field. 8+ years of product management, software engineering, technical marketing…
…data closer to GPU clusters, minimizing idle accelerator time and accelerating training and inference pipelines. We are seeking a seasoned Engineering Manager to lead…
We are now looking for a Senior Research Engineer passionate about Generative AI inference. Are you excited to change the way people infuse AI…
…across Crusoe Cloud’s GPU- and CPU-based infrastructure. You will be collaborating with hardware, software, infrastructure, and vendor engineering teams while working across…
…You'll work across our custom infrastructure — a hybrid training and inference stack spanning our own GPU data centers and the cloud — and the…
…for clusters ranging from 100 to 10,000+ GPUs - Develop deployment strategies for LLM training, inference, and HPC workloads - Present architectural recommendations to technical…