Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Engineer, GPU Inference”. A match may be a passing mention rather than the job itself. Titles only.
976 roles across 1,097 listings · show every listing · page 3 of 40
…not need specialized experience with GPUs or low-level hardware. QUALIFICATIONS - Five or more years of software engineering experience building production infrastructure. - Strong programming…
…Prometheus, Grafana, OpenTelemetry, and alerting people actually act on. - GPU and ML infrastructure exposure: GPUs, inference serving, or distributed training. - AI-assisted operations: Building…
…3+ years of product management experience shipping technical, platform, infrastructure, or enterprise products, or equivalent engineering-to-PM experience. Experience with AI inference platforms…
We are looking for a Senior Inference Engineer to own inference for real-time multimodal conversational AI. This is a full-stack inference role…
…data or software or machine learning engineering, with a strong understanding of distributed computing. (e.g. data pipelines, distributed training and inference, ML infrastructure…
…This senior software engineering role is part of the Machine Learning Inference Applications team and focuses on delivering high-performance model inference solutions for…
…Knowledge of GPU architecture, CUDA, Triton, custom kernels, or hardware-aware optimization. Familiarity with distributed training and model parallelism. Experience with AWS Trainium, Inferentia…
…mentor other engineers. Ways To Stand Out From The Crowd: Experience building platforms for AI/ML training, inference, model serving, GPU-accelerated workloads, distributed…
…8+ years of experience leading large, globally distributed engineering organizations. Experience delivering software platforms for complex SoC, accelerator, GPU, NPU, CPU, DSP, datacenter, or…
…increasingly demanding cloud native, AI, and GPU workloads. We are looking for a senior IC5 software engineer with deep Kubernetes expertise, required cloud infrastructure…
…The technical frontier you'll help define is heterogeneous, disaggregated inference - GPU on prefill, the RDU on decode - which explores hard problems across networking…
…inference systems. Familiarity with quantization, graph optimization, kernel fusion, and model partitioning. Experience with frameworks such as DeepSpeed, Megatron, vLLM, or TensorRT. Strong GPU…
…Engineering, or related field Experience with hardware-software co-design with non-GPU accelerators Publications or open-source contributions in LLM training or inference…
…own inference services - Support and debug production issues through on-call rotation Required Qualifications - Have 6+ years of experience in software engineering, with a…
…complex software systems and continuously improve the quality of other AI agents. ABOUT THE ROLE: We’re looking for an experienced research engineer to…
…As a Software Engineer II on the team, you will help design, build, and harden a secure, evaluable, vendor-agnostic agent platform that teams…
…1) AI Infrastructure with purpose-built chips like AWS Trainium and Inferentia, GPU-powered instances, and optimized frameworks like PyTorch and JAX, 2) AI…
By submitting your resume, you acknowledge that your 2027 Software Engineering internship application will be processed in accordance with NVIDIA’s Applicant Privacy Policy…
…are optimized for Training and Inferencing. Lead the development of end-to-end AI network architecture including gpu-gpu, server – storage, in-rack cabling…
…Experience building or supporting production AI/ML platforms (training, deployment, and model serving/inference), including GPU infrastructure/tooling. Strong DevOps/platform engineering practices: CI…
…You will own hard problems in inference efficiency, hardware/software codesign, systems architecture, and tooling, and help set the technical direction for the engineers…
…software engineering with deep specialization in compiler development, code generation, or performance optimization for accelerators Experience architecting production compiler infrastructure for ML accelerators, GPUs…
…GPU utilization and efficiency, interconnect (NVLink, InfiniBand, RoCE) fabric behavior, distributed training and inference throughput, and hardware degradation signals. Build the detection and diagnosis…
…Leadership & Strategy: - Lead, mentor, and grow a team of high-caliber software engineers on Crusoe’s - Partner with leadership to define and execute the…
…Infrastructure Engineering team, you will drive end-to-end performance characterization, bottleneck analysis, and optimization of large-scale AI training and inference clusters. In…