Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Engineer, GPU Inference”. A match may be a passing mention rather than the job itself. Titles only.
522 roles across 573 listings · show every listing · page 1 of 21
…We are hiring a Software Engineer to productionize and optimize our GPU serving stack, working across our custom inference APIs, the vLLM serving runtime…
…Develop and productize inference models on NVIDIA GPUs and Nvidia RTX Spark platforms as SDK/Microservices after optimization. Technical Mentorship: Mentor senior engineers and…
…Development experience with NVIDIA software libraries and GPUs, including CUDA and CUDA-X libraries. Experience with Kubernetes, distributed training, and large-scale inference. Experience…
…A minimum of 12+ years of overall professional experience in the technology industry in software engineering, developer relations, technical partnerships, solutions architect, product management…
…software and/or firmware engineering Strong analytical skills, technical leadership, creativity, agility, and ability to solve complex technical problems Deep understanding of DSP/GPU…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…Our workloads span thousands of GPUs, petabytes of driving data, and geographically distributed training and inference clusters. As Architect for AI Infrastructure, you will…
…Experience with AI inference acceleration features and accelerator or Graphics Processing Unit (GPU) performance analysis. Experience with the AI inference software stack, including compilers…
…with product experimentation, online evaluation, and A/B testing frameworks. - Strong software engineering skills with the ability to write clean, maintainable, and scalable code…
…A day in the life You work alongside customer engineering teams during live AI implementation sprints — debugging inference pipelines, optimizing RAG architectures, tuning agent…
…working within software development or Internet-related industries - Experience with Machine Learning and Large Language Model fundamentals, including architecture, training/inference lifecycles, and optimization…
…NVIDIA is looking for a Senior System Software Engineer with deep expertise in speech technologies to support enterprise and developer customers. This role involves…
NVIDIA is seeking a Senior Systems Software test (lead) Engineer to join our Cloud Service Provider (CSP) Engagements team, focusing on ML software stack…
…About the Role We're hiring a Software Engineer to help contribute to projects on our Inference Platform team. Our team primarily owns the…
…inference infrastructure, model serving systems, or GPU-accelerated workloads. Location: Sunnyvale or Toronto preferred Why Join Cerebras People who are serious about software make…
…host GPU-accelerated machine learning workloads Deploy and manage core infrastructure such as databases, monitoring and storage Closely collaborate with software engineers to create…
…ML training and inference workloads. Develop tooling that helps ML engineers debug, profile, optimize, and monitor model performance. Improve GPU and general resource utilization…
…with both engineers and non-engineers Desired Qualifications: Experience supporting AI/ML platforms, inference services, model-serving systems, data pipelines, or GPU-backed workloads…
…software development, cloud computing, systems engineering, infrastructure, security, networking, data & analytics) experience - Experience in external enterprise customer-facing role as a technical lead, with…
…Engineering Group, Engineering Group > Software Engineering General Summary: AI Inference Accelerator – Test & Automation - Senior Engineer you will be working for the Qualcomm Cloud AI100…
…fractional GPUs (MIG, MPS, time-slicing), GPU scheduling, training vs. inference fleets, and multi-tenant GPU isolation. --- The ideal candidate is a strategic thinker…
…software engineering and AI/ML engineering methodologies and best practices. - Collaborate with cross-functional teams, including Product Managers, Applied Scientists, and Data Engineers, to…
…Preferred qualifications, capabilities, and skills Foundational understanding of NVIDIA GPU infrastructure software (e.g., DCGM, BCM, Dynamo Inference). Proficiency with observability tools like Prometheus…