Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Engineer, GPU Inference”. A match may be a passing mention rather than the job itself. Titles only.
974 roles across 1,092 listings · show every listing · page 4 of 39
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…Collaborate with AI Infra, Hardware Engineering, Product, and Research; represent the team to leadership 3+ years managing software engineering teams with a track record…
…Experience with inference-serving frameworks, GPU-aware scheduling, or model-performance optimization. Expertise in vector databases, GPU-accelerated query engines, or distributed data platforms…
…for next-generation NVIDIA GPUs. Advance the state-of-the-art: Solve complex compilation problems for AI workloads (both inference and training) and successfully…
…Engineering Analyze software requirements and collaborate with architecture and hardware engineers to support AI workloads. Build, deploy, and operate components supporting LLM inference, agentic…
…and engineers to focus on AI workloads, not AI infrastructure, unleashing the full compute bandwidth of clustered GPUs. AI training and inference relies on…
…Proven ability to design, implement, and optimize scalable ML architectures, from distributed training to real-time inference. Strong software engineering skills in Python, C…
…LLM inference, an inter-agent message bus, parametric CAD pipelines, a print farm, and the infrastructure connecting it all. We need an engineer who…
…LLM inference, an inter-agent message bus, parametric CAD pipelines, a print farm, and the infrastructure connecting it all. We need an engineer who…
…develop the compiler for Apple's proprietary Neural Engine Accelerator, optimizing it for deep learning inference with a focus on performance, scalability, and power…
…key workloads with ultra high-speed inference. Key Responsibilities - Work with architects, designers, post silicon and software engineers to ensure a high-quality design…
…The Role As a Senior Software Engineer on the ML Infrastructure team, you will design and implement the core backend and infrastructure powering our…
Meta is seeking a Software Engineer to join the MTIA (Meta Training & Inference Accelerator) Software Tooling team, which develops and maintains the tooling ecosystem…
…Required - 10+ years of experience in software engineering with 5+ years in infrastructure, cloud infrastructure, or AI infrastructure roles - 3+ years of people management…
…engineers, and establishing engineering standards across a team or organization Desired Qualifications: Experience supporting AI/ML platforms, inference services, model-serving systems, GPU-backed…
…in Computer Science or a related technical discipline, or equivalent experience 15+ years of software engineering, systems engineering, or technical product/program management experience…
…Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is seeking talented and experienced Software Engineers…
…The role focuses on building, integrating and maintaining complex real-time systems within a large production software stack. Deep understanding of machine learning inference…
…Experience with large-scale distributed systems for AI/ML workloads running on GPUs or TPUs. Strong software engineering skills with experience developing and optimizing…
…GPU/CPU utilization to minimize cloud costs while maintaining low-latency inference for users Collaboration: Work closely with data scientists, data engineers, and software…
…Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About…
…Industry-leading acceleration technologies such as TPUDirect, TPUDirect Storage (TDS), and GPUDirect Storage (GDS). As a Senior Staff Software Engineer, you will lead the…
In this role, you will be a member of the MTIA (Meta Training & Inference Accelerator) Software team and part of the bigger AI and…
…across AWS’ GenAI offerings (ranging from Accelerated Compute, GPUs, Managed Inference, and distributed training and inference platforms), as well as underlying Core and Data…
…Background in software supply-chain security, SBOM tooling ecosystems, vulnerability management, and policy enforcement (OPA/Gatekeeper, Kyverno, etc.). Hands-on work with NVIDIA GPU…