Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Staff Software Engineer, GPU Inference”. A match may be a passing mention rather than the job itself. Titles only.
208 roles across 227 listings · show every listing · page 6 of 9
…Our first products are heavily focused on inference. Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is…
…large-scale AI systems and GPU compute infrastructure from the ground up. As a Staff Cloud Site Reliability Engineer at Wayve, you will build…
…SLOs) that help the team avoid repeating failures Minimum qualifications Significant software engineering experience building and operating production distributed systems Proficiency in at least…
…Or specialized experience in runtime optimizations, model quantization, compression, on-device inference, GPU inference, pytorch, kernel development Your Location: This position is US - Remote…
…The Role Senior / Staff Software Engineer on the QAIRT Windows platform team, responsible for the full-stack development and long-term health of the…
…Our first products are heavily focused on inference. Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is…
…As a Senior Reliability Engineer you will engage with an experienced cross-disciplinary staff to conceive and design infrastructure technologies. You will work closely…
…Engineering Group, Engineering Group > Software Engineering General Summary: As a leading technology innovator, Qualcomm pushes the boundaries of what is possible to enable next…
…and communicate effectively at all levels Preferred qualifications 10+ years of software engineering experience, including time as a technical lead setting direction for a…
…Key job responsibilities - Understand and improve state-of-the-art in Vector Search - Optimize inference for semantic matching and ranking models - Improve search engine…
…Solid understanding of on-device AI / edge AI technology stack , including inference engines, model quantization, and LLM deployment on embedded platforms. Familiarity with mainstream…
…analyzing deep learning workloads on hardware accelerators (GPUs, TPUs, ASICs, FPGAs, or others) - Solid software engineering fundamentals with an eye toward auditability and maintainability…
…Proven track record building or optimizing large-scale distributed ML systems (training/inference optimization, GPU utilization, multi-GPU/TPU setups, hardware co-design). Deep…
…software development, cloud computing, systems engineering, infrastructure, security, networking, data & analytics) experience - 3+ years of design, implementation, or consulting in applications and infrastructures experience…
Our Machine Learning Acceleration (MLA) team develops the Inferentia and Trainium SOCs that are used to power today’s AI workloads in datacenters all…
…Kubernetes, GPU scheduling, autoscaling inference workloads. QUALIFICATIONS - 3+ years of professional software engineering experience with meaningful work on ML inference or high-performance systems…
…Engineering Group, Engineering Group > Software Engineering General Summary: We are seeking an exceptional Staff Software Engineer to join our ML Platform team. This role…
…with GPU-based workloads and AI infrastructure, including training and inference characteristics, scheduling challenges, and performance considerations - Experience working across hardware and software boundaries…
…or system-level simulations for SoCs, ASICs, GPUs, or CPUs - Think of yourself as a software engineer first, with deep domain knowledge in chip…
…Our team builds C++ models of these custom SoCs that RTL designers, verification engineers, and software teams depend on throughout the silicon development lifecycle…
…We're looking for a Systems Software Engineer who wants to work at the boundary between hardware and software in both pre-silicon and…
…About the Role As a Senior/Staff Deep RL Engineer, you will design, train, and deploy deep reinforcement learning policies that make real-time…
…Our team builds C++ models of these custom SoCs that RTL designers, verification engineers, and software teams depend on throughout the silicon development lifecycle…
…engineering team - 7+ years of professional experience developing firmware, drivers, runtime software, or low-level systems software for custom hardware (SoCs, ASICs, GPUs, CPUs…
…Engineering Group, Engineering Group > Software Engineering General Summary: As a leading technology innovator, Qualcomm pushes the boundaries of what is possible to enable next…