Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Staff Software Engineer, GPU Inference”. A match may be a passing mention rather than the job itself. Titles only.
106 roles across 114 listings · show every listing · page 1 of 5
…with product experimentation, online evaluation, and A/B testing frameworks. - Strong software engineering skills with the ability to write clean, maintainable, and scalable code…
…About the Role We're hiring a Staff Engineer to help lead, drive, and contribute to projects on our Inference Platform team. Our team…
…ML training and inference workloads. Develop tooling that helps ML engineers debug, profile, optimize, and monitor model performance. Improve GPU and general resource utilization…
…high-performance C++ software architecture for mission-critical systems or ML inference. Proven track record of leading complex, cross-functional engineering projects as a…
About the Role We are looking for a Staff/Lead Software Engineer, AI Infrastructure, to play a critical role in building and scaling the…
…GPUs, Neuron, TPU or other AI acceleration hardware - Experience directly managing scientists or machine learning engineers - Experience debugging, profiling, and implementing best software engineering…
…About the team The team is comprise of both Hardware Design Engineers, System Design Engineers, Software Development Engineers and Technical Program Managers, all with…
…About the team The team is comprise of both Hardware Design Engineers, System Design Engineers, Software Development Engineers and Technical Program Managers, all with…
…We're looking for a Staff Engineer to be a technical lead for Inference Runtime: the team that owns the shared, accelerator-agnostic core…
…Shield AI is seeking a Senior / Staff C++ Software Engineer, Edge Systems to help define and build the software foundations that power autonomous systems…
The Senior Principal AI Agent / ML Software Engineer is a Senior Staff-level, hands-on technical leadership role responsible for defining, building, and operating…
…Proficiency with CUDA and NVIDIA GPU programming for accelerating quantum simulation, AI model training, or real-time inference workloads at scale. Widely considered to…
…drift behavior, enabling accurate performance prediction and parameter inference without full experimental overhead. Develop GPU-accelerated implementations to ensure the full simulation and modeling…
…Develop intelligent caching and tiered storage architectures to achieve extreme IOPS and cluster-wide throughput at GPU scale for training and inference workloads. Tune…
…Experience with AI-focused hardware(GPUs) or software(Training/Inferencing). Scripting/Automation tooling background Industry-standard data center certifications. NVIDIA is widely considered to…
…Demonstrated ability to work cross-functionally, collaborating effectively with ML engineers, data scientists, software engineers, product managers, and business stakeholders. The ability to thrive…
…experience focused on data, software, or ML engineering, with understanding of distributed computing (e.g., data pipelines, training and inference, ML infrastructure design) - 5…
…Mentor senior and staff engineers, conduct architecture reviews, provide design feedback, and improve technical standards. Bachelor’s degree in Computer Science, Software Engineering or…
…on AWS Inferentia and Trainium, our custom chips designed to accelerate deep-learning workloads This role is for a senior software engineer in the…
…About the team The team is comprised of both Hardware Design Engineers, System Design Engineers, Software Development Engineers and Technical Program Managers, all with…
…Prior professional experience as a Machine Learning Engineer, Data Scientist, or Backend Software Engineer before transitioning into Product Management. Understanding of the Ad Tech…
…We're looking for a Staff ML Engineer to drive the model serving layer for voice workloads. You'll work hands-on with inference…
Staff Software Engineer - AI Research Infrastructure P-1215 At Databricks, we are obsessed with enabling data teams to solve the world’s toughest problems…
…engineering team - Knowledge of Machine Learning and LLM fundamentals, including transformer architecture, training/inference lifecycles, and optimization techniques - 5+ years of full software development…
…a Staff ML Performance Engineer, you’ll play a key role in high-impact projects, optimising ML inference for edge accelerators and GPUs. The…