Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Development Engineer, AI/ML, AWS Neuron, Model Inference”. A match may be a passing mention rather than the job itself. Titles only.
49 roles · group by role · page 1 of 2
…AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon’s custom machine learning accelerators, Inferentia…
…on the customer AWS Trainium and Inferentia silicon and servers. Strong software development using Python, System level programming and ML knowledge are both critical…
…Join us to optimize the latest models to run really fast on the Trainium hardware. As a Sr. Software Development Engineer on the Inference…
…Key job responsibilities This role requires collaborating with other Neuron Software teams, Science, AWS AI Services, external partners and customers with a potential high…
…model forward pass. You'll work closely with engineers across inference, compilers, kernels, and ML systems to identify performance bottlenecks and build the software…
We are looking for an AI/ML Engineer to build efficient, stable foundation models for long-horizon agentic workloads. The role combines model development…
We are looking for an AI/ML Engineer to build efficient, stable foundation models for long-horizon agentic workloads. The role combines model development…
…We combine deep hardware knowledge with ML expertise to push the boundaries of what's possible in AI acceleration. The AWS Neuron SDK, developed…
…engineering, hardware design and verification, software, and operations. AWS Nitro, ENA, EFA, Graviton and F1 EC2 Instances, AWS Neuron, Inferentia and Trainium ML Accelerators…
…About AWS Neuron: AWS Neuron is the software of Trainium and Inferentia, the AWS Machine Learning chips. Inferentia delivers best-in-class ML inference…
…About AWS Neuron: AWS Neuron is the software of Trainium and Inferentia, the AWS Machine Learning chips. Inferentia delivers best-in-class ML inference…
…As an SDM for the LLM Inference Model Enablement team, you will lead a team of expert AI/ML engineers to onboard and optimize…
…engineering, hardware design, software and business development. Along with the AI Chips Inferentia and Trainium, Annapurna Labs has delivered advancements in Networking with AWS…
…Large Language Model fundamentals, including architecture, training/inference lifecycles, and optimization of model execution, or experience (non-internship) in professional software development - Experience engaging…
…software development, cloud computing, systems engineering, infrastructure, security, networking, data & analytics) experience - Experience in external enterprise customer-facing role as a technical lead, with…
…AWS Neuron is the software stack for Trainium and Inferentia, the AWS Machine Learning chips, delivering best-in-class ML performance in the cloud…
…software development, cloud computing, systems engineering, infrastructure, security, networking, data & analytics) experience - Experience communicating across technical and non-technical audiences and at C-level…
…best software engineering practices in large-scale systems - Experience in developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration…
…software engineering experience, with significant time as the technical lead or anchor on a platform, inference runtime, or ML infrastructure team Experience with ML…
…strong interest in LLM serving; prior inference or ML experience is not required Have significant software engineering experience, with a strong background in high…
…About the role The Cloud Inference team scales and optimizes Claude to serve the massive audiences of developers and enterprise companies across AWS, GCP…
…minds in the industry on next generation AI/ML hardware that powers AWS's training and inference infrastructure. Your analysis will directly shape architectural…
…Development: Designing high-performance kernels optimized for our ML accelerator architectures About the team AWS Neuron is the software of Trainium and Inferentia, the…
…AWS Neuron is the SDK that optimizes the performance of complex ML models executed on AWS Inferentia and Trainium, our custom chips designed to…
The AWS Neuron Compiler team is actively seeking skilled compiler engineers to join our efforts in developing a state-of-the-art deep learning…