Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Staff Machine Learning Engineer – Model Optimization & Quantization”. A match may be a passing mention rather than the job itself. Titles only.
19 roles across 20 listings · show every listing
…Work Optimize Nuro’s autonomy stack with cutting-edge optimization techniques like quantization, distillation, and model compression. Work with autonomy engineers to optimize, validate…
…Engineers here bring deep experience across compilers and program analysis, optimization algorithms, computer architecture, machine learning systems, and the practical craft of getting large…
…We are particularly excited about methods for post-training model optimization (pruning, quantization, NAS), efficient architecture design, adaptive/dynamic inference, resource-efficient training and…
…Prior experience in distributed training at scale and optimization techniques like model pruning, compression, quantization & distillation. Prior experience building AI/ML tooling for model…
…model architectures Good understanding of Quantization (8-bit, 4-bit) and Calibration algorithms Good understanding of machine learning compiler techniques and graphs optimizations Good…
…staff machine learning engineer, you will be responsible for fine-tuning state-of-the-art LLMs for diverse use cases while optimizing models for…
…a team of machine learning engineers in the design and implementation of state-of-the-art Visual Language Action Models (VLAs), VLMs, LLMs used…
…Familiarity with AI Inference Engines (TensorFlow Lite, ONNX Runtime, …), local LLM frameworks (Ollama, llama.cpp, …) and model optimization techniques (quantization, pruning, fine tuning, …). High…
…AI model deployment and optimization. The candidate will be responsible for evaluating, developing, and optimizing cutting edge machine learning and deep learning algorithms, and…
…Model inference Execution graphs Quantization or optimization (coursework or projects acceptable) Interest in Generative AI and agentic AI systems . Programming & Tooling Experience with Python…
…a Staff Machine Learning Scientist, you will use your experience to focus on designing, developing, and evaluating state-of-the-art foundation models, at…
…Engineering Group, Engineering Group > Machine Learning Engineering General Summary: THIS IS A FULL-TIME ONSITE ROLE REQUIRING 5 DAYS A WEEK IN OFFICE AT…
…and IOT products through machine learning hardware and software. We are looking for a Senior or Staff level Engineer to work on bleeding-edge…
…Ideal candidates will have a perspective across areas including machine learning fundamentals, quantization and numerical methods for machine learning model optimization, digital VLSI circuits…
…Engineering Group, Engineering Group > Software Engineering General Summary: Job Overview: The Qualcomm Cloud AI team is developing hardware and software for Machine Learning solutions…
…Our inference engine is designed to help developers run neural network models trained in a variety of frameworks on Snapdragon platforms at blazing speeds…
…We are particularly excited about methods for post-training model optimization (pruning, quantization, NAS), efficient architecture design, adaptive/dynamic inference, resource-efficient training and…
…Engineering Group, Engineering Group > Machine Learning Engineering General Summary: About Qualcomm AI Research Qualcomm AI Research is at the forefront of advancing the state…
…Engineering Group, Engineering Group > Machine Learning Engineering General Summary: About Qualcomm AI Research Qualcomm AI Research is at the forefront of advancing the state…