Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Staff Machine Learning Engineer – Model Optimization & Quantization”. A match may be a passing mention rather than the job itself. Titles only.
38 roles · group by role · page 1 of 2
…In this role you will develop tools to help developers optimize and deploy machine learning models on edge and mobile hardware. AIMET is Qualcomm…
…Optimize model latency, memory usage, and execution speed through quantization, distillation, and pruning. Develop Agentic Workflows: Design and implement robust agentic architectures, multi-agent…
…of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience. • 6+ months of experience developing and/or optimizing machine learning models, systems, platforms…
…model optimization for edge deployment — quantization, distillation, TensorRT, or equivalent - Publications, patents, or significant open-source contributions in computer vision, robotics perception, or machine…
…PhD in Computer Science, Machine Learning, or Robotics, with a research focus on Reinforcement Learning, Foundation Models, or Multi-Modal learning. Substantial involvement in…
…Experience with machine learning infrastructure, C++, performance, GPU programming, mobile GPU. Preferred qualifications: Master’s degree or PhD in Engineering, Computer Science, or a…
…deep generative models, Bayesian deep learning, equivariant CNNs, Bayesian optimizations, reinforcement learning, unsupervised learning, and graph NNs. Drives systems innovations for model efficiency advancement…
…Decisioning & optimization: causal inference, policy optimization, constrained optimization, or reinforcement learning. Training & inference efficiency: model sparsification, quantization, distillation, or parallelism and partitioning design. - Communication…
Meta is seeking a Staff Software Engineer to join the Core Machine Learning team, focused on building and scaling the foundational ML infrastructure and…
…About the role As a Senior Staff Machine Learning Scientist, you own the inference and optimization layer that makes AI in agentic workflows fast…
…You will join a team of world-class machine learning engineers hungry to apply leading-edge technologies to deliver extraordinary experiences to our customers…
…You will join a team of world-class machine learning engineers hungry to apply leading-edge technologies to deliver extraordinary experiences to our customers…
…and systems engineering activities. Job responsibilities include: Implement sensor signal processing and machine learning algorithms across various embedded SOCs. Debug, verify, optimize, and tune…
…As a Machine Learning and System Optimization Engineer, you will orchestrate and allocate overall system capacity to various core perception models running on-bot…
…Machine Learning, Robotics, or a related field. 5+ years of experience with deep learning architectures (especially Transformers, Diffusion Models, MoEs), algorithms, and optimization techniques…
…large-scale data engineering, hardware-savvy optimizations, and reproducible experimentation—researchers can produce impactful, trustworthy advancements in foundational deep learning. FOUNDATIONAL PAPERS This job…
…Work Optimize Nuro’s autonomy stack with cutting-edge optimization techniques like quantization, distillation, and model compression. Work with autonomy engineers to optimize, validate…
…Engineers here bring deep experience across compilers and program analysis, optimization algorithms, computer architecture, machine learning systems, and the practical craft of getting large…
…We are particularly excited about methods for post-training model optimization (pruning, quantization, NAS), efficient architecture design, adaptive/dynamic inference, resource-efficient training and…
…Prior experience in distributed training at scale and optimization techniques like model pruning, compression, quantization & distillation. Prior experience building AI/ML tooling for model…
…Prior experience in distributed training at scale and optimization techniques like model pruning, compression, quantization & distillation. Prior experience building AI/ML tooling for model…
…model architectures Good understanding of Quantization (8-bit, 4-bit) and Calibration algorithms Good understanding of machine learning compiler techniques and graphs optimizations Good…
…staff machine learning engineer, you will be responsible for fine-tuning state-of-the-art LLMs for diverse use cases while optimizing models for…
…a team of machine learning engineers in the design and implementation of state-of-the-art Visual Language Action Models (VLAs), VLMs, LLMs used…
…Familiarity with AI Inference Engines (TensorFlow Lite, ONNX Runtime, …), local LLM frameworks (Ollama, llama.cpp, …) and model optimization techniques (quantization, pruning, fine tuning, …). High…