Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “ML Engineer, Inference & Optimization”. A match may be a passing mention rather than the job itself. Titles only.
272 roles across 292 listings · show every listing · page 1 of 11
…level ML and AI architectures using the Databricks unified platform, including AI agents, end-to-end pipeline automation, and model training/inference optimization. Lead…
…inference techniques — including quantization, compression, and distillation — for large-scale generative and Mixture-of-Experts recommendation models. Collaborate with research scientists and engineers to…
…Collaborate closely with ML engineers and product stakeholders to productionize recommendation models —defining high-level interfaces, feature contracts, and deployment patterns for batch and…
…closely with engineers and researchers to measure impact and prevent regressions. - Help define the ML roadmap and technical direction for improving Engine agents, and…
…building/evaluating ML and agentic models to optimize guest, host, and business outcomes. A Typical Day: Inference: Develop and apply causal inference methods, including…
…or Kubernetes for ML job orchestration Experience with hyperparameter optimization and experiment tracking tools\ Background in ML Engineering, AI Engineering, MLOps, or LLMOps Prior…
…This role blends deep ML engineering expertise with strong analytical judgment to assess, interpret, and improve the behavior of advanced AI models. You will…
…engineering role. But it requires genuine curiosity about how AI inference platforms work. You'll need working fluency in model serving, inference optimization, and…
…etc.) to determine optimal configuration per model size. Requirements Bachelor's or Master's degree in Computer Science, Computer/Electrical Engineering, or a related…
…ML conferences. Software engineering skills in Python Experience in developing large computer vision and machine learning models, particularly on the hardware-aware model optimizations…
…Amazon.com's recommendations engine is driven by ML, as are the paths that optimize robotic picking routes in our fulfillment centers. Our supply…
…and deploy AI/ML systems that process threat data at scale, running over large-scale security logs with real-time inference - Expand and improve…
…organization, the MLS team is tasked with developing the next generation of EC2 Supercomputers, optimized for high-performance training and inference workloads. We are…
…Work may span the full lifecycle of modern ML systems: from architecture design/training to post-training optimization and inference acceleration. You will contribute…
…Deep understanding of AI/ML workflows, data pipelines, training/inference patterns, Vector DB, KV Cache and enterprise cluster deployments and ability to collaborate with…
…Partners with data science, ML engineering, and application teams to translate model and compute requirements into platform standards and deployment patterns. Optimizes platform reliability…
…Experience with modern natural language processing architectures such as transformer-based models and techniques for optimization and efficient inference. Experience with machine learning operations…
…Execute model optimization within strict millisecond latency budgets at the internet edge, uniquely balancing inference costs against incremental value while maintaining fleet-wide fail…
…challenges spanning ML workload analysis, graph-level optimizations, memory hierarchy management, and hardware-software co-design. This is a role for engineers who build…
…Query Optimization & Performance - Profile and tune query engines against columnar and time-series stores so that the observability layer meets its own strict P99…
…Experience mentoring engineers and setting technical direction across teams Track record of open-source contribution in the kernel, compiler, or ML systems ecosystem Experience…
…AI/ML systems, distributed/HPC systems, software tooling and infrastructure or model optimization with GPUs/accelerators Demonstrated ability to engage deep in technical work…
…prototyping, ablations, training, evaluation, optimization, and production deployment. - Work closely with engineering teams to integrate models into real-time systems, ensuring reliability, uptime, and…
…implement classical ML models and LLM based inferences to solve defined problems. You'll work closely with senior scientists and engineers, contributing to model…
…discussions with Principal Engineers and SDMs, identifying dependencies, scaling factors, boundary conditions, and risks across distributed systems serving ML inference at global scale - Represent…