Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “ML Engineer, Inference & Optimization”. A match may be a passing mention rather than the job itself. Titles only.
1,290 roles across 1,495 listings · show every listing · page 4 of 52
…fleets, AI/ML infrastructure, or large-scale inference or training systems. Experience with capacity planning, fleet management, or supply/demand optimization at hyperscale. Familiarity…
…or MLOps/model observability. Published work or open-source contributions in agentic systems or retrieval. Exposure to LLM fine-tuning or inference optimization in…
…or MLOps/model observability. Published work or open-source contributions in agentic systems or retrieval. Exposure to LLM fine-tuning or inference optimization in…
…SambaNova Suite™ is the first full-stack, generative AI platform, from chip to model, optimized for enterprise and government organizations. Powered by the intelligent…
…Scale E2E ML systems. Collaborate with engineering on data contracts, feature stores, distributed training/inference, and automated rollout/rollback; drive architectural investments that increase…
…as causal inference, experimentation, forecasting, propensity modeling, uplift modeling, ranking, or recommendation systems. • Strong judgment about when to use predictive ML, causal methods, generative…
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer…
…Experience with AI frameworks, AI inference systems, AI kernel development, and AI workload optimization. Software Engineering IC3 - The typical base pay range for this…
…validation, and performance optimization. Actively adopt and promote use large language models to automate data engineering tasks such as schema inference, pipeline generation, metadata…
…We work closely with ML researchers and developers to optimize and scale out model training and inference. The team operates at the intersection of…
…8+ years of experience in ML with a proven record of shipping large-scale models to production. Expertise in training and inference optimization. Proven…
…or machine learning engineering, with a strong understanding of distributed computing. (e.g. data pipelines, distributed training and inference, ML infrastructure design). - 3+ years…
…ML Engineer to build efficient, stable foundation models for long-horizon agentic workloads. The role combines model development with systems and hardware optimization. Qualifications…
…Device drivers and kernel integration AI runtimes and execution engines with emphasis on graphs spanning multiple NPU. Compiler technologies and graph optimization ML frameworks…
…and optimizations. Collaborate with product, engineering, and business teams to deliver scalable, production-ready AI systems. Conduct experiments using the latest ML technologies, analyze…
…Experience with AI/ML infrastructure, multi-node GPU clusters, accelerated compute, model training or inference platforms, GPU scheduling, device plugins, Karpenter, cluster autoscaling, CUDA…
…About the role We are seeking a talented and driven ML performance engineer to optimize and scale state-of-the-art foundation models on…
…engineering, with a demonstrated record of technical leadership on large-scale or novel ML systems Deep expertise in LLM training, fine-tuning, inference optimization…
…level ML and AI architectures using the Databricks unified platform, including AI agents, end-to-end pipeline automation, and model training/inference optimization. Lead…
…inference techniques — including quantization, compression, and distillation — for large-scale generative and Mixture-of-Experts recommendation models. Collaborate with research scientists and engineers to…
…Collaborate closely with ML engineers and product stakeholders to productionize recommendation models —defining high-level interfaces, feature contracts, and deployment patterns for batch and…
…closely with engineers and researchers to measure impact and prevent regressions. - Help define the ML roadmap and technical direction for improving Engine agents, and…
…building/evaluating ML and agentic models to optimize guest, host, and business outcomes. A Typical Day: Inference: Develop and apply causal inference methods, including…
…This role blends deep ML engineering expertise with strong analytical judgment to assess, interpret, and improve the behavior of advanced AI models. You will…
…engineering role. But it requires genuine curiosity about how AI inference platforms work. You'll need working fluency in model serving, inference optimization, and…