Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Site Reliability Engineer, Inference Infrastructure”. A match may be a passing mention rather than the job itself. Titles only.
14 roles across 15 listings · show every listing
…inference pipelines Own the full ML model lifecycle: training infrastructure, real-time serving, monitoring, retraining, and iteration Design and implement low-latency, high-reliability…
…reliability and performance, and participate in on-call rotation Required Qualifications 7+ years of experience in DevOps, Site Reliability Engineering (SRE), or Cloud Infrastructure…
…This comprehensive solution empowers sellers to reliably transport products from manufacturing sites to customers worldwide. The FBA team is the core group in charge…
…Experience in building and running cloud infrastructure (private or public cloud) or large-scale service systems that are highly available and reliable. Familiarity with…
…As the engine that enables Microsoft’s cloud-first mission, CO+I delivers reliable, trusted, secure, and sustainable cloud and AI infrastructure to meet…
…Ability to engage in site-reliability engineering practices. #azurecorejobs Software Engineering IC5 - The typical base pay range for this role across the U.S…
Overview The Infrastructure team in Finetuning, Inference and Training (FIT) group is looking for a Senior Software Engineer who loves to build scalable, highly…
…design of infrastructure and tooling for test integration and operations Evaluate proposed designs for manufacturability and integration Interface with both engineers, specialists, and technicians…
…and infrastructure for customer resource consumption, revenue processing, invoicing, and reporting. Our systems power Snowflake's business and enable every other engineering team — and…
…This role works closely with engineering teams, data center technicians, vendors, and operations teams to bring new sites and infrastructure online quickly and reliably…
Overview M365 Copilot inference is a high-impact engineering team advancing applied AI and large-scale machine learning across Microsoft. We design and operate…
…operating highly available distributed systems. - Experience with platform engineering, infrastructure engineering, production engineering, site reliability engineering, or similar disciplines. - Strong systems design, debugging, and…
…or machine learning engineering, with a strong understanding of distributed computing. (e.g. data pipelines, distributed training and inference, ML infrastructure design). - 3+ years…
…Engineering Group, Engineering Group > Software Engineering General Summary: Qualcomm is seeking an exceptional Senior Director, Software Engineering to lead software architecture, development, and execution…