Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Site Reliability Engineer, Inference Infrastructure”. A match may be a passing mention rather than the job itself. Titles only.
194 roles across 214 listings · show every listing · page 1 of 8
…As a Site Reliability Engineer you will: - Build self-service systems that automate managing, deploying and operating services. - This includes our custom Kubernetes operators…
…inference pipelines Own the full ML model lifecycle: training infrastructure, real-time serving, monitoring, retraining, and iteration Design and implement low-latency, high-reliability…
…reliability and performance, and participate in on-call rotation Required Qualifications 7+ years of experience in DevOps, Site Reliability Engineering (SRE), or Cloud Infrastructure…
…This comprehensive solution empowers sellers to reliably transport products from manufacturing sites to customers worldwide. The FBA team is the core group in charge…
…Experience in building and running cloud infrastructure (private or public cloud) or large-scale service systems that are highly available and reliable. Familiarity with…
…As the engine that enables Microsoft’s cloud-first mission, CO+I delivers reliable, trusted, secure, and sustainable cloud and AI infrastructure to meet…
…Ability to engage in site-reliability engineering practices. #azurecorejobs Software Engineering IC5 - The typical base pay range for this role across the U.S…
Overview The Infrastructure team in Finetuning, Inference and Training (FIT) group is looking for a Senior Software Engineer who loves to build scalable, highly…
…design of infrastructure and tooling for test integration and operations Evaluate proposed designs for manufacturability and integration Interface with both engineers, specialists, and technicians…
…and infrastructure for customer resource consumption, revenue processing, invoicing, and reporting. Our systems power Snowflake's business and enable every other engineering team — and…
…This role works closely with engineering teams, data center technicians, vendors, and operations teams to bring new sites and infrastructure online quickly and reliably…
Overview M365 Copilot inference is a high-impact engineering team advancing applied AI and large-scale machine learning across Microsoft. We design and operate…
…operating highly available distributed systems. - Experience with platform engineering, infrastructure engineering, production engineering, site reliability engineering, or similar disciplines. - Strong systems design, debugging, and…
…or machine learning engineering, with a strong understanding of distributed computing. (e.g. data pipelines, distributed training and inference, ML infrastructure design). - 3+ years…
…Engineering Group, Engineering Group > Software Engineering General Summary: Qualcomm is seeking an exceptional Senior Director, Software Engineering to lead software architecture, development, and execution…
…engineers and product stakeholders to productionize recommendation models —defining high-level interfaces, feature contracts, and deployment patterns for batch and/or real-time inference…
…external customers who need Harvey to be as reliable as the tools it replaces, and internal product engineering teams who need infrastructure they can…
…software development, cloud computing, systems engineering, infrastructure, security, networking, data & analytics) experience - 10+ years of IT development or implementation/consulting in the software or…
…training, to agentic execution infrastructure. - Collaboration and Influence: - Work cross-functionally with Product, Infrastructure, and GTM stakeholders. - Represent Engineering in strategic discussions to influence…
…infrastructure lifecycle management. Experience with inference-serving frameworks, GPU-aware scheduling, or model-performance optimization. Expertise in vector databases, GPU-accelerated query engines, or…
…AI inference systems in a datacenter environment. The engineer will support critical AI use cases by ensuring Qualcomm’s AI infrastructure is reliable, scalable…
…LLM inference, an inter-agent message bus, parametric CAD pipelines, a print farm, and the infrastructure connecting it all. We need an engineer who…
…LLM inference, an inter-agent message bus, parametric CAD pipelines, a print farm, and the infrastructure connecting it all. We need an engineer who…
…infrastructure initiatives 1 Year - London site operating as a high-functioning hub with 15-30+ engineers - Infrastructure teams delivering measurable impact on system reliability…
…Bachelor’s degree in Computer Science or a related field, or equivalent professional experience 5+ years of experience in Site Reliability Engineering, DevOps, Infrastructure…