Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Principal Engineer, AI Inference Reliability”. A match may be a passing mention rather than the job itself. Titles only.
50 roles across 53 listings · show every listing · page 1 of 2
…Experience building or supporting production AI/ML platforms (training, deployment, and model serving/inference), including GPU infrastructure/tooling. Strong DevOps/platform engineering practices: CI…
Meta is seeking a principal-level Compiler Architect to work across the full compiler stack for MTIA (Meta Training and Inference Accelerator). In this…
…for orbital AI inference. As we continue to upgrade and expand the constellation, we’re looking for best-in-class engineers to join the…
…Develop responsible-AI principles and standards addressing transparency, explainability, fairness, privacy, security, human oversight, reliability, and regulatory compliance. Maintain an enterprise inventory of material…
…AI inference systems in a datacenter environment. The engineer will support critical AI use cases by ensuring Qualcomm’s AI infrastructure is reliable, scalable…
…We are seeking an AI Engineer specializing in algorithm evaluation and agentic systems design for advanced computer vision and video understanding algorithms. What We…
…As a software engineer in AiDP reliability engineering you will work on one or many projects related to GenAI, ML, Inference and Big data…
…As the Principal Software Engineer on our team, you would have the opportunity to work on: ONNX: an open standard format for representing AI…
…robust systems for deployment, monitoring, scalability, and reliability of infrastructure supporting AI/ML workloads - Collaborate with engineering leaders globally to align on technical standards…
…As a Principal DevOps / Site Reliability Enginee r on the AI Efficiency team, you will own and evolve the operational foundations that allow the…
…AI-powered assistants, commerce experiences, or personalization platforms. Experience optimizing distributed training and inference systems on large GPU clusters. Experience mentoring principal-level engineers…
…and production-ready engineering. Our team is helping enable AI-native manufacturing by turning fragmented operational knowledge and data into reliable intelligence that improves…
…If you are passionate about reliability engineering, test infrastructure, and AI-scale systems, this role is for you. What you’ll be doing Design…
…scale AI inference acceleration across global deployments. We are seeking a Senior Data Center Operations Engineer to support Qualcomm’s Cloud AI data center…
…generative modeling and retrieval-augmented systems; Optimizing and scaling AI models for real-time inference, low-latency deployment, and cost-effective performance in cloud…
…reliability. Design, develop, and evaluate next-generation AI models and systems, from pre-training and post-training methodologies to novel architectures and scalable inference…
…behavior — so engineers can implement reliably without ambiguity Review and validate that production results match expected statistical behavior, partnering with engineers on edge cases…
…Experience with NVIDIA AI platforms, including CUDA, CUDA-X libraries, TensorRT-LLM, Triton Inference Server, NIM, NeMo, Megatron, Transformer Engine, NCCL, DGX, NVLink, InfiniBand…
…and the growing set of AI-powered experiences shipping to restaurants. Toast is seeking a Senior Principal Software Engineer to act as the technical…
…Define and own the technical architecture of critical machine learning systems, including model training pipelines, inference infrastructure, and feature engineering platforms, ensuring reliability and…
…a core part of how we approach AI safety Minimum qualifications Experience applying privacy engineering principles in production systems, including privacy by design, data…
…generation AI training and inference at hyperscale. The Platform Systems Engineering (PSE) team is seeking a Principal AI Network Hardware Systems Engineer to lead…
…of applied scientists, system engineers, and product managers to deliver state-of-the-art inference capabilities for Generative AI applications. Your work will involve…
…including model integrations, inference workloads, evaluation pipelines, observability, and the reliable execution of agentic workflows. QUALIFICATIONS - 8+ years of engineering experience. - Worked on Platform…
…infrastructure that powers the next generation of AI training and inference at scale. As a Staff Engineer on our Orchestration team, you will collaborate…