Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Machine Learning Engineer, Fleet Monitoring & Response”. A match may be a passing mention rather than the job itself. Titles only.
31 roles · page 1 of 2
…We support both Bare Metal and Virtual machine instances across a diverse fleet of HW, including clustered GPU platforms. In addition, our customers demand…
…As Director of Core Infrastructure Engineering, you will lead a high-performing engineering organization responsible for evolving RMC and delivering production-ready GPU infrastructure…
…The Engineering organisation within Region Services is structured across core capability pillars: Compute & Machine Learning, Security Identity & Compliance, Storage & Databases, and a growing capability…
…Design and develop responsive, high-performance web and Android applications that serve as the primary interface for monitoring and controlling robotic fleets. 3D & Map…
…autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design…
…autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design…
…testing of multi-disciplinary integrated systems - Experience with robotics, AI, or machine learning applications - Familiarity with software development in Rust, Python, or other languages…
…AWS) Hardware Engineering designs and delivers next-generation cloud infrastructure. Our team builds custom accelerator systems that power AI, machine learning, and compute workloads…
…manual workcell, fleet monitoring that detects performance degradation in real time and alerts Ops, and deep-dive tooling that enables engineers and scientists to…
…Partner closely with storage software, networking, control plane, Kubernetes, observability, compute, and fleet engineering teams to deliver cross-functional infrastructure initiatives, define and track…
…Lead incident response and recovery, then drive corrective actions to completion. What We Need to See: Bachelor’s degree in Computer Science, Engineering, or…
…fastest-growing infrastructure fleets in the industry — spanning multiple accelerator families, cpu families and clouds. The Capacity Engineering team is responsible for making sure…
…Lead incident response and recovery, then drive corrective actions to completion. What We Need to See: Bachelor’s degree in Computer Science, Engineering, or…
…autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design…
…autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design…
…Key job responsibilities Design & Build: Architect and implement the node-level runtime infrastructure that powers EKS compute. Design optimized Amazon Machine Images (AMIs), integrate…
…Responsibilities Develop, optimize, and productionize machine learning models for Cloudflare’s serverless inference platform, with a focus on performance, reliability, and model quality. Build…
…to onboard new workcell deployments without dedicated engineering support - Design fleet-wide observability systems: health monitoring, sensor anomaly detection, performance alerting, and proactive diagnostics…
…Engineering team, focused on building and scaling the infrastructure and software systems that power large-scale machine learning workloads across Meta's production fleet…
…Want to learn more? About the role As a Fleet Operations Specialist on the Autonomous Vehicle Operations team, you will oversee frontline operators and…
…The Engineering organisation within Region Services is structured across core capability pillars: Compute & Machine Learning, Security Identity & Compliance, Storage & Databases, and a growing capability…
…However, we do look to capture information about care responsibilities, and disabilities among other diversity information as part of an optional DEI Monitoring form…
…We are responsible for a service that: - Reliably manages a large fleet of cloud-native OpenSearch clusters and collections, freeing customers from sizing, scaling…
…Engineering, establishing incoming quality requirements and process controls at all supplier sites. - Build and maintain the reliability prediction and monitoring infrastructure — ensuring fleet performance…
…As a member of the Cloud-Scale Machine Learning Acceleration team, you will be the interface between the system engineering team and the ODM…