Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Manager, Site Reliability Engineering - Storage Layer Service”. A match may be a passing mention rather than the job itself. Titles only.
29 roles across 35 listings · show every listing · page 1 of 2
…As the Site Reliability Engineering Manager for SLS, you will partner with the teams building these storage services to define SLOs, shape capacity plans…
…Software Engineer , you will be responsible for the high-level orchestration of our grid-scale storage sites. Operating at the "System Coordination" layer, you…
…As an L3 engineer, you will act as a key driver of execution within our core services. You will be heavily involved in modernizing…
…WHAT YOU BRING Demonstrated expertise in software, production, or site reliability engineering, with a proven ability to design, build, and lead complex backend services…
…Evaluate emerging modular technologies and design methodologies to relentlessly improve energy efficiency and site performance. What You’ll Bring to the Team: - Engineering Foundation…
…Bonus Skills Strongly prefer an MS or PhD in Computer Science, ideally focusing on database management and/or storage engines. #LI-IM1 #LI-REMOTE…
…Exposure to site reliability engineering (SRE) practices. Exposure to AI-assisted development and data-driven engineering workflows. Knowledge of Azure resource providers, platform extensibility…
…Demonstrated project management experience, including planning, tracking, and delivering complex infrastructure projects, and infrastructure engineering knowledge are required as this role manages complex infrastructure…
…do, document, automate - Essential Skills & Experience - 8+ years experience in Site Reliability Engineering, Infrastructure Engineering, or similar roles in large-scale distributed production environments…
…the third-party services we depend on. You'll build the abstractions and standards that let dozens of adjacent engineering teams — Billing Infrastructure, Patient…
…compute, storage, and networking services - Experience building or maintaining data pipelines, ETL systems, or ML training/serving infrastructure - Understanding of system reliability principles including…
…Kubernetes, streaming data infrastructure, columnar lakehouse storage, and a TypeScript/React frontend. We’re looking for engineers willing and eager to work on the…
…You - 8+ years experience in incident management, site reliability engineering, or infrastructure operations - Experience managing incidents in large-scale distributed infrastructure environments - Strong understanding…
…You’ll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You…
…managers, site leads, and contracted services - Establish clear accountability structures, on-call protocols, and escalation paths for 24/7 operations - Source, negotiate, and manage…
…You'll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You…
…We seek a Distinguished Engineer to lead NVIDIA's storage strategy for AI Cloud across the Neocloud Provider (NCP) and Cloud Service Provider (CSP…
…storage, security, identity and infrastructure management - Optimize Linux system configuration including kernel, driver, filesystem and services to support workloads running via our orchestration layer…
…of core infrastructure services: DNS, DHCP, and network storage (NFS). - Hands-on experience with Palo Alto Networks firewalls, including policy management, threat prevention, and…
…key player in siting and scaling infrastructure to support high-performance computing and AI workloads, and helping Crusoe pioneer reliable, energy-first compute at…
…and storage layers Serve as the product owner for the source of truth for customer data across the platform Partner closely with engineering to…
…Kubernetes platforms - Manage storage and networking layers - Work with CSI drivers, persistent volumes, and cross-cloud networking to ensure data reliability and connectivity - Develop…
…building these storage services to define SLOs, shape capacity plans, and ensure the reliability, durability, and operational safety of the storage layer that underpins…
…Introduction), SRE (Site Reliability Engineering), Perception, Behavior, and Data Science teams to translate their performance analysis needs into robust, self-service infrastructure. About You…
…Introduction), SRE (Site Reliability Engineering), Perception, Behavior, and Data Science teams to translate their performance analysis needs into robust, self-service infrastructure. About You…