Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Senior Site Reliability Engineer - Managed Kubernetes”. A match may be a passing mention rather than the job itself. Titles only.
284 roles across 351 listings · show every listing · page 4 of 12
…in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure Strong hands-on experience operating AWS in production environments Deep expertise in Kubernetes, including…
…engineers; and as a Senior Engineer on the team, you will prioritize and tackle problems as they arise, working iteratively with product management, security…
…As a Lead Site Reliability Engineering at JPMorgan Chase within the Chief Technology Office, Identity & Access Management team, you are the non-functional requirement…
…site reliability, infrastructure, distributed systems, or production software engineering. - Have deep experience operating Kubernetes in production. - Understand Kubernetes architecture, scheduling, networking, resource management, upgrades…
…We welcome candidates from a variety of backgrounds, including software engineering, site reliability engineering, production engineering, infrastructure, and other roles focused on building reliable…
…Apple Services Site Reliability Engineering (SRE) teams are responsible for the systems and services that directly support those customers and their experiences. We are…
…Are you ready to be part of something outstanding? NVIDIA's Digital Marketing Organization seeks a senior Site Reliability Engineer (SRE) to join our…
…LLMOps, inference reliability, cost/performance management) Strong communication skills and the ability to set standards other engineers adopt Industry-recognised container / Kubernetes certification (e…
…Production on-call and incident management responsibilities. 12+ years of experience in Site Reliability Engineering, managing infrastructure and services at scale 5+ years of…
…experienced engineer to be a part of a team of Site Reliability Engineers. You will be working closely with engineering teams, product managers, as…
…site reliability improvements Guide and shape the future of technology at a globally recognized firm, driven by pride in ownership. As a Senior Manager…
…Escalation points for junior site reliability engineers during complex or high-impact incidents. Manage and execute complex manual Change Management tickets, by working closely…
…Leadership & Org Building Lead, grow, and mentor a team of Site Reliability Engineers, conducting regular 1:1s, performance reviews, and career development discussions Hire…
…to cloud on our core ground systems & Kubernetes infrastructure. ABOUT THE JOB As a Site Reliability Engineer on the Observability team, you will build…
…Kubernetes, deployment automation, cloud and on-premises infrastructure, developer environments, artifact management, and observability. You will help make the systems engineers depend on reliable…
…Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and…
…delivery and lifecycle management Strong understanding of site reliability engineering practices, including incident management, root-cause analysis, runbooks, and reliability patterns Practical knowledge of…
…SENIOR PRINCIPAL SOFTWARE ENGINEER – TS/SCI Based out of our headquarters in Long Beach, CA, we are seeking a highly skilled and experienced Senior…
…As a Senior SRE, you keep training and inference clusters reliable and fast, and you help redesign them for the next level of scale…
…large engineering organizations (500+ people) Experience leading without authority and influencing senior technical leaders Track record of successful change management in engineering contexts Demonstrated…
…Strong understanding of SRE principles, including SLIs, SLOs, error budgets, and incident management. Ability to manage highly technical team of Site Reliability Engineers. Experience…
…managing or developing within Kubernetes clusters and a deep understanding of container orchestration. - DevOps & Reliability Background: Proven experience in DevOps, Site Reliability Engineering (SRE…
…Strong expertise with orchestrating, scaling, and managing containerized applications in production environments using Kubernetes (K8s) and Terraform/Helm. Experience designing, building, and operating large…
…to guide architectural decisions and mentor senior engineers. - Experience operating highly available production services with strong reliability and operational excellence. - Experience building platforms that…
…Site Reliability, or ML Infrastructure Engineering, ideally supporting large‑scale cloud or service provider environments. Strong expertise in Linux systems, distributed computing, Kubernetes, containers…