Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Sr. Staff Technical Program Manager - Reliability”. A match may be a passing mention rather than the job itself. Titles only.
223 roles across 266 listings · show every listing · page 3 of 9
…We are looking for a Senior Software Development Manager to lead EMR's Workload Management organization. In this role, you will own the technical…
…serve people and organizations operating in locations without reliable connectivity. As an Amazon Leo Satellite Sr. RF Test Systems Development Engineer, you will own…
…We are looking for a Senior Software Development Engineer to be a technical leader and force multiplier: you will partner with the engineering manager…
…The Federal SRE Team We are looking for an experienced Staff Site Reliability Engineer to join Okta's Federal SRE team for the Emerging…
…We are seeking a Sr. Staff Technical Program Manager (L7) to lead a high-impact internal incubator within our Corporate Engineering organization. Operating as…
…The work also connects to other parts of Datadog, including observability, Cloud Cost Management, permissions, and AI-assisted workflows. As a Staff Product Designer…
…years of programming with at least one software programming language experience - 5+ years of leading design or architecture (design patterns, reliability and scaling) of…
…ENA Express is a next-generation EC2 networking feature that leverages the Scalable Reliable Datagram (SRD) protocol — developed by AWS — to deliver industry-leading…
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google…
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google…
…Key job responsibilities * Deliver high-performance models using distributed inference libraries * Drive technical excellence in performance optimization and system reliability across the Neuron ecosystem…
…About the team The Hardware Engineering AI/ML UltraServer platform team is a group of engineers and technical program managers directly responsible for launching…
…ENA Express is a next-generation EC2 networking feature that leverages the Scalable Reliable Datagram (SRD) protocol — developed by AWS — to deliver industry-leading…
…Manage and mentor a 7-member engineering team (including L5 Senior and L6 Staff Engineers), fostering a culture of technical excellence, operational rigor, and…
Amazon's Reliability & Maintenance Engineering (RME) organization is looking for a Product Manager - Technical III to own the product lifecycle, roadmap, and stakeholder alignment…
…QUALIFICATIONS: * 10+ years of experience in delivering and managing highly scalable and highly available distributed systems. * Strong knowledge of JAVA and object oriented programming…
…Provide technical leadership on high-impact projects. Manage project priorities, deadlines, and deliverables. Facilitate alignment and clarity across teams on goals, outcomes, and timelines…
…Our platform runs on GCP, focused on multi-tenant GKE clusters and managed cloud services. We manage cloud infrastructure with Terraform and generate Kubernetes…
…years of programming with at least one software programming language experience - 5+ years of leading design or architecture (design patterns, reliability and scaling) of…
…to manage device traceability and key quality/test metrics - Represent Leo Supplier Quality in cross-functional initiatives - Travel as needed to accomplish program objectives…
…network programming coordination/management, and availability, improving mechanisms. Work closely with other UTE team members, test engineers and Site Reliability Engineers (SREs) to ship…
…Role We are looking for a Sr. Staff Software Engineer (Full-Stack) to join our team. This is a Hybrid (San Jose, CA) role…
…intelligently route traffic based on reliability, latency, quality, compliance, and cost. - Develop systems for model provisioning, capacity management, failover, and traffic engineering across multiple…
…critical systems. - Architect for scale and reliability. Partner with tech leads and staff engineers to drive sound technical decisions on a polyglot stack (TypeScript…
…As a Staff Product Manager on our Cloud Infrastructure team, you will own the networking product roadmap and shape how customers experience performance, reliability…