Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Site Reliability Engineer, Observability”. A match may be a passing mention rather than the job itself. Titles only.
279 roles across 315 listings · show every listing · page 1 of 12
…Senior Site Reliability Engineer, Observability Engineer, Incident Management Engineer For positions that will be based in NY, the annual salary range for this position…
…A background in site reliability or observability engineering, including monitoring, alerting, SLOs, SLIs, runbooks, incident handling, and tools such as Prometheus, Grafana, and OpenTelemetry…
…The Global Information Security Office (GISO) at Everpure is seeking an IAM Site Reliability Engineer to operate, automate, and continuously improve our enterprise Identity…
…As a Lead Software Engineer - Site Reliability Engineer, Operations Excellence at JPMorganChase within the AI/ML & Data Platforms area , you will serve as a…
…Leads a team of Site Reliability Engineers and support critical application 24x7 Implements Site Reliability Engineering (SRE) best practices to ensure reliability, scalability, and…
…in related disciplines such as Software Engineering, Product Security, Application Security, Detection Engineering, Site Reliability Engineering, Security Engineering, or IT Infrastructure. About OpenAI OpenAI…
…You will lead the configuration, standardization, and continuous improvement of our NX and PLM ecosystem, ensuring that designers and engineers have reliable, performant, and…
…We are safe and reliable, backed by our Proof of Reserves. Across our multiple offices globally, we are united by our core principles: We…
…Cloud (AWS or GCP) plus resilience engineering : infrastructure-as-code (Terraform), CI/CD, HA, DR, SLOs, and observability, backed by automated testing and documentation…
…building the features, systems, and infrastructure that keep the site fast, reliable, and effective at scale. The ideal candidate is a full-stack engineer…
…You'll partner closely with AI Research, Product Engineering, Infrastructure, and external model providers to build a platform that is highly reliable, scalable, observable…
…functional off-sites or customer meetings. . See yourself at Twilio Join the team as Twilio’s next Staff, Business Intelligence Engineer, GTM Data Science…
…Experience with platform engineering, API gateways, service meshes, or developer platforms. Background in security engineering, application security, site reliability engineering, or incident response. Experience…
…Your work directly enables the engineering and operations teams that build every Apple product. Own the reliability, performance, and scalability of services and infrastructure…
…Zipline Ground Systems Teams build the equipment that lands, charges, and conditions autonomous aircraft at shipper sites. This work enables safe, reliable delivery in…
…Be accountable, with team and engineering leaders, for the long-term technical health of the systems in your scope across reliability, scalability, performance, cost…
Meta is seeking a Data Center Production Operations Engineer to support the reliability, efficiency, and scalability of our global data center infrastructure. In this…
…AI a reliable, scalable force multiplier across the entire product development lifecycle. Our products include the LLM Proxy, AI Key Management, Observability pipelines, the…
…execution, reliable data capture, accurate configuration records, and evidence engineers use for design and vehicle-readiness decisions. This role is based on site in…
…ABOUT THE JOB We are seeking a Manufacturing Optimization Manager to drive continuous improvement and operational excellence across our manufacturing sites in the US…
…mentor senior engineers. - Experience operating highly available production services with strong reliability and operational excellence. - Experience building platforms that require scalability, observability, automation, and…
…8+ years of experience in customer facing technical roles such as Solutions Engineering, DevOps, Site Reliability, or ML Infrastructure Engineering, ideally supporting large‑scale…
…Improve reliability, performance, and observability of critical infrastructure services. Lead architectural decisions for next-generation CAD infrastructure platforms. Mentor engineers and elevate engineering standards…
…New Deployments - Own the infrastructure deployment for new sites or site expansions end-to-end: chip vendor and OEM dependencies, architecture updates, cloud foundations…
…observability for GPU utilization and reliability, automation of cluster provisioning, cost optimization, and vendor relationships You may be a good fit if you have…