Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Site Reliability Engineer, Observability”. A match may be a passing mention rather than the job itself. Titles only.
417 roles across 464 listings · show every listing · page 1 of 17
…Senior Site Reliability Engineer, Observability Engineer, Incident Management Engineer For positions that will be based in NY, the annual salary range for this position…
We have a Lead Site Reliability Engineer (SRE) opportunity within our JPMC Google & Azure Cloud Site Reliability Engineering team. We have a Lead Site…
…Systematically incorporates external feedback (forums, live-site reports) into backlog refinement. Mentors junior engineers; coordinates cross-functionally on feature delivery. Software Development and Coding…
…e.g., online forums, live site maintenance reports) to refine application requirements. Collaborates with Designers, Product Managers, Software Engineers, and stakeholders when evaluating and…
…Infrastructure as Code (Terraform preferred) CI/CD pipelines Production monitoring and observability Metrics, logging, and operational dashboards Incident management Production readiness Site reliability and…
…field/customer, telemetry, and live-site signals. Institutionalizes feedback loops into planning cadences; measures impact. Mentors Staff engineers; partners with Product and Design to…
…dbt Labs pioneered analytics engineering, helping teams transform data into reliable, governed insights. Together, we support thousands of organizations as they build a trusted…
…ABOUT THE JOB We are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Costa Mesa, CA or…
…high degrees of freedom, dynamic constraints, environmental uncertainty, and reliability at scale. About The Work Own and drive the technical roadmap for collision-free…
…plan and execute complex hardware movements, site installations Site sustainment operations: Tier 1 troubleshooting and user engagements, issue resolution and escalation, reporting, working with…
…engineers, as well as data scientists across the company - Develop tools and frameworks for data ingestion, transformation, quality, and observability - Architect scalable, reliable systems…
…Ensure deep observability coverage by standardizing metrics, alerts, and distributed tracing across core data pipelines. CHAMPION ENGINEERING HEALTH: Advocate for a clean architectural foundation…
…Our work sits at the intersection of security, infrastructure, and software engineering. We build cloud-native platforms and services that are reliable, secure, and…
…part of the broader observability and data platform strategy. • Improves data quality and participates in data governance and reliability activities; develops and maintains data…
…The ideal candidate will bring strong experience in site reliability engineering, production operations, platform support, and automation , with the ability to manage day-to…
…services to ensure resilience, reliability, and availability at scale. Designs, governs, and signs off on advanced change activities and site augmentations, driving standardization of…
…on-call quality, response, post-mortems, and driving down incident count, time-to-detect, and time-to-resolve. - Automation & reliability engineering — automate low-frequency…
…As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale Plaid's reliability practices across product engineering. You'll architect…
…the workflow builder, execution engine, integrations, debugging tools, observability, evaluation systems, and feedback loops that help customers understand and improve what they deploy. This…
…measure performance and reliability across different configurations - Monitor and improve test coverage and reliability metrics - Collaborate with product and engineering teams to understand testing…
…This role is ideal for engineers who love building robust distributed systems, but who also want to run experiments, reason about tradeoffs in data…
…operating systems that process petabyte-scale datasets - Strong systems engineering skills, including reliability, observability, performance optimization, and debugging - Experience designing experiments and using data…
…adoption, sales engagement, and other conversion goals Platform Reliability & Performance: Own site performance, availability, observability, Core Web Vitals, CDN strategy, technical SEO, AI/search…
…We have strong teams across observability, release engineering, and incident management — and we're concentrating our reliability efforts into a dedicated SRE practice that…
…You will partner closely with software engineers, infrastructure teams, and cross-functional stakeholders to improve developer experience, platform reliability, deployment safety, observability, and operational…