Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Development Engineer, AWS Incident Tooling & Response”. A match may be a passing mention rather than the job itself. Titles only.
1,331 roles across 1,606 listings · show every listing · page 2 of 54
…the SDLC/TLM toolchain Applies knowledge of tools within the Software Development Life Cycle toolchain, including approved AI-assisted development and automation capabilities, to…
…lead design/code reviews and mentor engineers Hands-on experience using enterprise-authorized AI-assisted software development tools within the work environment (e.g…
…Applies knowledge of tools within the Software Development Life Cycle toolchain, including enterprise-authorized AI-assisted development and automation capabilities, to improve the value…
…Job responsibilities Guide and assist others in building appropriate level designs and gaining consensus from peers. Collaborate with software and infrastructure engineers to design…
…product and engineering stakeholders to advance key business outcomes. Job Responsibilities Design and deliver creative software solutions through hands-on development and technical troubleshooting…
…Software Engineer 3 – Cloud Automation & AI Systems Oracle Cloud Infrastructure (OCI) is building the future of cloud operations through Incident Response Center as a…
…engineers and raise the team's bar for technical rigor and system design. Own operational excellence for the experience platform - monitoring, observability, incident response…
…enabling developer velocity Operate and tune Cloud Security Posture Management (CSPM) tooling and coordinate remediation through engineering teams Investigate security events , triage incidents, identify…
…engineers best practices and hard-earned lessons from your time as a developer. Give feedback to other teams on designs for tools and software…
…software vs purely operational or governance security responsibilities. What We Are Looking For Technical Expertise in the security of cloud platforms (e.g., AWS…
…Drive reliability and performance under real operational load through observability, on-call operations, incident response, runbook development, and automated remediation. Lead maintenance and enhancement…
Role Overview As a Senior Software Engineer, AI Ops at Scale AI, you will own the long-term technical health, performance, and stability of…
…operations incident response OR Bachelor's Degree in Statistics, Mathematics, Computer Science, or related field AND 4+ years experience in software development lifecycle, large…
…Design, develop, and deliver software to dramatically improve the availability, scalability, latency, and efficiency of Klaviyo's asynchronous and queueing services. Design and develop…
…We believe in providing an exceptional developer experience that enables our engineering teams to build and ship high-quality software efficiently and with confidence…
…Front is building a serious amount of software for itself. GTM Engineers are shipping internal tools and AI agents on top of Salesforce, Gong…
…raise engineering standards: design reviews, reliability practices (SLOs, metrics, traces, alerts, runbooks, incident response), security posture, and the disciplined use of AI tools to…
…aggregate impact, while remaining responsible for developing, building, and launching solutions optimized for unique operating environments. The GTM Sales Engineer & AI solutions is the…
…6+ years of software engineering experience. Solid software engineering breadth across deployment and CI/CD, automated testing, and production reliability (observability, incident response). Experience…
…and post-incident improvement activities. Develop enrichment and automation workflows using Python, APIs, and security tooling to improve analyst efficiency and response consistency. Improve…
Join the team with a mission to drive higher availability for every AWS customer! As a Senior Software Development Engineer on the AWS Incident…
…the SDLC/TLM toolchain Applies knowledge of tools within the Software Development Life Cycle toolchain, including approved AI-assisted development and automation capabilities, to…
…integration and continuous delivery pipelines Collaborates with other software engineers and teams to design, develop, test, and implement availability, reliability, scalability, and solutions in…
…including leading post-incident analysis and resilience improvements. Deep expertise in public cloud platforms (AWS or equivalent), infrastructure automation tools (CloudFormation, Terraform), and capacity…
…Job responsibilities Executes standard software solutions, design, development, and technical troubleshooting Drives team adoption of enterprise-authorized AI-assisted engineering practices within the work…