Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Senior Product Manager, Infrastructure Observability | US | Remote”. A match may be a passing mention rather than the job itself. Titles only.
45 roles across 46 listings · show every listing · page 1 of 2
…Build secure, low-latency Go and gRPC services used across Coinbase's web, mobile, API, and partner experiences, with strong observability, failure testing, and…
…models used to run Tests and process their results - Improve the reliability, correctness, performance, and observability of the Tests runtime - Investigate complex production issues…
…take capabilities that exist in one product (access control, multi-cluster management, routing, catalog management, observability) and deliver them as a single implementation that…
…modular, granular, observable, and efficient. Working closely with data scientists, product managers and other XFN partners to build business insights, develop product features and…
…dbt, and Infrastructure as Code (IaC). Experience supporting machine learning platforms, LLM/AI access, or production ML services. Experience building or managing data contracts…
…observability Terraform at architectural scale — module design patterns, provider internals, remote state management, and multi-region provisioning Production Go or Python for infrastructure tooling…
…observability Terraform at architectural scale — module design patterns, provider internals, remote state management, and multi-region provisioning Production Go or Python for infrastructure tooling…
…Represent the team's capacity, priorities, risks, and delivery status to senior engineering leadership. What you’ll bring Engineering management experience leading remote, distributed…
…As a senior engineer on the ML Infrastructure and Platform group, you will help architect and build the foundational infrastructure for AI and machine…
…Collaborate closely with product managers, data scientists, and infrastructure engineers to deeply understand business needs and create impactful ML applications. Push the envelope on…
…and/or Senior Manager, Software Engineering Desire to grow in to a Tech-Lead-Manager role, with responsibility for line management of engineers, in…
…Experience using metrics, logs, traces, and observability tooling to understand production systems and make data-driven decisions. Clear communication skills and the ability to…
…pushing us toward a more intelligent, automated infrastructure future. WHAT YOU'LL DO - Architect and manage scalable, reliable cloud infrastructure across our production and…
…Partner closely with data scientists, ML engineers, and product engineers to productionize statistical and AI/ML capabilities into reliable, real-time backend infrastructure. Lead…
…optimization of production infrastructure. Hands-on experience operating distributed systems at scale, including diagnosing resource bottlenecks, tuning service performance, and managing capacity across heterogeneous…
…management, performance, reliability, and scalability problems, using data and benchmarking to guide decisions. Partner with engineers and teams across GitLab, including Git, Infrastructure, Site…
…You will work closely with Product Managers, AI Specialists, Data Analysts, and other Engineers on the team, and directly with customer executive sponsors and…
…You hold strong opinions from the user perspective, listen to peers, and negotiate trade-offs with product managers and users. Experience leading and supporting…
…We use Terraform to manage infrastructure, deploy containerized services on AWS, and rely on Postgres and modern observability tooling in production. Key Responsibilities Build…
…We use Terraform to manage infrastructure, deploy containerized services on AWS, and rely on Postgres and modern observability tooling in production. Key Responsibilities Lead…
…management, governance Establish company-wide standards for code quality, testing, MLOps (CI/CD), experimentation, model lifecycle management, and observability Lead adoption and advanced use…
…assistive and autonomous AI-native products to usher customers (IT departments using our IT service and operations management applications) into an agentic-first era…
…AI observability tooling (tracing, cost tracking, LLM-specific monitoring) Familiarity with cloud-native infrastructure, service observability, logging, monitoring, reliability engineering, and production troubleshooting Additional…
…Evaluate emerging technologies and advise management on industry trends, infrastructure risks, roadmaps, and investment decisions. Use telemetry, incident trends, and operational feedback to improve…
…This role exists to be the senior technical voice for Komodo's cloud infrastructure and shared services. Several of these platforms transferred to Infrastructure…