Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Principal Engineer, Cluster Orchestration”. A match may be a passing mention rather than the job itself. Titles only.
105 roles across 118 listings · show every listing · page 2 of 5
…Engineers define intent, author precise specifications, and orchestrate AI agents that execute at speed and scale. AI is embedded throughout our development process, helping…
…Your deep knowledge of software engineering, experience with DevOps and SRE principles, and passion for new technologies will help shape the future of our…
…specific logs, and experience managing multi-tenant Kubernetes clusters at scale. Experience with workflow/pipeline orchestration tools (Airflow, dbt) and understanding of data modeling…
Senior Solutions Principal, Data Management Our field sales professionals rely on proactive technical support during the sales process – and our expert Systems Engineering team…
…expand the scope of what engineering teams can deliver Define new metrics and data-driven decision-making principles for long-term ML projects, connecting…
…Help us build the Apple experience on a global scale! Apple is looking for a Senior Software Engineer with distributed systems and orchestration experience…
Amazon Web Services is looking for Software Development Engineers to join our growing Amazon Elastic Container Services for Kubernetes (EKS) team. Amazon EKS is…
Join NVIDIA, a leader in digital imagery, PC gaming, and advanced computing, as a Principal Software Engineer for DGX Cloud. At NVIDIA, our innovation…
…As a Principal Core Infrastructure Engineer at OCI, the ideal candidate will have 8-10+ years of software engineering experience and: Design and develop…
…As a Staff Engineer on our Orchestration team, you will collaborate to help drive the technical vision for Lambda's managed orchestration services, including…
NVIDIA is looking for an experienced HPC DevOps Engineer to help us build the supercomputers and HPC clusters of the future. As a Senior…
…As the Principal Systems Software Engineer, you will serve as the visionary lead for Crusoe’s next-generation AI infrastructure. This is a role…
…Help operate and improve the Kubernetes and cloud infrastructure that underpins the engineering workflow, including cluster behavior, node-pool strategy, workload scheduling, access patterns…
…Kubernetes and container orchestration. HPC cluster management platforms including Slurm, PBS, or Bright Cluster Manager. High-performance networking technologies including RDMA and InfiniBand. MPI…
…clusters demand. You will act as a senior technical contributor and a trusted voice in architecture decisions, partnering closely with Principal Engineers and engineering…
In this exciting Principal Core Infrastructure Engineer role you will play a critical role in designing, implementing, and maintaining the infrastructure that supports our…
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And…
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And…
…We’re looking for a Staff+-level engineer who has built production platform systems at scale before — not just consumed them. You’ve designed…
…the foundational compute and cloud components that product engineering teams build on: our Kubernetes clusters, related cloud infrastructure and networking, and the Infrastructure-as…
…BS/MS in Electrical Engineering, Computer Engineering, Computer Science, or a related field, with 5+ years of relevant industry experience. Ways to stand out…
NVIDIA is seeking a principal-level software engineer to build the next generation of our Kubernetes platform. Our teams build foundational capabilities for self…
…Knowledge of distributed inference architectures, including tensor parallelism, pipeline parallelism, expert parallelism/MoE serving, disaggregated serving, routing, autoscaling, and cluster-level orchestration. Experience with…
…Your deep knowledge of software engineering, experience with DevOps and SRE principles, and passion for new technologies will help shape the future of our…
…This is a hands-on, close-to-the-metal role for a first-principles Linux engineer. You'll be the final escalation for the…