Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Member of Technical Staff - GPU Performance Engineer”. A match may be a passing mention rather than the job itself. Titles only.
129 roles across 138 listings · show every listing · page 4 of 6
…Member of Technical Staff, IT Senior Applications Engineer, or a related occupation. - Required Skills - Infrastructure-as-Code and deployment automation: Terraform, AWS CloudFormation, AWS…
…reliability, performance, and cost-effectiveness of infrastructure to support large-scale AI and application workloads in secure, classified settings. Collaborate with SpaceXAI engineers to…
…Solve challenging technical problems across a wide range of modern technologies. Apply a software engineering mindset to automate operations and improve system reliability, scalability…
…the next generation of AI platforms powering advanced NLP applications? We are looking for a Lead Member of Technical Staff to join the Model…
…engines (vLLM or TensorRT). - Familiarity with Nvidia GPU architecture and CUDA. - Experience with ML performance engineering (tell us a story about boosting GPU performance…
…Innovate on algorithms, modeling approaches, hardware/software/algorithm co-design, and scaling paradigms for state-of-the-art performance. Build research tooling, user-friendly…
…training libraries - Background in HPC environments, parallel computing, and high-performance networking - Knowledge of infrastructure as code (Terraform, Ansible) and GitOps practices - Experience with…
…Kubernetes, GPU scheduling, autoscaling inference workloads. QUALIFICATIONS - 3+ years of professional software engineering experience with meaningful work on ML inference or high-performance systems…
…scale. - Build high-performance inference platforms capable of serving and evaluating models across thousands of GPUs. - Optimize throughput, latency, and GPU utilization for large…
…production-ready training systems. - Improve performance of distributed training workloads through optimization of communication, memory usage, and GPU utilization. - Build and maintain training pipelines…
…to maintain strong engineering standards and product quality. Collaborate effectively with cross-functional stakeholders and project team members to align technical execution with program…
Overview Microsoft AI is looking for a Member of Technical Staff – Capacity & Efficiency Infrastructure , to help us improve manage, and improve the efficiency of…
…petabyte-scale data replication, and GPU-to-GPU network performance. WHAT WE’RE LOOKING FOR • Systems-level engineering experience with a focus on cluster…
…deeply about performance, numerical stability, and reproducibility. - You thrive in high-agency environments and enjoy solving hard technical problems. WHAT WE OFFER: We believe…
…As a member of the UC organization, you’ll support the development and management of Compute, Database, Storage, Platform, and Productivity Apps services in…
…engineering and distributed systems fundamentals - Experience training large models in multi-node GPU environments - Deep understanding of parallelism strategies and performance trade-offs - Experience…
…of GPU execution constraints and memory trade-offs - Experience debugging performance issues in production ML systems - Ability to reason about system-level trade-offs…
…As a member of the UC organization, you’ll support the development and management of Compute, Database, Storage, Platform, and Productivity Apps services in…
…of groundbreaking GPU compute clusters that run demanding deep learning, high performance computing, and computationally intensive workloads. We seek engineers with deep technical expertise…
…member develop into a better-rounded engineer and enable them to take on more complex tasks in the future. - 5+ years of engineering team…
…Proven track record of managing significant technical budgets and scaling high-performance engineering and research teams. Experience deploying models on resource-constrained edge devices…
…You will work directly with the technical lead on problems that require deep understanding of both ML architectures and hardware constraints. This is high…
…Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts…
…Responsibilities include: - Identifying architectural changes to improve reliability and performance. - Fostering a culture of reliability across Modal’s engineering organization. - Defining and implementing operational…
…build the next generation of AI platforms powering advanced NLP applications? We are looking for Members of Technical Staff to join the Model Serving…