Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Performance Engineer (Inference, Training & GPU)”. A match may be a passing mention rather than the job itself. Titles only.
50 roles across 52 listings · show every listing · page 1 of 2
…HAVE… - Familiarity with GPU node orchestration & scheduling - You understand the intersection of GPU infrastructure, inference systems, Kubernetes, distributed systems, and performance optimization. FULL-TIME…
…leverage our large-scale GPU training and inference fleet through an observable, reliable and high-performance distributed AI/GPU communication stack. Currently, one of…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…Solve technically tests problems that exceed the scope of a generalist Software Engineers, specifically around optimizing Generative AI performance across heterogeneous hardware (CPUs, GPUs…
…AI training and inference? Want to do industry leading work delivering continuous price performance improvements in the cloud for AI model training for multi…
…such as Ray, Spark and in training/inference systems such as Ray, vllm/SGLang Solid grounding in engineering fundamentals and enterprise system design Preferred…
…Inferentia and Trainium Systems — delivering high-performance ML inference and training at cloud scale. We’re looking for a Senior Manager, Physical Design Engineering…
…GPU infrastructure, HPC, Security, Networking, or Storage - • Experience helping customers design and optimize infrastructure for AI/ML workloads (training clusters, inference optimization, GPU scheduling…
…Operator Development Parallel Computing AI Compiler Systems Large Language Model (LLM) Inference Optimization Reinforcement Learning Deep Learning End to End Training Performance Modelling, Analysis…
…for path generation Collecting training datasets and real-time inference run-times using simulators/gyms as well as performing in-vehicle tests Robotics: Embodied…
…Learning, GPU Computing, Accelerated Computing Validation Frameworks for Deep Learning, Deep Learning Frameworks and Libraries (NumPy, SciPy, cuBLAS, cuDNN) Data Preprocessing, Training Acceleration (CUDA…
…Experience building or supporting production AI/ML platforms (training, deployment, and model serving/inference), including GPU infrastructure/tooling. Strong DevOps/platform engineering practices: CI…
…kernels, runtimes, or performance engineering. - Strong systems programming fundamentals and experience writing performance-critical software. - Experience working with GPU, TPU, Trainium, or other specialized…
…We work closely with ML researchers and developers to optimize and scale out model training and inference. The team operates at the intersection of…
…industry-leading GPU fleet and AI Platform. Our core mission is to empower Google Cloud's most sophisticated training and inference customers by providing…
…GPU-accelerated compute clusters optimized for AI inference and training workloads at the tactical edge (NVIDIA H100/A100, AMD MI300, or similar) Engineer advanced…
…Experience with distributed training or HPC frameworks, inference serving, systems languages, high-performance networking, Unreal/client-server architecture, AI-assisted development tools Passion for…
…We’re hiring a PM to work on our core inference product and general platform, spanning inference performance and control, accounts and fraud reduction…
…with engineers, researchers, and external partners to troubleshoot issues, improve reliability, and optimize the performance of large-scale AI training and inference systems. • Build…
…with deep focus on GPU platforms, backend networks, and high-performance fabrics that power large-scale AI training and inference. This role will co…
…bare-metal performance requirements are the barrier to scaling AI training and inference workloads on CoreWeave. Translate customer requirements around GPU cluster topology, RDMA…
…Hands-on experience with next-generation AI accelerators beyond standard GPUs (e.g., AWS Trainium, Google TPUs, or custom ASICs) for training and inference…
…We are looking for engineers who treat fellow teammates with fairness, respect, and support. Our team maintains a high-performance AI inference platform that…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…HAVE… - Familiarity with GPU node orchestration & scheduling - You understand the intersection of GPU infrastructure, inference systems, Kubernetes, distributed systems, and performance optimization. FULL-TIME…