Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Compiler Engineer - AI Inference”. A match may be a passing mention rather than the job itself. Titles only.
16 roles
…You will join a dynamic team building and applying AI agents to simplify and accelerate customer adoption of Trainium and Inferentia. As an SDE…
…This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch, enabling unparalleled ML inference…
…We are the first inference-focused frontier AI system. Our addressable market is the entirety of inference, unlike many of our competitors. We are…
…PhD in Computer Science, Engineering, or a related field. - Comfort navigating the full AI toolchain: Python modeling code, compiler IRs, performance profiling, etc. - Strong…
…scale AI training and inference. Your work will range from prototyping system software on new accelerators to enabling performance optimizations across our AI workloads…
…Solid experience in large AI job performance analysis for training/inference workload Knowledge of Linux device drivers and/or compiler implementation Knowledge of GPU…
…We are the first inference-focused frontier AI system. Our addressable market is the entirety of inference, unlike many of our competitors. We are…
…This is an opportunity to work with top engineers, researchers, and partners across NVIDIA and leave a mark on the way generative AI reaches…
…The Distributed Training team works side by side with chip architects, compiler engineers and runtime engineers to create, build and tune distributed training solutions…
…Electrical Engineering, or a related field - Experience optimizing large models for training and inference (LLMs, VLMs, or video models) - Knowledge of compiler stacks or…
…design. - Skills in Design Compiler, Fusion Compiler, ICC2 or similar physical design tools. - BS or MS in Electrical Engineering. Preferred: - Knowledge of CPU/GPU…
…bottleneck analysis, co-optimize inference performance for GPUs, TPUs, or custom accelerators. Work closely with AI researchers and infrastructure engineers to develop efficient model…
…8+ years of hands-on validated ML/DL performance engineering experience with focus on improving GPU compute efficiency of training and inferencing workloads. Experience…
…Position Overview We are looking for a Software Engineer to work at the forefront of deploying our cutting-edge AI models, enhancing the performance…
…Building and operating ML inference services in production Designing scalable API architectures with async processing Optimizing GPU workloads (batching, quantization, compilation, CUDA) Managing distributed…
About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance…