Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Compiler Engineer - AI Inference”. A match may be a passing mention rather than the job itself. Titles only.
296 roles across 342 listings · show every listing · page 4 of 12
…Inference Engineer, for GPU Kernel Optimization! What does it take to push every LLM inference operation to its performance ceiling? Our LLM Inference Performance…
…As a Qualcomm Datacenter AI Systems Engineer focused on AI Inference Platforms and Competitive Analysis, you will evaluate, benchmark, optimize, and validate end-to…
…engineer on our team, you will independently lead the SW driven early stages test development for our Trainium and Inferentia AI acceleration engines , drive…
…Nice to Have - Experience with compilers and exporters (torch.compile, TensorRT, ONNX, XLA). - Experience optimizing inference workloads for latency and throughput. - Triton compiler and…
Meta's Silicon Engineering organization is building custom silicon solutions that power the infrastructure underpinning Meta's AI and data center workloads at scale…
…engineering, performance engineering, ML systems engineering, infrastructure engineering, or related areas. - Hands-on experience with large-scale GPU or accelerated computing infrastructure for AI…
…Experience in scale-up networking in AI GPU/xPU systems. Experience in memory systems for AI inference/training infrastructure. Experience in building high-performance…
…including silicon engineering, hardware design and verification, software, and operations. Amazon Nitro, ENA, EFA, Graviton and F1 EC2 Instances, Amazon Neuron, Inferentia and Trainium…
…Working alongside experienced kernel, compiler, performance, and hardware engineers, you will learn how algorithms are mapped to specialized hardware and contribute to software that…
…You will partner across compiler, runtime, cloud infrastructure, hardware architecture, product management, and AI research teams while helping shape the future of AI inference…
…Job Summary Etched is building a new category of AI hardware: frontier inference clusters. As a top-level PD engineer, you will own full…
…Job Summary Etched is building a new category of AI hardware: frontier inference clusters. As a Physical Design Methodology Engineer, you will architect our…
…Our group is seeking an Engineering Manager to lead the Performance Tools and Services team, with a focus on the tools, services, and infrastructure…
…In-depth knowledge of ML converters/compilers and runtimes, and hardware-accelerated ML inference techniques. Strong understanding of generative AI model architectures and their…
…the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is…
…The profiler plays a crucial role to internal and external customers in optimizing AI workloads across hardware platforms such as Trainium and Inferentia devices…
…As an SDM for the LLM Inference Model Enablement team, you will lead a team of expert AI/ML engineers to onboard and optimize…
…You will develop model inference and fine-tuning services, onboard and benchmark new accelerators, and work closely with foundation model researchers and engineers to…
…Debug intricate integration issues between system applications, AI inference engines, edge hardware, the factory network, and physical test fixtures during the initial line ramp…
…We are looking for someone who understands compilers and can operate at the intersection of systems architecture, framework engineering, and customer-facing product strategy…
…The work sits at the intersection of distributed systems, AI inference, compilers and runtimes, performance engineering, security, and external partnerships. About the Role We…
…Engineering and 1+ years of industry experience 3+ years experience working on compilers for parallel architectures 1+ years experience working with ML inference or…
…in-class Qualcomm AI inference accelerators for data center, and hybrid AI applications. Minimum Qualifications: • Bachelor's degree in Engineering, Information Systems, Computer Science…
…rated AI-cloud for high-performance GPU infrastructure across AI/ML, visual effects, rendering, and real-time inference. Our stack is engineered for speed…
…These tools give kernel developers, compiler and runtime teams, architects, and design verification engineers visibility into both correctness and performance. Your work will help…