Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Compiler Engineer - AI Inference”. A match may be a passing mention rather than the job itself. Titles only.
296 roles across 342 listings · show every listing · page 2 of 12
…engineers who build the foundational software stack that unlocks the full performance potential of custom silicon for large-scale AI workloads. Design compiler architecture…
Meta designs and deploys its own AI systems. MTIA — the Meta Training and Inference Accelerator — is Meta's family of in-house AI accelerator…
…AI frameworks Familiarity with AI compilers, high-performance kernel development, or hardware enablement Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent…
…of AI and hardware optimization, we want you to join our team! As a Machine Learning Compiler Engineer on the Apple Neural Engine (ANE…
…registers, DMA, command queues, or similar accelerator interaction patterns Experience shipping production system software Experience building or extending ML runtimes, inference engines, or accelerator…
…help engineers debug, profile, measure, and monitor AI workloads running on MTIA hardware at scale. You will work at the intersection of compilers, runtime…
…and compilers. ONNX Runtime: ONNX based cross-platform, high performance ML inferencing and training accelerator. Foundry Local: an on-device AI inference solution offering…
…Maintain an in-house ML inference platform to serve large language models efficiently. Maintain an in-house ML compiler platform to compile, deploy, and…
…Maintain an in-house ML inference platform to serve large language models efficiently. Maintain an in-house ML compiler platform to compile, deploy, and…
…Operating at the intersection of AI research and infrastructure engineering, you will define the long-term strategic outlook and architectural roadmap for our future…
…MTIA (Meta Training & Inference Accelerator) Software team and part of the bigger AI and Compute Foundations team. The Graph Compiler team drives the development…
…This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch, enabling unparalleled ML inference…
…of applied scientists, system engineers, and product managers to deliver state-of-the-art inference capabilities for Generative AI applications. Your work will involve…
…Strong experience building AI agents or AI-backed engineering systems. Familiarity with modern model architectures and AI inference workloads. Experience evaluating AI models or…
…design, compilers and lower-level runtime features, information retrieval structures and algorithms, language tooling and static analysis, or interfaces between probabilistic inference and deterministic…
…The Inference Enablement and Acceleration team works side by side with compiler engineers and runtime engineers to create, build and tune distributed inference solutions…
NVIDIA is seeking a Senior MLOps Engineer to join our DSX Enablement team, collaborating closely with strategic customers to implement and enhance groundbreaking AI…
…Model compression, hardware-aware model optimizations, hardware accelerators architecture, GPU architecture, machine learning compilers, or ML systems, AI infrastructure, high-performance computing, performance optimizations…
…of AI Accelerators architectures Contribute to the development of the PyTorch AI framework core compilers to support new state of the art inference and…
…engineers actually use them. Manage the interface between silicon and its consumers: hardware systems for package/board/bring-up, and inference/kernels/compiler teams…
…of AI and hardware optimization, we want you to join our team! As a Machine Learning Compiler Engineer on the Apple Neural Engine (ANE…
…You'll collaborate with hardware teams, compiler engineers, and ML researchers to unlock capabilities that few organizations can deliver at Apple's scale and…
…Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About…
…You will join a dynamic team building and applying AI agents to simplify and accelerate customer adoption of Trainium and Inferentia. As a Sr…
…compiler engineering, or AI framework development Proficient in Python and C++ Solid understanding of ML compiler concepts (graph IRs, operator fusion, shape inference, lowering…