Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Compiler Engineer - AI Inference”. A match may be a passing mention rather than the job itself. Titles only.
53 roles across 58 listings · show every listing · page 1 of 3
…Job Summary Etched is building a new category of AI hardware: frontier inference clusters. As a top-level PD engineer, you will own full…
…Job Summary Etched is building a new category of AI hardware: frontier inference clusters. As a Physical Design Methodology Engineer, you will architect our…
…Job Summary Etched is building a new category of AI hardware: frontier inference clusters. As a Physical Design Engineer, you will own block-level…
…Our group is seeking an Engineering Manager to lead the Performance Tools and Services team, with a focus on the tools, services, and infrastructure…
…In-depth knowledge of ML converters/compilers and runtimes, and hardware-accelerated ML inference techniques. Strong understanding of generative AI model architectures and their…
…the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is…
…The profiler plays a crucial role to internal and external customers in optimizing AI workloads across hardware platforms such as Trainium and Inferentia devices…
…As an SDM for the LLM Inference Model Enablement team, you will lead a team of expert AI/ML engineers to onboard and optimize…
…You will develop model inference and fine-tuning services, onboard and benchmark new accelerators, and work closely with foundation model researchers and engineers to…
…Debug intricate integration issues between system applications, AI inference engines, edge hardware, the factory network, and physical test fixtures during the initial line ramp…
…We are looking for someone who understands compilers and can operate at the intersection of systems architecture, framework engineering, and customer-facing product strategy…
…The work sits at the intersection of distributed systems, AI inference, compilers and runtimes, performance engineering, security, and external partnerships. About the Role We…
…Engineering and 1+ years of industry experience 3+ years experience working on compilers for parallel architectures 1+ years experience working with ML inference or…
…in-class Qualcomm AI inference accelerators for data center, and hybrid AI applications. Minimum Qualifications: • Bachelor's degree in Engineering, Information Systems, Computer Science…
…rated AI-cloud for high-performance GPU infrastructure across AI/ML, visual effects, rendering, and real-time inference. Our stack is engineered for speed…
…These tools give kernel developers, compiler and runtime teams, architects, and design verification engineers visibility into both correctness and performance. Your work will help…
…GitHub) Experience with PyTorch, TensorFlow or similar machine learning toolsets Experience or knowledge of training/inference of Large scale AI models - CV and/or…
…world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10…
…Drive the strategy, roadmap, and execution of NVIDIA’s inference frameworks engineering, focusing on Client AI. Partner with internal compiler, libraries, and research teams…
…The emphasis is on AI Data Center infrastructure software, accelerator enablement, workload orchestration, inferencing optimization, workload deployment, platform readiness, and regional product positioning for…
…Experience in machine learning (ML) infrastructure development or ML performance engineering. Preferred qualifications: Experience with ML compilers and their internals, experience writing compiler optimization…
…The profiler plays a crucial role to internal and external customers in optimizing AI workloads across hardware platforms such as Trainium and Inferentia devices…
…GPU performance work — CUDA/Triton kernels, torch.compile, operator fusion, quantization — and interest in inference-efficiency domains such as AI-RAN. Experience benchmarking AI…
…model training and inference. Solid software engineering fundamentals and system architecture thinking, with the ability to build modules and drive engineering practices in complex…
…Deep Learning Training and Inference Frameworks, with the goal of supporting NVIDIA's top AI researchers and software engineers in driving the future of…