Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Senior Software Engineer, ML Compiler, Frameworks and Performance”. A match may be a passing mention rather than the job itself. Titles only.
118 roles across 128 listings · show every listing · page 4 of 5
…Partner with PMs responsible for compiler, NKI, runtime, and infrastructure. Drive trade-offs between training performance, scalability, developer experience, and AI/ML ecosystem compatibility…
…Perform rigorous parity checking, accuracy recovery, and latency benchmarking between PyTorch frameworks and compiled edge binaries. Develop and optimize custom ML OPs and TensorRT…
…identify performance gaps relative to reference frameworks. Develop detailed performance plans based on profiling findings and collaborate with NVIDIA's kernel engineering and OSS…
…Familiarity with deep learning architectures and the latest LLM developments. Background with NVIDIA hardware and software, performance tuning, and error diagnostics. Hands-on experience…
…and software stack optimization to technical leadership with customer architects, engineering teams, and senior decision makers. You will engage directly with developers, ML engineers…
…architectures + ML compiler workload synthesis, a plus Prior working experience of hardware accelerators and hardware software co-design Experience of profiling software and optimization…
…Collaborate with cross-functional teams to ensure seamless integration of AI/ML components within the broader framework. Mentor and coach junior engineers, providing development…
…a software engineer who will work on all parts of the runtime stacks, supporting AI, ML, and scientific applications in high-performance distributed systems…
…engineers, ML SW engineers, and compiler engineers to evaluate architecture and microarchitecture tradeoffs and help make hardware design decisions - Contribute to cycle-approximate performance…
…functional models and emulators, to performance analysis on live silicon - Collaborate with chip architects, RTL designers, modelers, compiler engineers, and ML framework teams to…
…and analyze application/ML model performance for complex optimization opportunities, and prototype/generalize solutions at the application, compiler, or infrastructure level (firmware, runtime, framework…
…and Python to integrate ML capabilities seamlessly into Apple's ecosystem. As a senior technical contributor, you will set engineering standards, mentor engineers, and…
…FlashInfer, Flash Attention) Expertise in inference engines like vLLM and SGLang Expertise in machine learning compilers (e.g. Apache TVM, MLIR) Open source project…
…engines like vLLM and SGLang Expertise in machine learning compilers (e.g. Apache TVM, MLIR) Strong experience in GPU kernel development and performance optimizations…
…Develop Proofs of Readiness (PORs) and collaborate closely with our compiler team on Torch-TRT, MLIR-TRT, and related frameworks to bridge performance gaps…
…Experience with compiler technologies (e.g., MLIR, LLVM, XLA, Triton, etc.). Excellent C/C++ and Python programming and software design skills, including debugging, performance…
…inference High-performance computing (HPC) and collective communications ML systems, runtimes, or compilers Performance modeling, benchmarking, and systems analysis Hardware–software co-design for…
…Staff Software Engineer in the Qualcomm AI Stack SDK Software team, you will design, develop, and deliver advanced AI/ML software solutions for Generative…
…Contributions to opensource ML frameworks, compilers, or runtime systems, particularly in areas related to performance or scalability. Demonstrated research impact, such as publications or…
…software performance benchmarking, profiling, and optimizations. Background in compiler development Experience in working with TensorRT, PyTorch, TensorFlow, ONNX Runtime or other ML frameworks. NVIDIA…
…edge software stack, the AWS Neuron Software Development Kit (SDK), which includes an ML compiler, runtime and natively integrates into popular ML frameworks, such…
…edge software stack, the AWS Neuron Software Development Kit (SDK), which includes an ML compiler, runtime and natively integrates into popular ML frameworks, such…
…Familiarity with popular LLM frameworks and libraries such as TensorRT, TensorRT-LLM, vLLM, SGLang, MLC-LLM, or FlashInfer. A track record of strong software…
…Neuron is a Software that include ML compiler and native integration into popular ML frameworks. Our products are being used at scale with external…
…Neuron is a Software that include ML compiler and native integration into popular ML frameworks. Our products are being used at scale with external…