Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Senior Software Engineer, ML Compiler, Frameworks and Performance”. A match may be a passing mention rather than the job itself. Titles only.
22 roles across 23 listings · show every listing
Overview The Microsoft AI Frameworks team develops the software, performance systems, and engineering tools that enable state-of-the-art AI models to run…
…frameworks (i.e. TensorFlow, PyTorch) is a plus. Experience with common compiler development practices and methodologies. Excitement about high-performance systems engineering and performance…
…emulation, performance modeling — and with hardware/software co-design cycles Familiarity with ML framework internals: PyTorch dispatch and eager execution, torch.compile / Inductor, custom…
…and validate tooling effectiveness and systems performance improvements Familiarity with ML framework internals (PyTorch graph execution, torch.compile, operator dispatch) and AI compiler stacks…
…Maintain an in-house ML inference platform to serve large language models efficiently. Maintain an in-house ML compiler platform to compile, deploy, and…
…Define and drive the technical roadmap and architecture for the hardware/software stack, ensuring unparalleled performance for the training and serving of large ML…
…an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch, enabling unparalleled ML inference and training performance. As…
…Strong software development using Python, System level programming and ML knowledge are both critical to this role. Our engineers collaborate across compiler, runtime, framework…
NVIDIA is seeking a Senior MLOps Engineer to join our DSX Enablement team, collaborating closely with strategic customers to implement and enhance groundbreaking AI…
…Responsibilities As a Senior Software Engineer: Design, implement, test, and operate production-quality components across AI frameworks, runtimes, benchmarking systems, performance tooling, and service…
…About the team The Edge AI ML Platform and Infrastructure team brings together software engineers, ML infrastructure engineers, and GPU performance specialists. We build…
…and performance. Background in deep learning compilers and ML systems, including graph-level and codegen tools (e.g., Triton, XLA, torch.compile) and highly…
…Profile and analyze AI/ML workloads across Apple Silicon compute engines (GPU, ANE, and CPU) to help identify performance bottlenecks, working alongside senior engineers…
…and MLIR dialects, conduct performance optimizations and analysis, implement compiler optimizations and kernel generation for neural networks, and contribute to other general software engineering…
…Work multi-functionally with compiler, architecture, and platform software teams to ensure performance achievements hit target expectations on schedules strictly linked to the CUDA…
…Proven experience leading and managing engineers including hiring, performance management, and technical mentorship of senior ICs and managers. Track record of shipping petabyte-scale…
…Senior/Staff AI Software Engineer to own developer productivity for AIMS: the build, test, and iteration loop that our ML researchers and engineers rely…
…Minimum Qualifications: • Bachelor's degree in Computer Science, Engineering, Information Systems, or related field and 2+ years of Hardware Engineering, Software Engineering, Systems Engineering…
…Our team works closely with compiler, kernel, hardware, and framework organizations across NVIDIA to surface bottlenecks and ship measurable gains. If driving GPU performance…
…You will tackle complex problems rooted across hardware and software domains, develop innovative validation methodologies, and set the standard for engineering excellence across our…
…fabric-level performance analysis. - Experience with inference serving frameworks, training frameworks, or ML compiler/runtime stacks. - Familiarity with both x86 and ARM-based server…
…Experience in building high-performance LLM inference systems using SGLang or vLLM. Publications in top computer architecture, systems, and/or ML conferences. Research Sciences…