Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Senior Software Engineer, Deep Learning Inference - TensorRT”. A match may be a passing mention rather than the job itself. Titles only.
15 roles
…demonstrating strong foundations in computer science and machine learning. - 2+ years of professional software engineering experience, with demonstrated experience in distributed systems, backend platforms…
…or certification on software engineering concepts and 7+ years applied experience Deep, hands-on experience with LLM inference systems — vLLM, TensorRT-LLM, SGLang, LLM…
…and production inference. The ideal candidate is a senior technical DevRel leader who can earn credibility with ML researchers and platform engineers. This person…
…Strong knowledge of machine learning, deep learning, generative AI, real-time inference, data engineering, MLOps, and cloud-native systems. Experience with regulated environments, high…
…Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon’s custom machine learning accelerators, Inferentia and Trainium. The…
SENIOR SOFTWARE ENGINEER, MACHINE LEARNING INFRASTRUCTURE - GENERATIVE AI ABOUT THE TEAM Deliveroo's GenAI Platform team sits within Machine Learning Platform and builds the…
…Make Wayve the experience that defines your career! The role We’re looking for a Senior Machine Learning Engineer to join a high-ownership…
…Deep expertise in performance internals and execution graphs of major deep learning training and inference frameworks (e.g., PyTorch, JAX, TensorRT, vLLM, sgLang, Nemo…
…Experience with Inference deployment and optimization software (ex. vLLM, SGLang, FlashInfer, TensorRT-LLM, Triton, Dynamo, TorchAO, etc.) Demonstrable knowledge of GenAI or machine learning…
…Accelerate distributed inference using NVIDIA technologies such as NIM, TensorRT-LLM, vLLM, and SGLang. Collaborate with business, engineering, and product teams while providing technical…
…Software Development Engineer to join the AWS Mantle team and drive the technical vision for our distributed inference engine that serves millions of customers…
NVIDIA is seeking a Senior Deep Learning Algorithms Engineer to advance Dynamo, our open-source distributed inference platform for large-scale, low-latency AI…
…Deep familiarity with NVIDIA's AI software stack, including CUDA, TensorRT-LLM, NeMo, RAPIDS, Dynamo, Triton, or Isaac. Experience driving the implementation of AI…
…We are seeking a high-caliber Machine Learning Engineer to help us build, optimize, and deploy enterprise-scale AI solutions. Working closely with senior…
…A love of research publications in the machine learning and software engineering communities Effective communicator with experience collaborating cross-functionally with other teams For…