Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Senior Performance Engineer - LLM Inference Frameworks”. A match may be a passing mention rather than the job itself. Titles only.
8 roles
…Senior or Staff Infrastructure Engineer to act as a primary technical lead, engineering the 'paved road' for our knowledge retrieval and inference engines. You…
…platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. The Together Cloud…
…of DNN frameworks, or deep learning training and inference workloads. Experience in evaluating, analyzing, and optimizing LLM training and inference performance of state-of…
…accelerated computing, distributed inference, deep learning frameworks (PyTorch, TensorFlow, JAX), and inference-specific frameworks & optimizations (Dynamo, Triton Inference Server, TensorRT-LLM, vLLM, SGLang) Market…
P-1377 Mission As a Senior ML and AI Technical Solutions Engineer, you play a critical role by helping customers debug and maintain stable…
…8+ years of hands-on validated ML/DL performance engineering experience with focus on improving GPU compute efficiency of training and inferencing workloads. Experience…
…tuning LLMs using popular frameworks such as TensorFlow, PyTorch, or Hugging Face Transformers. Proficiency in model deployment and optimization techniques for efficient inference on…
…We’re hiring Staff and Senior Software Engineers to work on one of the teams that powers the entire company: Infrastructure, Platform, and Leverage…