Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Senior Software Engineer, RL Post-Training Frameworks”. A match may be a passing mention rather than the job itself. Titles only.
14 roles across 15 listings · show every listing
…team - Manage pre/post training runs and continue improve system stability and throughput - Prototype new acceleration approaches using emerging compilation frameworks - 5+ years of…
…engineering models on SpaceX data. This will consist of training models from scratch, fine tuning models, and using Reinforcement Learning (RL) to post train…
…Analyze complex hardware-software interactions to identify and resolve performance bottlenecks in both training and inference pipelines. Collaborate closely with AI researchers, HW and…
…Experience with DL and RL algorithms and frameworks such as PyTorch. Enjoy working with multiple levels and teams across organizations (engineering/research, product, sales…
…for training software on Trainium. This includes distributed training libraries, post-training workflows (RLHF, DPO, fine-tuning), reinforcement learning frameworks, and training performance optimization…
…Strong engineering and R&D experience in LLM post-training, advanced RL-based methods to improve LLM models’ safety and quality using RLHF/RLAIF…
…record in post-training large-scale models, CPT, SFT, RL Hands-on experience with production ML pipelines, including dataset creation, training frameworks, and metrics…
…scale distributed AI training, post-training (RLHF, alignment), inference, or robotics workloads (multi-node, multi-GPU) Experience driving framework or software bring-up on…
…Fine-tune, and deploy large VLMs and LLMs, utilizing prompt optimization and advanced post-training techniques (SFT, RL, etc.), to solve complex, open-set…
…and senior engineers to implement and improve workflows for LLM pretraining, fine-tuning, and reinforcement learning-based post-training. This includes building training pipelines…
…You'll work at the intersection of cutting-edge research and production engineering, implementing, scaling, and improving post-training techniques like Constitutional AI, RLHF…
…Senior Machine Learning Engineer, you will join end-to-end development of large language models and agentic systems, from training pipelines to evaluation frameworks…
…This position centers on developing a Closed-Loop Simulation-based Reinforcement Learning (RL) framework in order to train advanced end-to-end AV models…
…in computer architecture - Previous software engineering expertise with Pytorch/Jax/Tensorflow, Distributed libraries and Frameworks, End-to-end Model Training. Amazon is an equal…