Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Senior Software Engineer, RL Post-Training Frameworks”. A match may be a passing mention rather than the job itself. Titles only.
42 roles across 50 listings · show every listing · page 2 of 2
…optimization libraries, or distributed computing frameworks with public upstream records. Solid experience in Agentic AI, RL post-training or long-context LLM workload optimization…
…or defect/quality analysis. - Familiarity with modern training/inference infrastructure (e.g., distributed training, RL frameworks, model serving). Amazon is an equal opportunities employer…
…scale distributed AI training, post-training (RLHF, alignment), inference, or robotics workloads (multi-node, multi-GPU) Experience driving framework or software bring-up on…
…Experience building internal ML platforms or research clusters at a company doing large-scale training Familiarity with agentic AI: RL training with rollouts, agent…
…team - Manage pre/post training runs and continue improve system stability and throughput - Prototype new acceleration approaches using emerging compilation frameworks - 5+ years of…
…engineering models on SpaceX data. This will consist of training models from scratch, fine tuning models, and using Reinforcement Learning (RL) to post train…
…Experience with DL and RL algorithms and frameworks such as PyTorch. Enjoy working with multiple levels and teams across organizations (engineering/research, product, sales…
…for training software on Trainium. This includes distributed training libraries, post-training workflows (RLHF, DPO, fine-tuning), reinforcement learning frameworks, and training performance optimization…
…Strong engineering and R&D experience in LLM post-training, advanced RL-based methods to improve LLM models’ safety and quality using RLHF/RLAIF…
…record in post-training large-scale models, CPT, SFT, RL Hands-on experience with production ML pipelines, including dataset creation, training frameworks, and metrics…
…Fine-tune, and deploy large VLMs and LLMs, utilizing prompt optimization and advanced post-training techniques (SFT, RL, etc.), to solve complex, open-set…
…and senior engineers to implement and improve workflows for LLM pretraining, fine-tuning, and reinforcement learning-based post-training. This includes building training pipelines…
…You'll work at the intersection of cutting-edge research and production engineering, implementing, scaling, and improving post-training techniques like Constitutional AI, RLHF…
…Senior Machine Learning Engineer, you will join end-to-end development of large language models and agentic systems, from training pipelines to evaluation frameworks…
…This position centers on developing a Closed-Loop Simulation-based Reinforcement Learning (RL) framework in order to train advanced end-to-end AV models…
…in computer architecture - Previous software engineering expertise with Pytorch/Jax/Tensorflow, Distributed libraries and Frameworks, End-to-end Model Training. Amazon is an equal…
…About Data Engine Our Generative AI Data Engine powers the world’s most advanced LLMs and generative models through world-class RLHF (Reinforcement Learning…