Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Research Engineer, Code RL (Reinforcement Learning)”. A match may be a passing mention rather than the job itself. Titles only.
17 roles across 18 listings · show every listing
…build, and scale our Reinforcement Learning Environments (RLE) platform and the infrastructure that powers it - Partner closely with Engineering, Research, Product, and Operations teams…
…As an Applied Machine Learning Engineer, you will serve as a vital bridge between cutting-edge AI research and practical, real-world applications. Your…
…As an Applied Machine Learning Engineer, you will serve as a vital bridge between cutting-edge AI research and practical, real-world applications. Your…
…Our research team pushes the frontier of post-training and reinforcement learning. Our applied AI team sits side-by-side with customers as they…
…YOUR RESPONSIBILITIES We are looking for a Senior Research Scientist to lead fine-tuning, post-training, and reinforcement learning for the next generation of…
…YOUR RESPONSIBILITIES We are looking for a Senior Research Scientist to lead fine-tuning, post-training, model-steerability, and reinforcement learning for the next…
…other scientists, interns, engineers in the use of ML techniques. - Publishing innovation in research forums. - 6+ years of building machine learning models for business…
…As a Reinforcement Learning Engineer you will advance the frontier of AI Reinforcement Learning development and delivery. You will build the infrastructure and tooling…
…developed code - Publication record in top tier venues in generative AI reasoning or classical planning - PhD in a relevant field (reinforcement learning, neurosymbolic AI…
…Developing systems that enable models to use computers effectively Advancing code generation through reinforcement learning Pioneering fundamental RL research for large language models Building…
…DPO, RLHF, or KTO), reward model design and training, and reinforcement learning to improve output quality, controllability, and human preference adherence. Editing — Research in…
…Build continuous learning systems where Advisor improves from every interaction — leveraging reinforcement learning from human feedback (RLHF), outcome-driven reward signals, and retrieval-augmented…
…Background in coding models or code-generation systems, and/or research experience in reinforcement learning (RLHF, RLAIF, preference modeling) or related model-training techniques…
…research, machine learning, or a related quantitative field - Experience with optimization solvers (e.g., Gurobi, OR-Tools) or reinforcement learning frameworks (e.g., RLlib…
…NVIDIA is building an RL Frameworks engineering team to develop the open-source tools and infrastructure that AI researchers and post-training teams depend…
…You'll adopt and adapt state-of-the-art techniques — including supervised fine-tuning, reinforcement learning, preference optimization, and knowledge distillation — running rigorous experiments…
…agents, reinforcement learning, post-training workflows. Ability to work effectively in a fast-moving environment that blends research exploration with product and engineering rigor…