Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Research Engineer, Code RL (Reinforcement Learning)”. A match may be a passing mention rather than the job itself. Titles only.
114 roles across 119 listings · show every listing · page 2 of 5
…Scale's coding data products, RL environments, and agentic coding evaluations. You will work across AI Product Management, ML Researchers, Engineering, Operations, and Go…
…reinforcement learning. This role sits at the intersection of research and large-scale systems engineering: a builder who understands both the algorithms behind RL…
…We're looking for individuals who have a background in reinforcement learning research, are able to iterate quickly, and who can convert scientific rigor…
…reinforcement learning, reward modeling, RLHF/RLAIF, or alignment techniques applied to generative models - Experience with real-time interactive systems — models that handle concurrent input…
…researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role Code RL at Anthropic drives reinforcement learning…
…for machine learning research or RL workflows, and familiarity with agentic systems or LLM training pipelines Experience building agent frameworks, orchestration engines, or multi…
…integrating inference engines (vLLM, SGLang), weight synchronization, and asynchronous/off-policy schemes. - Design RL environments for agentic, multi-step tasks — sandboxed code execution, tool…
…build, and scale our Reinforcement Learning Environments (RLE) platform and the infrastructure that powers it - Partner closely with Engineering, Research, Product, and Operations teams…
…As an Applied Machine Learning Engineer, you will serve as a vital bridge between cutting-edge AI research and practical, real-world applications. Your…
…As an Applied Machine Learning Engineer, you will serve as a vital bridge between cutting-edge AI research and practical, real-world applications. Your…
…Our research team pushes the frontier of post-training and reinforcement learning. Our applied AI team sits side-by-side with customers as they…
…YOUR RESPONSIBILITIES We are looking for a Senior Research Scientist to lead fine-tuning, post-training, and reinforcement learning for the next generation of…
…YOUR RESPONSIBILITIES We are looking for a Senior Research Scientist to lead fine-tuning, post-training, model-steerability, and reinforcement learning for the next…
…other scientists, interns, engineers in the use of ML techniques. - Publishing innovation in research forums. - 6+ years of building machine learning models for business…
…As a Reinforcement Learning Engineer you will advance the frontier of AI Reinforcement Learning development and delivery. You will build the infrastructure and tooling…
…developed code - Publication record in top tier venues in generative AI reasoning or classical planning - PhD in a relevant field (reinforcement learning, neurosymbolic AI…
…Developing systems that enable models to use computers effectively Advancing code generation through reinforcement learning Pioneering fundamental RL research for large language models Building…
…DPO, RLHF, or KTO), reward model design and training, and reinforcement learning to improve output quality, controllability, and human preference adherence. Editing — Research in…
…Build continuous learning systems where Advisor improves from every interaction — leveraging reinforcement learning from human feedback (RLHF), outcome-driven reward signals, and retrieval-augmented…
…Background in coding models or code-generation systems, and/or research experience in reinforcement learning (RLHF, RLAIF, preference modeling) or related model-training techniques…
…research, machine learning, or a related quantitative field - Experience with optimization solvers (e.g., Gurobi, OR-Tools) or reinforcement learning frameworks (e.g., RLlib…
…NVIDIA is building an RL Frameworks engineering team to develop the open-source tools and infrastructure that AI researchers and post-training teams depend…
…You'll adopt and adapt state-of-the-art techniques — including supervised fine-tuning, reinforcement learning, preference optimization, and knowledge distillation — running rigorous experiments…
…agents, reinforcement learning, post-training workflows. Ability to work effectively in a fast-moving environment that blends research exploration with product and engineering rigor…
…RL phases. Strong software engineering fundamentals; comfortable in Python and Bash, comfortable reading and refactoring large internal codebases. 5+ years experience in Machine Learning…