Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Research Engineer, Performance RL (Reinforcement Learning)”. A match may be a passing mention rather than the job itself. Titles only.
28 roles across 29 listings · show every listing · page 1 of 2
…Experience with reinforcement learning from human feedback (RLHF) and preference optimization. Expertise in statistical analysis, A/B testing, and experimental design at scale.
…generation, or human evaluation. - Hands on experience with reinforcement learning, preference optimization, SFT, RLHF/RLAIF, or other post-training techniques for LLMs. - Experience optimizing…
…Our research team pushes the frontier of post-training and reinforcement learning. Our applied AI team sits side-by-side with customers as they…
…Define and maintain clear research quality standards and engineering best practices and engineering standards for the team PhD degree in Computer Science, Machine Learning…
…OUR IDEAL STAFF RESEARCH SCIENTIST, EXOTIC AI WILL HAVE: - 8+ years of relevant experience in machine learning engineering, AI research, or a closely related…
…engineers - Experience with sim-to-real transfer at scale (domain randomization, system identification, etc) - Experience designing reward functions and training curricula for reinforcement learning…
…high-fidelity simulations where AI learns to perform real-world tasks through reinforcement learning. We work with the leading AI labs to help them…
…high-fidelity simulations where AI learns to perform real-world tasks through reinforcement learning. We work with the leading AI labs to help them…
…high-fidelity simulations where AI learns to perform real-world tasks through reinforcement learning. We work with the leading AI labs to help them…
…Robotics, Mechanical Engineer, Electrical Engineering, or a related field with a focus on reinforcement learning, robot learning, or control - Experience applying RL to physical…
…WHOOP is hiring a Senior Product Manager to work at that layer, alongside our AI engineers, research scientists, and data scientists. The work spans…
…and multiprocessing) and demonstrated excellence in related performance analysis and tuning. Prior experience with Reinforcement Learning algorithms and compute patterns Expertise in distributed computing…
…of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the Team Our Reinforcement Learning teams are…
…reinforcement learning and related learning-based control for high-DOF, multi-fingered robotic hands—translating state-of-the-art research (reinforcement / imitation learning, teleoperation…
…Our research team pushes the frontier of post-training and reinforcement learning. Our applied AI team sits side-by-side with customers as they…
…We aim to push the envelope by combining traditional autonomous systems algorithms with deep reinforcement learning-based solutions to deliver unmatched capability, agility, and…
…engineers - Hands-on experience building real-time AI systems — speech, audio, or video - Track record with post-training methods: reinforcement learning, reward modeling, RLHF…
…s coding data products, RL environments, and agentic coding evaluations. You will work across AI Product Management, ML Researchers, Engineering, Operations, and Go-To…
…reinforcement learning. This role sits at the intersection of research and large-scale systems engineering: a builder who understands both the algorithms behind RL…
As an Applied Scientist on the Science SW team, you will collaborate closely with other scientists and engineers to bring Reinforcement Learning (RL) research…
…collaborations with other research teams in Alphabet. AI Foundations areas that we are currently focusing on include reinforcement learning, learning from demonstration, generative modeling…
…reinforcement learning for LLMs (RLHF, RLAIF), model alignment, and advanced context techniques. Experience building systems for multi-intent understanding, multi-hop reasoning, deep research…
…train LLMs using techniques like SFT, DPO, Reinforcement Learning (RLHF and RLAIF) for supporting model performance specific to a customer's location and language…
…reinforcement learning, reward modeling, RLHF/RLAIF, or alignment techniques applied to generative models - Experience with real-time interactive systems — models that handle concurrent input…
…collaborations with other research teams in Alphabet. AI Foundations areas that we are currently focusing on include reinforcement learning, learning from demonstration, generative modeling…