Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “AI Research Engineer - Reinforcement Learning”. A match may be a passing mention rather than the job itself. Titles only.
652 roles · group by role · page 1 of 27
…Are passionate about keeping up to date with current research and enjoy reimplementing and extending state-of-the-art approaches in deep reinforcement learning…
…dexterous manipulation and relevant frontier research problems - Conduct original research on world-action foundation models and reinforcement learning, delivering manipulation policies for humanoid robotic…
…engineering, research, and everything in between. Your contributions will span model architecture, data curation, training and inference infrastructures, evaluation protocols, alignment and reinforcement learning…
…record in machine learning or AI research, with the depth to push at least one important area forward: pretraining, reinforcement learning / post-training, evals…
The mission of Thinking Machines is to build AI that extends human will and judgment. ABOUT THE ROLE Our team scales reinforcement learning for…
…to drive independent research initiatives, work with teams on AI, and develop solutions to fundamental questions in machine learning and AI. Artificial intelligence will…
…By bridging multi-modal telemetry, EHR data, and reinforcement learning, we are shifting healthcare from reactive observation to proactive intervention. As a Research Scientist…
…We work on a range of unique problems focused on research topics that maximize scientific and real-world impact, aiming to push the state…
Overview We are seeking a Systems Research Engineer to accelerate our Future AI Infrastructure research area. You will bring key skills and experience in…
…The role Join our GPU simulation team and build the simulator that trains our reinforcement-learning agents for air autonomy. You will shape the…
…US: $174000 - $252000 (USD) + 15% bonus target + equity + benefits Learn more about benefits at Google . Drive post-training research and engineering using reinforcement learning…
…As an AI Agents Applied Research/Engineering Executive Director in our Digital Team, you will work with the team to shape how millions of…
…Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About…
…Run tight feedback loops into Product, Engineering, and Research, shaping the product based on what you learn measuring our most demanding deployments. Your background…
…Experience in engineering automated evaluation harnesses, reward models, or testing frameworks (e.g., LLM-as-a-judge, visual regression testing, Reinforcement Learning from Human…
…agentic AI (planning, tool use, memory), reinforcement learning (RLHF, RLVF, RLGF, offline RL). One or more scientific publication submissions for conferences, journals, or public…
…Preferred/Additional Qualifications Proven software engineering skills, evidenced by professional experience, internships, and impactful open-source contributions. Experience with reinforcement learning, imitation learning, or…
…automation and AI tooling that measurably increases the velocity of engineering, research, and go-to-market teams. - Partner with research engineers to scope product…
…Fluency with AI productivity tools such as Claude Code and Codex, with a track record of thoughtfully integrating AI into engineering and research workflows…
…To do this, we believe that many technical breakthroughs are needed in generative modeling, reinforcement learning, large-scale optimization, active learning, and other areas…
ABOUT ELEVENLABS ElevenLabs is an AI research and product company transforming how we interact with technology. We launched in January 2023 with the first…
…the future of AI—no bureaucracy, just results. - Learn from the Best: Collaborate with world-class engineers and AI researchers who thrive on curiosity…
…Hands-on experience implementing Reinforcement Learning from Human/AI Feedback (RLHF/RLAIF) or direct preference optimization (DPO) loops. Multimodal Experience: Experience working with multimodal…
…reinforcement learning alignment loops (RLHF/DPO) to guarantee model safety and predictability in high-stakes environments. Collaborate and Pioneer: Work closely with AI Researchers…
…The team includes full-time researchers and engineers, along with postdocs, interns, and research fellows in fixed-term roles. Learn more about Microsoft’s…