Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “AI Researcher - Reinforcement Learning”. A match may be a passing mention rather than the job itself. Titles only.
627 roles across 690 listings · show every listing · page 1 of 26
…ABOUT THE TEAM The Reinforcement Learning team teaches NEO new capabilities, training policies for manipulation and locomotion tasks across simulation and real-world environments…
…This role sits at the intersection of agentic AI and reinforcement learning, where your research will directly shape how enterprises leverage intelligent automation. AS…
…Are passionate about keeping up to date with current research and enjoy reimplementing and extending state-of-the-art approaches in deep reinforcement learning…
…dexterous manipulation and relevant frontier research problems - Conduct original research on world-action foundation models and reinforcement learning, delivering manipulation policies for humanoid robotic…
…alignment and reinforcement learning from human feedback (RLHF), and many other exciting topics at the cutting edge of AI. Microsoft AI is building foundational…
…record in machine learning or AI research, with the depth to push at least one important area forward: pretraining, reinforcement learning / post-training, evals…
The mission of Thinking Machines is to build AI that extends human will and judgment. ABOUT THE ROLE Our team scales reinforcement learning for…
…to drive independent research initiatives, work with teams on AI, and develop solutions to fundamental questions in machine learning and AI. Artificial intelligence will…
…By bridging multi-modal telemetry, EHR data, and reinforcement learning, we are shifting healthcare from reactive observation to proactive intervention. As a Research Scientist…
…Leverage broader expertise to participate in a wide variety of research, including learning from simulation, reinforcement learning, learning from demonstrations, vision-language-action models…
…We work on a range of unique problems focused on research topics that maximize scientific and real-world impact, aiming to push the state…
…Future AI Infrastructure research area. You will bring key skills and experience in operating systems, performance, distributed systems, CPUs, GPUs, and machine learning. Good…
…The role Join our GPU simulation team and build the simulator that trains our reinforcement-learning agents for air autonomy. You will shape the…
…US: $174000 - $252000 (USD) + 15% bonus target + equity + benefits Learn more about benefits at Google . Drive post-training research and engineering using reinforcement learning…
…As an AI Agents Applied Research/Engineering Executive Director in our Digital Team, you will work with the team to shape how millions of…
…About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push…
…GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with…
…translating them into practical messaging, learning experiences, and tools that sales teams can absorb quickly and apply consistently in live customer conversations. What You…
…tight feedback loops into Product, Engineering, and Research, shaping the product based on what you learn measuring our most demanding deployments. Your background looks…
…Experience in engineering automated evaluation harnesses, reward models, or testing frameworks (e.g., LLM-as-a-judge, visual regression testing, Reinforcement Learning from Human…
…agentic AI (planning, tool use, memory), reinforcement learning (RLHF, RLVF, RLGF, offline RL). One or more scientific publication submissions for conferences, journals, or public…
…Our research spans a diverse range of areas, including computer vision, generative AI, 3D perception, and robotic action learning, among others relevant to Embodied…
…By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge…
…Fluency with AI productivity tools such as Claude Code and Codex, with a track record of thoughtfully integrating AI into engineering and research workflows…
…To do this, we believe that many technical breakthroughs are needed in generative modeling, reinforcement learning, large-scale optimization, active learning, and other areas…