Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Reinforcement Learning Engineer”. A match may be a passing mention rather than the job itself. Titles only.
1,162 roles across 1,285 listings · show every listing · page 2 of 47
…ABOUT THE ROLE Our team scales reinforcement learning for frontier models. Progress in RL is increasingly set by how well it scales: more rollouts…
…Experience with reinforcement learning (RL) and post-training with LLMs, especially spoken LLMs. Excellent engineering skills in Python and deep learning frameworks (e.g…
…By bridging multi-modal telemetry, EHR data, and reinforcement learning, we are shifting healthcare from reactive observation to proactive intervention. As a Research Scientist…
Company Description It all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker, Phyllis. She cried tears of…
…Speech/audio (e.g., technology duplicating and responding to the human voice), reinforcement learning (e.g., sequential decision making), ML infrastructure, or specialization in…
…to the state of the art in technical innovation, be it reinforcement learning, distributed systems, generative AI, or deployment infrastructure. The defence industry is…
…generative AI, deep learning, reinforcement learning, or specialization in another machine learning field. Preferred qualifications: Master’s degree or PhD in Engineering, Computer Science…
…for aerospace, industrial, and commercial type structural engineering projects Perform all calculations necessary for steel and reinforced concrete design Drive coordination and design necessary…
…Reinforce a positive environment by applying best practices and high-quality engineering standards. Gain deep expertise in one (or more) subareas of research, and…
…The role Join our GPU simulation team and build the simulator that trains our reinforcement-learning agents for air autonomy. You will shape the…
…alike Stay curious about new technologies and enjoy learning enough to speak credibly with software engineers about their work Manage multiple sourcing pipelines simultaneously…
…US: $174000 - $252000 (USD) + 15% bonus target + equity + benefits Learn more about benefits at Google . Drive post-training research and engineering using reinforcement learning…
…on cost savings, schedule optimization, value engineering, customer service, and asset management. Capture, integrate, and share lessons learned across project teams and relevant stakeholders…
…Machine Learning Compiler Engineer on the NKI team, you will be a thought leader supporting the ground-up development and scaling of a compiler…
…Apply reinforcement learning and preference optimization to improve personalization and dialogue policies. Scale LLM systems through caching, batching, prompt governance, and evaluation frameworks. Implement…
…Natural Language Processing, Computer Vision, Speech Recognition, Reinforcement Learning, Ranking and Recommendation, or Time Series Analysis. Solid understanding of Cloud (AWS/ Azure/ GCP) and…
…Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About…
…We doubled down on everything we learned and built Bolt.new — the fastest way to go from idea to production without writing traditional code…
…Run tight feedback loops into Product, Engineering, and Research, shaping the product based on what you learn measuring our most demanding deployments. Your background…
…As a Solutions Engineer, you'll serve as a trusted technical voice, guiding customers toward clarity, feasibility, and alignment on the end-to-end…
…As a Solutions Engineer, you'll serve as a trusted technical voice, guiding customers toward clarity, feasibility, and alignment on the end-to-end…
…Experience in engineering automated evaluation harnesses, reward models, or testing frameworks (e.g., LLM-as-a-judge, visual regression testing, Reinforcement Learning from Human…
…agentic AI (planning, tool use, memory), reinforcement learning (RLHF, RLVF, RLGF, offline RL). One or more scientific publication submissions for conferences, journals, or public…
…Speech/audio (e.g., technology duplicating and responding to the human voice), reinforcement learning (e.g., sequential decision making), ML infrastructure, or specialization in…
…Our work frequently takes us right up to the state of the art in technical innovation, be it reinforcement learning, distributed systems, generative AI…