Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Research Scientist - RL Training”. A match may be a passing mention rather than the job itself. Titles only.
84 roles across 89 listings · show every listing · page 1 of 4
…RL training recipes that inform what data Snorkel ships as part of its data-as-a-service deliveries. Work closely with research scientists, ML…
As a Senior Applied Scientist in the Alexa AI team, you will define and drive the science roadmap for state-of-the-art conversational…
…You will leverage Reinforcement Learning (RL), sim-to-real transfer, and other learning-based architectures to train policies that produce stable, dynamic gaits across…
…Trainium/Inferentia, SageMaker, EKS, PCS, EC2) to shape product vision, prioritize features, and represent the voice of the customer - Work with account teams, research…
…demonstrated research experience with a strong portfolio of independent, published work - Expert-level knowledge of post-training methods including SFT, RLHF, RLAIF, DPO, GRPO…
…in Computer Science, Machine Learning, Statistics, Operations Research, or related quantitative field. - Deep expertise in large language models, post-training techniques (RLHF, fine-tuning…
…and RL research domains - Publish research findings at top-tier venues such as NeurIPS, ICML, ICLR, and ACL OUR IDEAL AI RESEARCH SCIENTIST WILL…
…research at top-tier venues and represent Amazon Robotics in the broader academic and industry community • Mentor and develop a team of applied scientists…
…solutions, translating ambiguous operational challenges into well-scoped research with clear success criteria Develop RL-based control strategies that enable self-optimizing data center…
…training, inference serving, data loading, checkpointing, memory usage, and GPU utilization. * Design, implement, and evaluate novel approaches to LLM fine-tuning, alignment (RLHF, DPO…
…stability Optimize RL post-training efficiency (GPU utilization, batching, sequence packing, async rollouts) Partner with research scientists to translate new RL algorithms into scalable…
…We're seeking an exceptional Research Scientist to join our Life Sciences team at Anthropic. Our team is building a world-class research group…
…fast-moving AI, ML, or research environment. - Deep familiarity with post-training and alignment concepts (supervised fine-tuning, RLHF, AI safety frameworks, LLM evaluation…
…You will be able to reshape the landscape of frontier AI research directions by pre-training/post-training frontier multi-modal generation models with…
…combination of ambitious research vision and practical impact. We leverage Amazon's computational infrastructure and rich real-world datasets to train and deploy state…
…integration—someone equally strong in (1) LLM training (domain-adaptive continual pretraining, post-training, preference optimization / RL such as GRPO-style methods), (2) agentic…
The AWS Neuron Science Team is looking for talented scientists to enhance our software stack, accelerating customer adoption of Trainium and Inferentia accelerators. In…
…Working knowledge of frontier model development (benchmarks, evals, RL, post-training, agentic systems) and the data and environments behind them. Technical or research background…
…world‑class researchers and engineers to advance the state of the art in large‑scale machine learning, focusing on post-training, RL and inference…
Scale Labs, Research Scientist — Safety Post Training As the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in…
…Make Wayve the experience that defines your career! The Role We’re looking for Research Scientists to join Wayve Labs and help build the…
…Experience with large-scale distributed training for RL - History of technical leadership and cross-functional collaboration - Experience bridging research with practical engineering implementation in…
Meta’s Reality Labs Research (RL-R) brings together a team of researchers, developers, and engineers to create the future of Mixed Reality (MR…
…You will also work with researchers and data scientists to develop, fine-tune, and evaluate domain specific Large Language Models for various tasks and…
…Reinforcement Learning (RL) post-training of frontier Large Language Models (LLMs) to revolutionize customer service? Come join the world class researchers and academics in…