Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Research Scientist - RL Training”. A match may be a passing mention rather than the job itself. Titles only.
36 roles across 40 listings · show every listing · page 1 of 2
…Mentor research scientists and engineers on the team, upleveling internal capabilities in post-training and agentic AI. PhD in Computer Science, Machine Learning, Natural…
…of model training, data recipe, and their research. Have experience with synthetic data generation, data curation and cleaning, Prior training and RL, evaluation, or…
…Collaborate with Data Scientists, Researchers, and Engineers to drive improvements across our platforms. We are looking for a Machine Learning Engineer focused on Evaluation…
…You will collaborate closely with applied scientists, research scientists, economists, product managers, and software engineers. We are seeking innovative thinkers who can balance theoretical…
…JPMorganChase AI Research is a global team of research scientists, engineers and product managers that develops novel AI capabilities and partners across the firm…
…Build, mentor, and grow a team of research scientists and research engineers Technical Strategy & Execution: Oversee work across the full LLM post-training stack…
…We are hiring a Staff Research Scientist, Physical AI for our AI Research team. You will build the next-generation training and learning platform…
…Key job responsibilities As an Applied Scientist with the FAI team, you will support the development of RL Gyms, assess their usefulness for the…
…mentoring scientists or engineers - Experience with sim-to-real transfer at scale (domain randomization, system identification, etc) - Experience designing reward functions and training curricula…
The Catalog Services Product Knowledge team is seeking an Applied Scientist for the Catalog Services organization. Our vision is simple: build AI systems that…
…scale, and the ability to write research code for data processing, training and evaluation. Multi-GPU training (FSDP, DeepSpeed) and evaluation-harness engineering. Layered…
We are seeking an Applied Scientist to join the SAF Lab. In this role, you will lead the effort in safe reinforcement learning (RL…
…Meta is seeking a HW Research Scientist to join our Research and Development teams. The candidate will have experience working on AI models, hardware…
…WHOOP is hiring a Senior Product Manager to work at that layer, alongside our AI engineers, research scientists, and data scientists. The work spans…
…Post-train and adapt open-source LLMs for SCC use cases using SFT, LoRA, and preference-tuning methods (RLHF, RLAIF, RLVR). Design and build…
…to address identified limitations - Guide fellow scientists in solving complex technical challenges, from sim2real transfer to training RL policies - Mentor team members while maintaining…
…training, evaluation, and deployment to help researchers move faster and tackle increasingly ambitious problems. About the Role We’re hiring research scientists, research engineers…
We are looking for a Senior Applied Scientist to help drive the research and development of real-time multimodal conversational AI. You will contribute…
…models (training, fine-tuning, or evaluation) Preferred Qualifications - 5+ years of relevant industry or academic research experience - Experience with LLM alignment techniques (RLHF, DPO…
As an Applied Scientist on the Science SW team, you will collaborate closely with other scientists and engineers to bring Reinforcement Learning (RL) research…
…This includes scaling ladders for pre-training, data mix optimizations and RL recipes (preferences) for post-training. Engage with the wider research community on…
…About the Role As a Research Engineer / Research Scientist on the Personalization-Memory team, your work will span memory architecture, post-training, and developing…
…tune/post-train LLMs using techniques like SFT, DPO, RLHF, and RLAIF. * Collaborate with partner teams on evaluation frameworks and post-training methodologies. * Communicate…
…research or industry experience - Foundational knowledge of state-of-the-art LLM architectures, training, evaluation, and post-training techniques (SFT, DPO, RLHF, RLAIF) - Knowledge…
We are looking for a Principal Applied Scientist to drive the research and development of real-time multimodal conversational AI. You will operate across…