Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Applied Scientist, RL post-training, AWS”. A match may be a passing mention rather than the job itself. Titles only.
8 roles across 9 listings · show every listing
…prototyping, ablations, scaling experiments, evaluation, and delivery into production, with rigorous and reproducible evaluation. - Partner closely with post-training, RL/RLHF, and instruction-following…
…for a Research Engineer / Scientist to join the Future of Computing Research team to work on RLHF and post-training for personalized, multimodal AI…
…platform — including emerging directions such as reinforcement learning (RLHF/RLVR), agent optimization, and other post-training and agentic techniques — enabling the next generation of…
…platform — including emerging directions such as reinforcement learning (RLHF/RLVR), agent optimization, and other post-training and agentic techniques — enabling the next generation of…
…Strong engineering and R&D experience in LLM post-training, reinforcement learning for language models (RLHF, RLAIF), reward modeling, policy optimization, model alignment, grounding…
…Bigtable, Redis, or feature platform infrastructure. - Experience with post-training techniques such as fine-tuning, RLHF, reinforcement learning, or reward modeling. - Experience building agentic…
…Turing accelerates frontier research with high-quality data, specialized talent, and training pipelines that advance thinking, reasoning, coding, multimodality, and STEM. For enterprises, Turing…
…platform — including emerging directions such as reinforcement learning (RLHF/RLVR), agent optimization, and other post-training and agentic techniques — enabling the next generation of…