Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Applied Scientist, RL post-training, AWS”. A match may be a passing mention rather than the job itself. Titles only.
30 roles across 32 listings · show every listing · page 1 of 2
…RL) post-training of frontier Large Language Models (LLMs) to revolutionize customer service? Come join the world class researchers and academics in the AWS…
…manipulation and multi-fingered hands - VLA/WAM and post-training for robotics - Large-scale RL in robotic simulation, and sim-to-real transfer - Imitation…
…pretraining and post-training frameworks. Design, develop and maintain large-scale multimodal model inference and serving frameworks. Work with research scientists and product engineers…
…researchers with the technical depth to move the frontier in pretraining, reinforcement learning (RL) / post-training, or evals; and researchers with real depth in…
…We work on all topics of multimodal modelling, pre/post-training and design agents, we build scalable training and evaluation loops, and partner closely…
…Develop, pre-train, and fine-tune in-house LLMs and multimodal foundation models. Apply SOTA post-training alignment techniques (SFT, RLHF, DPO) to maximize…
AWS Agentic AI is looking for a Principal Applied Scientist who will partner with world-class scientists and engineers to pioneer the next generation…
…post-training pipelines (SFT through RL) — parallelism strategies, training stability, and multi-accelerator communication (NCCL, NVLink) - Familiarity with multiple hardware backends (NVIDIA GPU, AWS…
…platform — including emerging directions such as reinforcement learning (RLHF/RLVR), agent optimisation, and other post-training and agentic techniques — enabling the next generation of…
…Frontier AI labs currently improve model capabilities with various data-intensive post-training techniques. We believe that data spend for AI training will increase…
…Frontier AI labs currently improve model capabilities with various data-intensive post-training techniques. We believe that data spend for AI training will increase…
…prototyping, ablations, scaling experiments, evaluation, and delivery into production, with rigorous and reproducible evaluation. - Partner closely with post-training, RL/RLHF, and instruction-following…
…for a Research Engineer / Scientist to join the Future of Computing Research team to work on RLHF and post-training for personalized, multimodal AI…
…platform — including emerging directions such as reinforcement learning (RLHF/RLVR), agent optimization, and other post-training and agentic techniques — enabling the next generation of…
…platform — including emerging directions such as reinforcement learning (RLHF/RLVR), agent optimization, and other post-training and agentic techniques — enabling the next generation of…
…Strong engineering and R&D experience in LLM post-training, reinforcement learning for language models (RLHF, RLAIF), reward modeling, policy optimization, model alignment, grounding…
…Bigtable, Redis, or feature platform infrastructure. - Experience with post-training techniques such as fine-tuning, RLHF, reinforcement learning, or reward modeling. - Experience building agentic…
…Turing accelerates frontier research with high-quality data, specialized talent, and training pipelines that advance thinking, reasoning, coding, multimodality, and STEM. For enterprises, Turing…
…platform — including emerging directions such as reinforcement learning (RLHF/RLVR), agent optimization, and other post-training and agentic techniques — enabling the next generation of…
…traces drive automated evaluation, agent simulation, and post-training techniques (e.g., reward modeling and RLHF/RLVR evaluation) — enabling the next generation of AI…
…Build collateral (notebooks, github repos, demos, etc.) applied to workflows such as AV and GenAI data curation, model training and validations, LLMs, VFMs, video…
…Define requirements for reusable training building blocks that compose into end-to-end workflows. Post-Training, RL & Emerging Workflows Drive strategy for post-training…
…Post-Training Excellence: Build the infrastructure required for sophisticated Reinforcement Learning (RL) and RLHF pipelines, enabling labs to refine foundation models with maximum efficiency…
…RL, VLA post-training, large-scale closed-loop RL based on neural simulation with applications to autonomous driving - Work closely with Research Scientists and…
…AT APPLIED INTUITION, YOU WILL: - Conduct research on reinforcement learning (RL) related topics including large-scale closed-loop RL and VLA post-training with…