Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Research Engineer, Code RL (Reinforcement Learning)”. A match may be a passing mention rather than the job itself. Titles only.
19 roles across 20 listings · show every listing
…vetting design, solution approaches, results and code. Deep research expertise in multi-agent systems and reinforcement learning including agent co-ordination, negotiation, communication, simulations…
…Define and maintain clear research quality standards and engineering best practices and engineering standards for the team PhD degree in Computer Science, Machine Learning…
…OUR IDEAL STAFF RESEARCH SCIENTIST, EXOTIC AI WILL HAVE: - 8+ years of relevant experience in machine learning engineering, AI research, or a closely related…
…ML codebases and distributed systems. - Experience improving model behavior through data, reward modeling, or RL techniques. - Evidence of owning ambitious research or engineering agendas…
…high-fidelity simulations where AI learns to perform real-world tasks through reinforcement learning. We work with the leading AI labs to help them…
…We build training gyms for AI agents — high-fidelity simulations where models learn to solve economically valuable problems through reinforcement learning. We work with…
…high-fidelity simulations where AI learns to perform real-world tasks through reinforcement learning. We work with the leading AI labs to help them…
…high-fidelity simulations where AI learns to perform real-world tasks through reinforcement learning. We work with the leading AI labs to help them…
…You’ll gain hands on experience with real data, production infrastructure and real deadlines, while learning from and working alongside the researchers and engineerings…
…of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the Team Our Reinforcement Learning teams are…
…reinforcement learning and related learning-based control for high-DOF, multi-fingered robotic hands—translating state-of-the-art research (reinforcement / imitation learning, teleoperation…
…engineers - Hands-on experience building real-time AI systems — speech, audio, or video - Track record with post-training methods: reinforcement learning, reward modeling, RLHF…
…Scale's coding data products, RL environments, and agentic coding evaluations. You will work across AI Product Management, ML Researchers, Engineering, Operations, and Go…
…reinforcement learning. This role sits at the intersection of research and large-scale systems engineering: a builder who understands both the algorithms behind RL…
…We're looking for individuals who have a background in reinforcement learning research, are able to iterate quickly, and who can convert scientific rigor…
…reinforcement learning, reward modeling, RLHF/RLAIF, or alignment techniques applied to generative models - Experience with real-time interactive systems — models that handle concurrent input…
…researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role Code RL at Anthropic drives reinforcement learning…
…for machine learning research or RL workflows, and familiarity with agentic systems or LLM training pipelines Experience building agent frameworks, orchestration engines, or multi…
…integrating inference engines (vLLM, SGLang), weight synchronization, and asynchronous/off-policy schemes. - Design RL environments for agentic, multi-step tasks — sandboxed code execution, tool…