Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Member of Technical Staff - RL Inference”. A match may be a passing mention rather than the job itself. Titles only.
39 roles · page 1 of 2
…Design and optimize our inference stack for all shapes of RL workloads at SpaceXAI, from small scale ablations to production training runs. Analyze, profile…
…You will work on systems that determine inference latency, throughput, stability, and the reliability of RL and post-training training loops. Magic’s long…
Overview Microsoft AI is looking for a Member of Technical Staff, Multimodal Infrastructure to help build the next wave of capabilities of our personalized…
…Set technical direction for experimentation, measurement, and evaluation frameworks, including offline metrics, online experimentation, and causal inference for discovery and generative use cases. Provide…
…A day in the life Amazon offers a full range of benefits that support you and eligible family members, including domestic partners and their…
…frontier of post-training, from data curation to large-scale optimization. - Develop data generation pipelines, reward models, reinforcement learning algorithms, and inference-time scaling…
…Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving. - Build What’s Next: Work with bleeding-edge…
…Our approach combines frontier-scale pre-training, domain-specific RL, ultra-long context, and inference-time compute to achieve this goal. ABOUT THE ROLE…
…budgets for training and inference, sized to support frontier-scale experimentation including large-model pre-training and post-training, RL training runs, and large…
…at Snowflake. - Support team members in delivering a high level of technical quality. IDEAL REQUIREMENTS & QUALIFICATIONS: - Have 7+ years of industry experience designing, building…
…AI & GPU LANDSCAPE - Strong working knowledge of the modern AI stack - open model families, finetuning techniques (LoRA, QLoRA, full FT, RLHF/RLAIF), inference engines…
…Optimization and integration of inference systems into our RL training stack. CORE TECHNICAL RESPONSIBILITIES LLM Serving - Multi‑tenant LLM Serving: Build a multi-tenant…
…the intersection of frontier research, real infrastructure, and go-to-market for a category that does not fully exist yet. Core Technical Responsibilities This…
…Come be a part of what’s next. We're seeking a technical leader to help shape the strategy, development & delivery of GenAI tools…
…As a Member of the Research Staff, this individual should have extensive experience working on the hard technical aspects of LLMs, such as data…
…A day in the life Amazon offers a full range of benefits that support you and eligible family members, including domestic partners and their…
…at least one of: JAX / PyTorch / XLA. Proven track record building or optimizing large-scale distributed ML systems (training/inference optimization, GPU utilization, multi…
AWS Trainium is deployed at scale, with millions of chips in production, used for training and inference of frontier models. AWS Neuron is the…
…Experience with alignment or RL techniques beyond basic supervised fine-tuning. - Familiarity with on-device or low-latency inference constraints. WHAT SUCCESS LOOKS LIKE…
…IDEAL EXPERIENCE - Experience deploying and operating large-scale GPU systems for inference or model serving. - Several years of hands-on experience building and running…
Overview Microsoft AI is looking for a Member of Technical Staff – Capacity & Efficiency Infrastructure , to help us improve manage, and improve the efficiency of…
…paper. - You have deep experience in at least one of the following: Distributed Training & Inference or Data Infrastructure - You enjoy working at the boundary…
…Implement efficient algorithms for state-of-the-art model performance, including real-time inference, distillation, and scalable serving for visual content. Develop scalable data…
…Our approach combines frontier-scale pre-training, domain-specific RL, ultra-long context, and inference-time compute to achieve this goal. ABOUT THE ROLE…
…in principal/staff-level technical leadership. Experience with large-scale, real-time ML systems (recommendations, personalization, matchmaking). Expertise in graph ML, RL, and representation…