Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Engineer, RL Training Infra”. A match may be a passing mention rather than the job itself. Titles only.
211 roles across 223 listings · show every listing · page 4 of 9
…real infrastructure, and go-to-market for a category that does not fully exist yet. ROLE IMPACT This is a generalist software engineering role…
…CORE TECHNICAL RESPONSIBILITIES This hybrid role spans across our AI platform software engineering and infrastructure. You'll be instrumental in: PLATFORM DEVELOPMENT - Build intuitive…
…engineering action - Strong fundamentals in cryptography, identity/access management, and secure software development lifecycle NICE TO HAVE - Experience securing GPU infrastructure or ML training…
…It is not a traditional solutions engineering role. You will help define how Prime Intellect turns frontier post-training infrastructure into a product customers…
…operating roles - Product management for AI, infra, devtools, or enterprise software - ML engineering, applied research, or AI engineering with customer exposure - Venture/investing roles…
…Collaborate with ML researchers, software developers, and product managers across Apple to translate product requirements into scalable, reliable, and efficient model evaluation infrastructure. Bachelor…
…training to production serving, to optimizing the inference engine for RL training workloads. You will collaborate closely with our product, research, and engineering teams…
…in Computer Science or equivalent 6+ years of industry experience in software engineering Deep backend engineering fundamentals, especially in Python and distributed systems. Track…
We are looking for a Software Test development engineer in NVIDIA’s AI SWQA team. The position is in NVIDIA AI Software Quality Assurance…
…AI training and/or inference infrastructure — RL/post-training, training frameworks, or inference serving. Proficiency in Python (plus scripting), and solid software engineering practices…
…training/inference, Agentic AI, gaming AI, and distributed computing infrastructure. You will accelerate the mass deployment and performance maximization of NVIDIA GPU software/hardware…
…NVIDIA is building an RL Frameworks engineering team to develop the open-source tools and infrastructure that AI researchers and post-training teams depend…
…In this role, you will collaborate closely with the Onboard Perception, Cost Planner, Simulation, Validation, Data Science, Systems Engineering, QA, and ML Infra teams…
…3+ years of full-time engineering experience, post-graduation with specialities in infrastructure and identity systems. Infrastructure expertise – IAM controls, Infrastructure as Code (Terraform…
…training and infrastructure stack to improve our production training setup Partner with product teams to bring research advances into production Minimum Qualifications: Software engineering…
…judge, or defect/quality analysis. - Familiarity with modern training/inference infrastructure (e.g., distributed training, RL frameworks, model serving). Amazon is an equal opportunities…
…Transformer Engine and heterogeneous compute targets (GPU, CPU, LPU) Experience driving performance optimization efforts for large-scale distributed AI training, post-training (RLHF, alignment…
…a backend software engineer to own and evolve our internal LLM agent framework. This role sits at the intersection of backend infrastructure, applied AI…
…post-training stack, including RL, data pipelines, graders, reward signals, evals, and behavioral analysis. You will work with researchers, engineers, product teams, infrastructure teams…
…This work sits at the intersection of frontier model training, product behavior, evaluation, and systems engineering, and will directly shape the computer-use capabilities…
…enterprise software, turning connected tools into a powerful action surface for our agents. You will work with researchers, engineers, product teams, infrastructure teams, and…
…of model training with a clear product interface for iterative deployment (Codex Chronicle). You will work with researchers, engineers, product teams, infrastructure teams, and…
…You will work with researchers, engineers, product teams, infrastructure teams, and safety/alignment partners to decide what should go into major model runs, measure…
…researchers, engineers, API/product teams, Codex, infrastructure, and safety/alignment partners to decide which behaviors matter, how to measure them, how to train them…
…control For sandbox infrastructure: own the Python SDK and work in a tight loop with the backend team, enabling RL training runs to spawn…