Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Research Engineer - Distributed Training”. A match may be a passing mention rather than the job itself. Titles only.
1,527 roles across 1,708 listings · show every listing · page 3 of 62
…distributed systems. Optimize system performance, minimize latency, and maximize throughput. Collaborate effectively with cross-functional teams, including product managers, research scientists, and other engineering…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…Amazon Web Services has supported over 10,000 local startups and has provided cloud skills training to over 700,000 talents. Amazon’s first…
…autonomous driving - Track record of successful production ML deployments - Experience with large-scale distributed environments for ML training and inference - History of impactful first…
…and non-technical audiences and at C-level, including training, workshops, publications - Knowledge of distributed systems design and implementation or equivalent - Knowledge of large…
…task orchestration. - Leadership experience with of science or research teams in scalable ML infrastructure, distributed training, or optimization of large models. - Deep knowledge of…
…As a Lead Software Engineer at JPMorgan Chase within the Commercial & Investment Bank's Markets Research Technology Team, your role will be pivotal in…
…We're looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data…
…About this role As the Senior Staff Machine Learning Platform Engineer, you will own the technical vision and evolution of Faire’s ML platform…
…You will work on: - Flow-level observability across payout → settlement → reconciliation - Reliability engineering for distributed payment systems - Instrumentation and monitoring of financial primitives (ledger…
…group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role Training and serving frontier…
…group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role Training and serving frontier…
…PyTorch), with a solid foundation in software engineering practices and comfort with large-scale training. Strong ownership: research-literate and pragmatic, able to drive…
…Experience leading or partnering with distributed teams across UK and US time zones. MS or PhD in Computer Science, Engineering, or a related field…
…medical researchers deserve the same cutting-edge technological innovations that have transformed other industries. We're on the lookout for passionate software engineers who…
…to support inference, pre-training, and post-training workloads for state-of-the-art AI models. • Collaborate with engineers, researchers, and external partners to…
…Responsibilities Design and deliver next-generation distributed storage systems optimized for AI/ML workloads, from training to inferencing Provide technical leadership across architecture, development…
…5+ years in a data engineering, data science, technical architecture, or similar pre-sales/consulting role Experience building distributed data systems Comfortable programming in…
…We're looking for a Machine Learning Engineer to own the data foundations that power our multimodal agent research—building the pipelines, datasets, and…
…Familiarity with distributed training frameworks (PyTorch FSDP, Megatron-LM, DeepSpeed) and how cluster architecture decisions impact training throughput, fault tolerance, and overall MFU. Understanding…
…Deep experience with containerized deployments (Docker, Kubernetes), GPU scheduling/orchestration, and distributed training frameworks (e.g., PyTorch Distributed, Ray, Slurm, or Megatron-LM). Data…
…in engineering sciences and machine learning to push the frontier of AI-accelerated simulation. Within AI4Engineering Science, you will research and train foundational physics…
…this work in how engineers actually work, and with the broader research team to translate that domain grounding into training signal and evaluation benchmarks…
…Built on DeepL’s world-leading translation quality, the Voice product line brings together Research, Engineering, Product, and GTM teams to deliver real-time…
…s Enablement and Marketing teams to ensure partner sellers and engineers are trained and certified on DeepL solutions. - Support partner sellers with opportunity qualification…