Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Engineer, Productivity - Inference Runtime”. A match may be a passing mention rather than the job itself. Titles only.
27 roles · page 1 of 2
…experience reasoning about model routing, inference cost, and latency tradeoffs in production Strong software engineering fundamentals: distributed systems, API design, and testing discipline Comfort…
…hybrid environments Collaborate with AI and software engineering teams to understand their needs, provide golden paths to production, and build internal tools that accelerate…
…software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products…
…spanning multiple engineering and science teams, product organizations, and VP orgs, driving cross-functional alignment on AIR's most critical runtime and platform initiatives…
…The Qualcomm Cloud AI team develops hardware and software platforms enabling efficient inference of large-scale foundation models. We are seeking a Staff Engineer…
…We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom…
…inference infrastructure, developer tooling) at an AI lab or an AI-native product company, or led adoption of AI-driven development inside an engineering…
…in Computer Science, AI/ML, Engineering, or a related field, or equivalent experience. 8+ years of software engineering experience, including ownership of production systems…
…About the Role As a Software Engineer, Trainium, you will help bring OpenAI's inference workloads to AWS Trainium and build the software stack…
…experience reasoning about model routing, inference cost, and latency tradeoffs in production Strong software engineering fundamentals: distributed systems, API design, and testing discipline Comfort…
…On-device client frameworks hand requests to a cloud service that attests, routes, and orchestrates them; an inference engine serves them; and model runtimes…
…We are seeking a Systems Development Engineer to develop automation software, diagnostic tooling, and fleet health infrastructure for our accelerated (AI/ML) server platforms…
…including Python, Ansible, Container Runtimes, Kubernetes, and data center deployments Familiarity with AI workloads, including agentic & RAG-based workflows, inference at scale, large scale…
…ABOUT THE ROLE Commure is hiring a Senior Software Engineer to own the platform powering our fleet of OpenClaw agents. Across Commure, we use…
…The Inference Enablement and Acceleration team works side by side with compiler engineers and runtime engineers to create, build and tune distributed inference solutions…
…As a Senior Research Engineer, you will lead engineering across model development and runtime systems, building capabilities and turning research into systems that work…
…As a Research Engineer, you will contribute across model development and runtime systems, building capabilities and helping turn research into systems that work in…
…Hardware-software co-design Systems for AI/ML S ystems Infrastructure for large LLM training and inference Systems/AI algorithms codesign (e.g., for…
…Hardware-software co-design Systems for AI/ML Systems Infrastructure for large LLM training and inference Systems/AI algorithms codesign (e.g., for sparsity…
…networking fundamentals (TCP/IP, HTTP, gRPC). - Software engineering: 5+ years in Python, Go, C++, or Rust, writing production-grade tools and systems code. - Cloud…
…Prototype and validate software capabilities across kernels, compiler and runtime layers, distributed training and inference frameworks, and serving infrastructure. Analyze workload behavior at scale…
We are looking for a Senior Inference Engineer to own inference for real-time multimodal conversational AI. This is a full-stack inference role…
…software systems that have been successfully delivered to customers, or experience with Machine Learning and Large Language Model fundamentals, including architecture, training/inference lifecycles…
…This senior software engineering role is part of the Machine Learning Inference Applications team and focuses on delivering high-performance model inference solutions for…
…Act as the key technical liaison between Chinese CSP customers and NVIDIA Networking BU (NBU) global engineering, product and R&D teams, collect high…