Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Engineer, Model Routing & Inference”. A match may be a passing mention rather than the job itself. Titles only.
54 roles across 58 listings · show every listing · page 1 of 3
…runtime, inference platform, or workflow engine. Experience with OpenAI, Gemini/Vertex AI, AWS Bedrock, or similar platforms. Familiarity with model and provider routing based…
…As a Software Engineer on the Cluster Deployment Automation team, you will help build the pushbutton tooling that makes large-scale cluster deployments faster…
…Our first products are heavily focused on inference. Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is…
…Our first products are heavily focused on inference. Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is…
…Our first products are heavily focused on inference. Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is…
…By combining powerful local inference (Nemotron models) with strong privacy routers and sandboxed execution, you will help develop the foundation of the desktop AI…
…By combining powerful local inference (Nemotron models) with strong privacy routers and sandboxed execution, you will help develop the foundation of the desktop AI…
…By combining powerful local inference (Nemotron models) with strong privacy routers and sandboxed execution, you will help develop the foundation of the desktop AI…
…proxy, edge, gateway, routing, or traffic management systems. Familiarity with the current AI ecosystem: model providers, model capabilities, agent frameworks, inference patterns, and the…
…model lifecycle, inference routing, and agent execution. - Own the quality of what you ship — code, test coverage, documentation, operability, and rollout safety. This is…
…enable other engineering teams. NICE TO HAVE - Experience with AI infrastructure, LLM serving, or machine learning platforms. - Experience with model routing, inference gateways, or…
…WHAT YOU'LL DO - Lead and grow a high-performing team of software engineers responsible for Harvey's Model Infrastructure platform. - Define the technical…
…models. - Experience with CUDA or comparable technologies. - A strong command of software engineering fundamentals, with a record of building and shipping AI/ML inference…
…12+ years of software engineering with depth in GPU computing, ML systems, or high-performance inference Strong Python or C++ programming, software design, and…
…As a Software Engineer III at JPMorganChase within the Firmwide LLM Serving Platform team, you are an integral part of an agile team that…
…calling, routing, orchestration, and state handling to support multi-step agentic task execution Build and maintain inference-time integrations such as model gateways, APIs…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…SKILLS & QUALIFICATIONS - 7+ years of relevant industry experience in software integration, development, or quality engineering, including prior release qualification experience. - Proven ownership of end…
…Create internal APIs and abstractions that allow Software Engineers to provision AI-ready environments (complete with model weights, vector DBs, and event streams) with…
…real-time GPU serving, high-throughput batch inference, and model fine-tuning. You’ll work across model serving and inference engines, fine-tuning and…
…real-time GPU serving, high-throughput batch inference, and model fine-tuning. You’ll work across model serving and inference engines, fine-tuning and…
…Behind every routing decision, capacity plan, and network design is a modeling problem — and the science behind how those models learn, generalize, and improve…
…As senior software development engineer in Test, we are looking for a candidate who can make a big impact on how we test and…
…workloads with ultra high-speed inference. About the Role: We are looking for a hands-on Infrastructure Engineer to join our team and support…
…Work closely with customer AI researchers, engineers, and developers to build, prototype, and deploy Agentic solutions on NVIDIA platforms. Optimize inference performance and total…