Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Applied AI Engineer, Inference”. A match may be a passing mention rather than the job itself. Titles only.
2,812 roles · group by role · page 1 of 113
…We are looking for an Applied AI Engineer to help us understand, measure, and improve the real-world performance of our inference platform. In…
…engineering fundamentals, with a record of building and shipping AI/ML inference systems. - Experience with Docker and Kubernetes. - Prior work building or tuning AI…
…We measure how effectively AI is applied to deliver results, and consistent, creative use of the latest AI capabilities is key to success here…
…You will work closely with research scientists and product engineers on multimodal data processing, model training, inference and serving tasks . As a contributing member…
…Knowledge and experience with Docker, Kubernetes, High-performance application development. Technical leadership skills and ability to mentor early-in-profession engineers. #MicrosoftAI Software Engineering…
…Make Wayve the experience that defines your career! The Role As a Data Scientist supporting AI engineers, you will partner with one or more…
…Job Description About the Team The Agentic Engineering organization at ServiceNow is the customer-obsessed engineering group that builds a conversational AI experience that…
…AI - No CS degree or formal engineering tenure required. Shipping real things fast with AI-assisted tooling clears our bar - Open to using AI…
…inference: numerics, quantization, and their implications for RL. - Hands-on work with LLM serving stacks (e.g., SGLang, vLLM, TokenSpeed, or custom engines). - Experience…
…based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration…
…turn generative AI into one of their most valuable assets. About the team The Cloud Platform team owns the production inferencing service that serves…
…Solve technically tests problems that exceed the scope of a generalist Software Engineers, specifically around optimizing Generative AI performance across heterogeneous hardware (CPUs, GPUs…
…Lawful permanent residents, refugees, and asylees may verify status using other documents, where applicable. Ability to work in an "AI first" environment using modern…
…critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research…
…The Marketplace Intelligence (MI) team is looking for an Applied Science Manager to lead a team of scientists and engineers in building production ML…
You will join a dynamic team working at the cutting edge of the GenAI revolution by applying AI to AI. You will work on…
…The ideal candidate has a PhD in Economics and deep expertise in causal inference and applied econometrics. Experience with large-scale data, proficiency in…
…of Generative AI cloud at AWS? Do you want to build the future of the cloud for AI training and inference? Want to do…
…with Principal Engineers, SDMs, and Principal Scientists, identifying dependencies, scaling factors, boundary conditions, and risks across horizontal distributed systems serving AI inference at global…
…The Qualcomm Cloud AI team is developing hardware and software solutions for Inference Acceleration. We are hiring LLM Serving Engineers at multiple levels to…
…The Qualcomm Cloud AI team is developing hardware and software solutions for Inference Acceleration. We are hiring an AI Performance Engineers at multiple levels…
…The Qualcomm Cloud AI team develops hardware and software platforms enabling efficient inference of large-scale foundation models. We are seeking a Staff Engineer…
…Engineering Group, Engineering Group > Software Engineering General Summary: Job Description The Qualcomm Cloud AI team is developing software solutions for Inference Acceleration. We are…
…as Ray, Spark and in training/inference systems such as Ray, vllm/SGLang Solid grounding in engineering fundamentals and enterprise system design Preferred qualifications…
…cycle for applications natively powered by foundation models (Large Language Model (LLMs)/Lower Mass Market (LMMs) or agentic architecture. Experience prototyping AI features using…