Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Manager, Large Language Model Inference”. A match may be a passing mention rather than the job itself. Titles only.
565 roles across 664 listings · show every listing · page 5 of 23
…From training the world's largest models to enabling autonomous AI agents that reason, plan, and act, Azure Storage provides the critical data foundation…
…Key job responsibilities - Develop and optimize compute kernels for a custom ML accelerator architecture, targeting production-level performance for large language model inference. - Implement…
…You will design RAG architectures, agentic AI workflows, model customization strategies, responsible AI implementations, and production-scale inference pipelines. You will interact with other…
…You will design RAG architectures, agentic AI workflows, model customization strategies, responsible AI implementations, and production-scale inference pipelines. You will interact with other…
…1) API-driven Inference Services like Amazon Bedrock 2) Agentic AI including Agentic platforms, tool usage, evaluations, governance and cost management along with use…
…experience with Machine Learning and Large Language Model fundamentals, including architecture, training/inference lifecycles, and optimization of model execution - Master's degree in a…
…variety of LLM model families, including massive scale large language models like the Llama family, DeepSeek and beyond. The Inference Enablement and Acceleration team…
…including product managers, applied scientists, and other engineers to define and execute on our product roadmap • Implement and optimize machine learning models for healthcare…
…Job Responsibilities Build generative AI, agentic AI, and large language model solutions in Python from proof of concept through production deployment with measurable outcomes…
…machine learning, statistics, optimization, causal inference, and experimentation Experience deploying models in production and adjusting model thresholds to improve performance Experience designing, running, and…
…Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving. - Build What’s Next: Work with bleeding-edge…
…technology management . Proficient for AI & ML models (e.g., Agentic AI framework, AI Foundry, Semantic Kernel, Foundry SDK, Responsible AI, fine-tuning/inferencing, Azure…
…model servers inference, CI deployments, observability, and large-scale cloud integration Collaborate with the ML and platform teams behind Voyage to bring new models…
…models in an applied environment - Experience in Python, Perl, or another scripting language - Experience in a ML or data scientist role with a large…
…Agentic AI and large language models are shaping how enterprise software gets built, deployed, and operated at scale. We're hiring a Senior Software…
…machine learning, statistics, deep learning, natural language processing, or information retrieval - Experience applying causal inference techniques in artificial intelligence and machine learning systems Amazon…
…and implement Generative AI workflows using large language models, including evaluation methods and feedback loops for model and pipeline improvement Translate experimental results into…
…Apply advanced ML methods—including anomaly detection, cross-signal analysis, large language models (LLMs), and other modern AI techniques—with clear evaluation frameworks, robust…
…Direct experience optimizing infrastructure for Large Language Model (LLM) training and inference at scale. - Peer-reviewed publications at top systems venues (OSDI, SOSP, NSDI…
…AI infrastructure powering frontier model development. The Compute Orchestration & Scheduling team own cluster orchestration, workload scheduling, resource allocation, quota management, and the systems that…
…getting and analyzing large amounts of data, generate insights and opportunities, design simulations and experiments, and develop statistical and ML models. The team is…
…language - Hands-on experience training large-scale foundation models — direct involvement in model training, not just using pre-trained models - Experience with multimodal model…
…or natural language processing - Experience with Machine Learning and Large Language Model fundamentals, including architecture, training/inference lifecycles, and optimization of model execution Amazon…
…planning, forecasting, and infrastructure scaling - Experience with language-side server architectures (NLU, ASR, TTS, or LLM inference serving) Amazon is an equal opportunity employer…
…AWS Neuron is the software of Trainium and Inferentia, the AWS Machine Learning chips. Inferentia delivers best-in-class ML inference performance at the…