Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Engineering Manager, Inference Benchmarking — AI Perf”. A match may be a passing mention rather than the job itself. Titles only.
127 roles across 134 listings · show every listing · page 1 of 6
…NVIDIA’s open-source benchmarking platform, AIPerf, is the growing standard for assessing LLM serving performance across various inference frameworks. Hyperscalers, cloud providers, and…
…As the Engineering Manager for this team, you will lead a group focused on model optimization, training efficiency, GPU enablement, load testing, model performance…
…Running perf benchmarks for both training and inference. Collaborate with AE, FAE, and Solution Architect teams on validation for customer issues and technical documentation…
…Working knowledge of DL inference pipelines and on-device performance profiling. Ability to collaborate effectively across a globally distributed, multi-timezone engineering organization. Level…
…Develop tooling that helps ML engineers debug, profile, optimize, and monitor model performance. Improve GPU and general resource utilization through scheduling, resource management, caching…
…Ensuring the reliability, scalability, and performance of the ML systems by writing automated tests, monitoring performance, and implementing best practices for model management. Participating…
…The Principal Product Manager, Marketing Technology AI Opportunity Okta is building the next generation of AI-native marketing technology, transforming how Demand Generation, Marketing…
…cost optimization for training and inference - Provide deep technical guidance on inference optimization model serving architectures (self-managed on EKS, SageMaker endpoints, Sagemaker Hyperpod…
…About the team The Rufus Features Science team, based in London, works alongside ~150 engineers, designers and product managers, shaping the future of AI…
…team, based in London, works alongside ~150 engineers, designers and product managers, shaping the future of AI-driven shopping experiences at Amazon. The team…
…pivotal performance engineering group at the intersection of state of the art ML research and production-scale infrastructure. You'll build and manage a…
…frameworks Inference & Performance Optimization Optimize model inference for latency, throughput, and cost Implement advanced techniques such as caching, quantization, batching, and routing Benchmark and…
…Network Engineering, Validation and Reliability Oversee lab validation, scale testing, performance benchmarking, and failure scenario analysis. Lead root cause analysis for complex network incidents…
…on production traffic patterns. ➢ Engineering High-Performance Model Serving and Inference Infrastructure (25% of time) Designing multi-stage inference pipelines that handle both real…
…You won't just manage; you'll architect and guide a brilliant team of engineers who are pushing the performance of LLM inference. Your…
…engines, databases, analytics platforms, data processing frameworks, and AI data infrastructure. Partner with ISVs on discovery, architecture reviews, technical deep dives, POCs, benchmarks, demos…
…matching frontier performance through training and harness engineering. (blog https://fireworks.ai/blog/open-source-agents-frontier-advisors) - The fine-tuning bottleneck is not…
…inference, and data processing pipelines for our generative AI platform. You'll architect scalable, resilient backend infrastructure, lead technical design discussions, mentor engineers, and…
…in AI. Just like frontier open-source inference enables sustainable unit economics for our customers, we are building a sustainable growth engine to drive…
…The Field Engineering organization at CoreWeave is dedicated to ensuring every customer running AI workloads at scale has a seamless, reliable, and high-performance…
…AI inference serving, performance optimization, and scalability across heterogeneous hardware Experience with MLOps practices for AI application development, deployment, monitoring, and lifecycle management Familiarity…
Oracle is seeking a Principal Software Engineer (IC4) to join our AI Networking team and help build the high-performance communication infrastructure that powers…
…matching frontier performance through training and harness engineering. (blog https://fireworks.ai/blog/open-source-agents-frontier-advisors) - The fine-tuning bottleneck is not…
…benchmarks, and leaderboard communities across both general and vertical domains; as well as the associated software ecosystems, performance acceleratio n techniques, and AI-driven…
…product and engineering with specificity and urgency. - Help customers get live quickly by coordinating onboarding, benchmarking, and integration efforts. Deal Management & Commercial Execution - Lead…