Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Engineering Manager, Inference Benchmarking — AI Perf”. A match may be a passing mention rather than the job itself. Titles only.
250 roles across 267 listings · show every listing · page 2 of 10
…the adoption and performance of our advertising platform through data-driven insights. You will work closely with product managers, engineers, and other data scientists…
…Build the benchmarking, observability, and regression-detection tooling that keeps performance from silently degrading as models and code evolve - Collaborate with engineers across functions…
…We are seeking an AI Engineer specializing in algorithm evaluation and agentic systems design for advanced computer vision and video understanding algorithms. What We…
…and manage high-performance GPU infrastructure for serving open-source LLMs. - Build evaluation frameworks (Evals) that enable engineers and data scientists to benchmark model…
…Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute…
…You will operate with significant autonomy, owning the scientific direction of your projects while collaborating with applied scientists, software engineers, product managers, technical, and…
…We value rigorous experimentation, strong data foundations, clear documentation, and production-ready engineering. Our team is helping enable AI-native manufacturing by turning fragmented…
…team expectations for validating AI outputs for correctness, performance, and security Strong understanding of responsible AI use in engineering workflows, including data sensitivity considerations…
…Engineer at JPMorganChase within the AI/ML Data Platform team, you will serve as the firm's deepest technical voice on LLM inference performance…
…Partner with Product and Engineering to turn analysis into shipped features—embedded analytics, benchmarks, insights, and AI-powered experiences—that deliver value directly to…
…Working knowledge of AI inference or training, finetuning concepts. Experience building, benchmarking, or operating AI platforms and LLM inference stacks. Experience developing AI agents…
…Identify and accelerate high-value workloads such as foundation model training, fine-tuning, speech AI, retrieval augmented generation, multimodal AI, inference optimization, and production…
…Strong knowledge of machine learning, deep learning, generative AI, real-time inference, data engineering, MLOps, and cloud-native systems. Experience with regulated environments, high…
…Experience with NVIDIA AI platforms, including CUDA, CUDA-X libraries, TensorRT-LLM, Triton Inference Server, NIM, NeMo, Megatron, Transformer Engine, NCCL, DGX, NVLink, InfiniBand…
…Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation. Fireworks AI is an equal-opportunity employer. We celebrate diversity…
…Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation. Fireworks AI is an equal-opportunity employer. We celebrate diversity…
…based AI networking solutions spanning TCP/IP, UDP, routing, congestion management, flow control, QoS, and traffic engineering. Analyze transport-layer behavior and performance characteristics…
…infrastructure underpinning Meta's AI and data center workloads at scale. As an ASIC Engineer specializing in architecture, performance and modeling, you will define…
…size, and real-time performance - Design model architectures informed by hardware constraints and inference requirements, working with inference engineers to ensure models are servable…
…AI. THE OPPORTUNITY We're looking for a Lifecycle Marketing Sr. Manager to build and scale the customer journeys that power Fireworks' growth engine…
…on ML infrastructure, inference systems, or AI platforms - Experience supporting enterprise or strategic customers - Prior experience in infrastructure, DevEx, or performance optimization initiatives - Experience…
The Prime Video Science team leverages the latest in machine learning and AI techniques combined with causal inference to bring scientific rigor to the…
…initiatives in GPU or AI infrastructure, distributed training or inference, model evaluation, or performance benchmarking. Experience architecting workflow engines, schedulers, experiment platforms, test frameworks…
…Manager to own the products that help customers extract the best possible performance from AI models and applications running on NVIDIA hardware. Every inference…
…size, and real time performance. - Design model architectures informed by hardware constraints and inference requirements, working with inference engineers to ensure models are servable…