Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Performance Engineer (Inference, Training & GPU)”. A match may be a passing mention rather than the job itself. Titles only.
423 roles across 455 listings · show every listing · page 1 of 17
…we index on inference, serving, GPU optimization, and training performance. Distributed-systems breadth is welcome, but secondary. Strong performance-engineering foundations: profiling, roofline analysis…
…Develop and productize inference models on NVIDIA GPUs and Nvidia RTX Spark platforms as SDK/Microservices after optimization. Technical Mentorship: Mentor senior engineers and…
…Development experience with NVIDIA software libraries and GPUs, including CUDA and CUDA-X libraries. Experience with Kubernetes, distributed training, and large-scale inference. Experience…
ABOUT THE ROLE We are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…Our workloads span thousands of GPUs, petabytes of driving data, and geographically distributed training and inference clusters. As Architect for AI Infrastructure, you will…
…As the Engineering Manager for this team, you will lead a group focused on model optimization, training efficiency, GPU enablement, load testing, model performance…
…Develop robust evaluation frameworks and quality signals to measure real-world model performance. - Systems & Infrastructure: Lead the design of efficient training and inference systems…
…A day in the life You work alongside customer engineering teams during live AI implementation sprints — debugging inference pipelines, optimizing RAG architectures, tuning agent…
…generation high-performance training and inference platforms. You will work at the intersection of hardware and software, validating stable and performant technical solutions from…
…Translate customer needs into clear product requirements for rendering quality, sensor fidelity, performance, scalability, usability, and workflow integration. Partner with engineering teams on path…
…into business outcomes (cost efficiency, performance, scalability) Partner with solutions architects to design high-performance architectures (AI training, inference, HPC) Build ROI/TCO models…
…Our Industrial Compute organization develops and deploys large-scale AI campuses designed to support the next generation of frontier model training and inference workloads…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…of ML training and inference workloads. Develop tooling that helps ML engineers debug, profile, optimize, and monitor model performance. Improve GPU and general resource…
…or Cloud DevOps Engineering), or Bachelor's degree - Experience with Machine Learning and Large Language Model fundamentals, including architecture, training/inference lifecycles, and optimization…
…fractional GPUs (MIG, MPS, time-slicing), GPU scheduling, training vs. inference fleets, and multi-tenant GPU isolation. --- The ideal candidate is a strategic thinker…
…g., scientists, product managers, data engineers) to create enterprise-scale AI/ML systems that handle high-volume inference workloads, implement comprehensive model and AI…
…GPU architectures (NVIDIA A100/H100/B200, AWS Trainium/Inferentia), NVLink, EFA networking, storage hierarchies (FSx for Lustre, S3), and how they interact at scale…
…Required qualifications, capabilities, and skills Formal training or certification on software engineering concepts and 5+ years applied experience Hands-on experience in system design…
…scale, high-performance C++ software architecture for mission-critical systems or ML inference. Proven track record of leading complex, cross-functional engineering projects as…
…Learning Engineer focused on research enablement and performance, turning promising experiments into stable, scalable, user-facing capabilities while making training and inference faster, cheaper…
…Bonus Points: - Experience working with teams building platforms or services for AI inference and/or training. - Direct experience governing model onboarding programs across GPU…
…data quality, training efficiency, model architecture, or inference performance - Research depth plus engineering rigor conducting frontier research and builds systems others depend on; doesn…