Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “ML Engineer, Inference & Optimization”. A match may be a passing mention rather than the job itself. Titles only.
64 roles across 65 listings · show every listing · page 1 of 3
…Responsibilities Build and optimize global and local request routing, ensuring low-latency load balancing across data centers and model engine pods. Develop auto-scaling…
…AWS Neuron is the software of Trainium and Inferentia, the AWS Machine Learning chips. Inferentia delivers best-in-class ML inference performance at the…
…Working at the hardware-software boundary, our engineers craft high-performance kernels for ML functions, ensuring every FLOP counts in delivering optimal performance for…
…Our first products are heavily focused on inference. Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is…
…Our first products are heavily focused on inference. Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is…
…WHAT ARE WE LOOKING FOR? - Previous founding or startup experience - Experience optimizing ML inference or engineering systems for research teams - Fluency in Python and…
…Understanding of model internals, inference pipelines, evaluation techniques, and prompt engineering Ability to thrive in ambiguous, fast-changing spaces and have a product-oriented…
…structures You are proficient in C++ programming Experience in training and inferencing ML models. Experience with robotics Experience with ETA modeling At Nuro, your…
…About the Role We are seeking a versatile and experienced engineer to join our Inference Core Model Bringup team. This team is responsible to…
…the-baseten-inference-stack/ - Driving model performance optimization https://www.baseten.co/blog/driving-model-performance-optimization-2024-highlights/ RESPONSIBILITIES Core Engineering Responsibilities - Design…
…for connected fleets - Familiarity with the NVIDIA Jetson ecosystem, including optimizing ML inference and other GPU workloads on embedded compute (e.g., CUDA, TensorRT…
…needed) Evaluation & Inference: Implement algorithms and software to analyse and evaluate the performance of AI models. Optimising performance of AI/ML models such as…
…scale AI training and inference. Your work will range from prototyping system software on new accelerators to enabling performance optimizations across our AI workloads…
…and distributed inference platforms to large-scale data pipelines, interactive analytics, and advanced developer tooling. We collaborate closely with hardware, robotics, ML, design, and…
…inference, and data processing pipelines. - Lead technical design discussions, mentor other engineers, and establish best practices for building and operating large-scale ML infrastructure…
…Responsibilities - Tackle embedding models and retrieval systems optimized for grounding, relevance, and adaptive reasoning. - Collaborate with a team of researchers and engineers building end…
…Work cross-functionally with engineers across behavior, perception, mapping, and ML research to advance the state-of-the-art in autonomous driving. Implement practical…
…Partner with infrastructure engineers to develop and optimize systems for training, inference, monitoring, and deployment. Explore new ideas at the edge of what’s…
…datasets and optimize data pipelines for model training and evaluation. - Work closely with the model serving team to ensure that inference is fast and…
…Kernel Engineer, you'll be responsible for identifying and addressing performance issues across many different ML systems, including research, training, and inference. A significant…
…In this role, you will: - Design and implement inference infrastructure for large-scale multimodal models. - Optimize systems for high-throughput, low-latency delivery of…
…The Distributed Training team works side by side with chip architects, compiler engineers and runtime engineers to create, build and tune distributed training solutions…
…Act as a virtual member of CoreWeave's Security product and engineering teams, identifying opportunities for product enhancement and collaborating with engineers to implement…
…Act as a virtual member of CoreWeave's Storage product and engineering teams, identifying opportunities for product enhancement and collaborating with engineers to implement…
…AI/ML applications and associated technologies such as Infiniband and NVIDIA Collective Communications Library (NCCL) Preferred: Code contributions to open-source inference frameworks Experience…