Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Staff ML Engineer, Generative Model Performance & Efficiency”. A match may be a passing mention rather than the job itself. Titles only.
25 roles
…that deliver exceptional performance and efficiency for inference workloads. You will work on cutting-edge architectural problems and performance modeling with deep cross-functional…
…Model performance increasingly depends on purpose-built datasets. We need ML-minded engineers who can collect, filter, and synthesize high-quality data at scale…
ABOUT LIQUID AI Spun out of MIT CSAIL, we build general-purpose AI systems that run efficiently across deployment targets, from data center accelerators…
…We are seeking a Research Engineer to join our Pre-training team, responsible for developing the next generation of large language models. In this…
…to enhance model performance. By combining research and engineering, you will bridge the gap between raw data and cutting-edge AI models, directly contributing…
…You’ll partner closely with engineering partners, product teams, and infrastructure stakeholders to design solutions that balance performance, cost-efficiency, and operational simplicity across…
…High performance, large-scale ML systems GPU/Accelerator programming ML framework internals OS internals Language modeling with transformers The annual compensation range for this…
…High performance, large-scale ML systems GPU/Accelerator programming ML framework internals OS internals Language modeling with transformers The annual compensation range for this…
…About the role As a Research Engineer on our team you will work end to end across the whole model stack, identifying and addressing…
…We're looking for a Software Engineer focused on Performance Optimization to help push the boundaries of speed and efficiency across our AI infrastructure…
…Build tooling to efficiently evaluate the effectiveness of novel LLM-generated jailbreaks. Write scripts and prompts to efficiently produce evaluation questions to test models…
…model quality Design, build, and run robust, efficient pipelines for model fine-tuning and evaluation Develop tools to measure and improve model performance across…
…Build tooling to efficiently evaluate the effectiveness of novel LLM-generated jailbreaks. Write scripts and prompts to efficiently produce evaluation questions to test models…
…ML) engineers, software engineers, and ML research engineers. We develop industry-leading simulation solutions using advanced generative and reconstructive ML algorithms, to model the…
…diagnosing performance issues, and identifying root causes. Demonstrated problem-solving ability and understanding of data engineering best practices to ensure reliable, efficient workflows. Solid…
…The team combines expertise in software engineering, machine learning, and low-level kernel design and development to design robust systems and enhance model performance…
…and optimising large multimodal models, and have experience building evaluations to measure their performance. - Are comfortable diving into complex ML codebases to identify and…
…Machine Learning and Generative AI solutions, enabling organizations to securely and cost-effectively own and host ML and Generative AI models, augmented or trained…
…company-wide platform for ML training, evaluation, and deployment - Experience with performance engineering and compute acceleration for large-scale ML training, including profiling, bottleneck…
…About the Role We're hiring a Staff Engineer to own major areas of the architecture of our Inference Cloud Platform. This team owns…
…and ML growth - Design inference infrastructure that keeps AI workloads fast, reliable, and cost-efficient - Build the foundational infrastructure for next-generation song creation…
…High performance, large-scale ML systems GPU/Accelerator programming ML framework internals OS internals Language modeling with transformers Representative projects: Implement low-latency high…
…to understand our model’s performance in the real world. - Design sampling algorithms to improve serving efficiency of large generative models. WHO YOU ARE…
…Model architecture design for Transformers or other large neural nets. Distributed systems / high‑performance computing for ML. Are comfortable working from algorithms to engines…
…models and deploy them efficiently in our robotaxi. Collaborate closely with x-functional teams, including ML researchers, software engineers, data engineers, and hardware engineers…