Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Staff ML Engineer, Generative Model Performance & Efficiency”. A match may be a passing mention rather than the job itself. Titles only.
166 roles across 176 listings · show every listing · page 1 of 7
…Mentorship & Career Growth We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll…
…This team partners with application teams to deliver ML algorithms and models that drive efficiency, innovation, and growth. They identify business needs and define…
…the latest AI/ML models. Veza can only do what its integrations enable; you will work closely with all other engineering teams as you…
…measure real-world model performance. - Systems & Infrastructure: Lead the design of efficient training and inference systems for large-scale generative models. Architect scalable data…
…About the Role We're hiring a Staff Engineer to help lead, drive, and contribute to projects on our Inference Platform team. Our team…
…efficiency of ML training and inference workloads. Develop tooling that helps ML engineers debug, profile, optimize, and monitor model performance. Improve GPU and general…
…Mentor Staff+ Engineers: Act as a force multiplier by coaching the next generation of technical leaders and influencing company-wide engineering standards. Basic Qualifications…
…kernel performance tuning, or eBPF. AI/ML Infrastructure: Experience building or managing compute platforms tailored for GPU scheduling and large-scale model training. Community…
…As a Staff ML Engineer , you will be at the forefront of Physical AI , building advanced autonomy algorithms and models to add rich semantics…
…Drive alignment across engineering, product, and executive stakeholders while raising the performance bar by mentoring Staff and Senior engineers. Roles & Responsibilities Strategic Technology Direction…
…of model’s data generation subsystem between model training data preparation and model inference. Develop Simulation solutions to enable robust testing and efficient evaluation…
…modal AI model infrastructure (LLMs, generative models, video/image/speech models) - Background in building infra for multi-tenant SaaS, enterprise AI/ML platforms, or…
…Working alongside a tight-knit team of researchers and production engineers, your work will directly define how the next generation of AI comprehends the…
…Work across the full modeling lifecycle: problem formulation, feature engineering, training, calibration, deployment, monitoring, and iteration in production. Build agentic engineering workflows that accelerate…
…qualifications Deep background in systems engineering or ML infrastructure, with the ability to go hands-on with performance profiling, latency and throughput optimization, and…
…new generations of silicon and compute to create an outsized impact on accelerating several customer workloads including AI/ML, core computing, High Performance Computing…
…processing pipelines for our generative AI platform. You'll architect scalable, resilient backend infrastructure, lead technical design discussions, mentor engineers, and establish best practices…
…generation infrastructure that powers breakthrough innovation in AI/ML and HPC workloads. If you’re passionate about pushing the limits of performance, efficiency, and…
…generation infrastructure that powers breakthrough innovation in AI/ML and HPC workloads. If you’re passionate about pushing the limits of performance, efficiency, and…
…Agent / ML Software Engineer is a Senior Staff-level, hands-on technical leadership role responsible for defining, building, and operating next-generation AI systems…
…Deep understanding of the security of AI and ML models, agents, and associated systems. BONUS POINTS IF… - Security Research: Proven experience contributing to or…
…designing, deploying, and operating AI/ML systems in production Solid understanding of Generative AI architectures, including transformers, diffusion models, and hybrid systems (LLMs, LVMs…
…Our Technology Stack Our engineering team works with a modern tech stack designed for scalability, performance, and developer efficiency: Frontend: React.js with Redux…
…Push GPU efficiency and training performance, raising utilization (such as model FLOPs utilization and end-to-end throughput) and lowering cost per training run…
…engineers like you to (1) develop methods for efficiently and continuously learning from large scale real-world data, to (2) develop models and model…