Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Performance Engineer, Inference Systems”. A match may be a passing mention rather than the job itself. Titles only.
469 roles across 507 listings · show every listing · page 1 of 19
…The Inferentia chip delivers best-in-class ML inference performance at the lowest cost in cloud. Trainium will deliver the best-in-class ML…
…6+ years of experience in AI infrastructure, systems engineering, high-performance computing, networking, site reliability engineering, or a related technical role. Deep understanding of…
…Engineering Group, Engineering Group > Systems Test Engineering General Summary: We are looking for a highly motivated Systems Test Engineer with 4+ years of hands…
…As a Software Engineer who has deep systems thinking to design, build, and enhance scalable and highly concurrent ML and AI serving platform. Knowledge…
…Architect High-Performance Inference Systems: Design, optimize, and deploy enterprise-scale LLM serving infrastructures. You will push the boundaries of throughput and latency. Own…
…door to work for all business systems. By combining ServiceNow’s leading workflow automation with Moveworks’ Reasoning Engine and natural language capabilities, we deliver…
…We are the first inference-focused frontier AI system. Our addressable market is the entirety of inference, unlike many of our competitors. We are…
…engineering who have led cross-functional hardware programs through NPI and production. - Experience managing manufacturing readiness for complex electromechanical systems or high-performance compute…
…Partner with Revenue Operations, Sales, Finance, HR, Payroll, and Business Systems to operationalize compensation plans, troubleshoot data or system issues, and ensure accurate, timely…
…identifiers, personal records, professional or employment information, and inferences drawn from your PI. We collect your PI for our purposes, including performing services and…
…the functionality and performance of kernel libraries. - Collaborate with kernel, compiler, performance, and hardware engineers to improve software and system performance. - Study emerging machine…
…Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference…
…This role focuses on advancing the science and systems behind ML measurement, feature understanding, and causal inference at scale. The work spans areas such…
…Job Summary Every model release presents a new opportunity to push the frontier on kernel engineering. Future performance breakthroughs will come from AI systems…
…Our Analytics Engineers own the data end to end — ingesting from source systems, modeling through a medallion architecture on BigQuery, and producing the trusted…
…establish engineering standards. - Drive support for emerging LLM architectures and inference workloads. Team Leadership - Hire, mentor, and grow a high-performing engineering team. - Develop…
…a formal engineering technical leadership role, leading software engineering teams. 2 years of experience in LLM training or inference, including performance optimizations, distributed execution…
…engineering management role, guiding software engineering teams, including hiring or team development. 2 years of experience in LLM training or inference, including performance optimizations…
…bringing high-performance intelligence to the edge. You will have an opportunity to drive distributed inference technology that powers real-time systems where latency…
…bringing high-performance intelligence to the edge. You will have an opportunity to drive distributed inference technology that powers real-time systems where latency…
…Expert-level software engineering fundamentals using Python, PyTorch, or JAX, with a track record of building reliable, highly scalable ML systems. Proven ability to…
…We are the first inference-focused frontier AI system. Our addressable market is the entirety of inference, unlike many of our competitors. We are…
…Integrate and enable new machine learning models into the existing platform or client environments. - Performance Optimizations: Improve system performance, efficiency, and scalability of deployed…
…matching frontier performance through training and harness engineering. (blog https://fireworks.ai/blog/open-source-agents-frontier-advisors) - The fine-tuning bottleneck is not…
…Integrate and enable new machine learning models into the existing platform or client environments. - Performance Optimizations: Improve system performance, efficiency, and scalability of deployed…