Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Engineer - Training/Inference (C++)”. A match may be a passing mention rather than the job itself. Titles only.
687 roles across 771 listings · show every listing · page 1 of 28
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…Develop, test, and maintain production software systems powering automated battery sorting, spanning ML inference, image acquisition, sensor integration, and hardware-adjacent control interfaces Train…
…Engineering, or a related field—or equivalent industry experience. Desirable Experience with multi-cloud orchestration, particularly in latency- or cost-sensitive training and inference…
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer…
…Optimize model performance for on-device use cases (memory, power, compute constrained environments). Engage directly with research, software engineering, hardware engineering, and product teams…
…post-training. - Experience with product experimentation, online evaluation, and A/B testing frameworks. - Strong software engineering skills with the ability to write clean, maintainable…
…A day in the life You work alongside customer engineering teams during live AI implementation sprints — debugging inference pipelines, optimizing RAG architectures, tuning agent…
…training/inference lifecycles, and optimization of model execution, or experience (non-internship) in professional software development - Cloud Technology Certification, or AWS Professional level certification…
…enabling cloud service providers with next-generation high-performance training and inference platforms. You will work at the intersection of hardware and software, validating…
…Minimum Qualifications: • Bachelor's degree in Computer Science, Engineering, Information Systems, or related field and 4+ years of Hardware Engineering, Software Engineering, Systems Engineering…
…training (SFT, RLHF/DPO), RAG over proprietary technical data, or multi-agent orchestration - Deep software engineering: C++ or Rust, developer-facing internal platforms, CI…
…post-training, and inference optimization. Experience leading research collaborations with engineering, product, and research teams, including architectural design, model evaluation, benchmark design, code reviews…
…About the Role We're hiring a Software Engineer to help contribute to projects on our Inference Platform team. Our team primarily owns the…
…Bonus Points - Domain knowledge in fraud, risk, or cybersecurity. - Background in Software Engineering - Familiarity with CI/CD, Docker, Kubernetes and the modern devops framework…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…Experience with building real time ML inference applications What Success Looks Like ML engineers can move from idea to experiment faster. Training and inference…
…Cloud Security Professional or Cloud DevOps Engineering), or Bachelor's degree - Experience with Machine Learning and Large Language Model fundamentals, including architecture, training/inference…
…record of handling challenges at scale. In this role, you’ll be working directly with architects, designers, verification engineers, software teams, and Physical Design…
…AWS Neuron is the software stack for Trainium and Inferentia, the AWS Machine Learning chips, delivering best-in-class ML performance in the cloud…
…cloud for AI training and inference? Want to do industry leading work delivering continuous price performance improvements in the cloud for AI model training…
…high performance computer that we are building for machine learning (ML) inference and training workloads. We are looking for accomplished engineers to set the…
…5+ years in machine learning engineering, backend software engineering, MLOps, or a closely related field Production ML service experience — deploying, serving, and operating models…
…As a Senior Software Engineer , you will design and build highly scalable systems to help creators to continuously deliver experiences that feel native and…