Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Engineer, AI Inference”. A match may be a passing mention rather than the job itself. Titles only.
2,336 roles · group by role · page 1 of 94
…Position Overview We are looking for a Software Engineer to work at the forefront of deploying our cutting-edge AI models, enhancing the performance…
We are seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme…
Help us push the boundaries of AI inference at NVIDIA — where your systems expertise shapes both the technology and the teams building on top…
We are seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme…
We are seeking highly skilled and motivated software engineers to join us and build AI inference systems that serve large-scale models with extreme…
…SOFTWARE ENGINEER, INFERENCE (AI DATA ENGINEERING) The application software team is the central nervous system of SpaceX – we create mission critical applications that are…
…Inferentia silicon and servers. Strong software development using Python, System level programming and ML knowledge are both critical to this role. Our engineers collaborate…
…Inferentia silicon and servers. Strong software development using Python, System level programming and ML knowledge are both critical to this role. Our engineers collaborate…
…Software Development Engineer on the Inference Model Enablement team, you will onboard and optimize state-of-the-art open-source and customer LLMs, both…
…NVIDIA's cutting-edge innovations in AI and high-performance computing. We are seeking a Senior Software Engineer to design, build, and optimize highly…
…This role is for a senior software engineer in the Machine Learning Inference Applications team. This role is responsible for development and performance optimization…
…inference stack to power Voice AI models - from product roadmap through engineering implementation. You’ll partner closely with Forward Deployed Engineers, Model Performance Engineers…
…This team sits at the intersection of AI infrastructure, distributed systems, compilers, runtimes, kernels, and hardware/software co-design. The Innovation Engine for Inference…
…Build, lead and scale world-class engineering teams in Vietnam, collaborating with global counterparts across system software, data science, and AI platforms. Drive the…
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators…
…research engineers through other technical leaders — developing tech lead managers or equivalent, setting direction across sub-teams — while staying technically influential yourself. - AI as…
…inference scaling across multi-node clusters using Ray Serve and Triton Experience in leading technical projects and supporting architectural decisions with data Software Engineering…
…Technical leadership skills and ability to mentor early-in-profession engineers. #MicrosoftAI Software Engineering IC5 - The typical base pay range for this role across…
About us Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to…
…experience reasoning about model routing, inference cost, and latency tradeoffs in production Strong software engineering fundamentals: distributed systems, API design, and testing discipline Comfort…
In this role, you will be a member of the AI Networking Software team and part of the bigger DC networking organization. The team…
…Experience with Machine Learning (ML) hardware accelerators and ML inference software. Preferred qualifications: PhD in Computer Engineering, Computer Science, or a related field. 2…
…world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10…
…seamlessly connect our multi-cloud and hybrid environments Collaborate with AI and software engineering teams to understand their needs, provide golden paths to production…
…Solve technically tests problems that exceed the scope of a generalist Software Engineers, specifically around optimizing Generative AI performance across heterogeneous hardware (CPUs, GPUs…