Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Solutions Architect (Inference)”. A match may be a passing mention rather than the job itself. Titles only.
1,184 roles across 1,384 listings · show every listing · page 1 of 48
About the Role As a Solutions Architect (Inference) at Together AI, you will work with customers and prospects to create business value through Generative…
…Enabling a workload aware network for routing, securing and load balancing inference. As part of the Telco Solution Architecture team, you’ll be immersed…
…As a Solutions Architect focused on inference, you’ll collaborate closely with our engineering, DevOps, and customers to develop enterprise AI solutions. Together, we…
…You will solve complex architecture problems with solutions that are extensible and scale, yet always look for ways to simplify. You will work with…
…About the team The Cloud Platform team owns the production inferencing service that serves SambaNova's models to customers on RDU accelerators, including capacity…
…Knowledge of ML converters/compilers and runtimes, and hardware-accelerated ML inference techniques. Understanding of Generative AI model architectures and their optimization for on…
…ML compilers, production coding agents, GenAI model architecture, model training, neural network optimization, or alternatively applied math. - Passion for customer experience and usability, including…
…with server design and the knowledge of various teams to architect the solutions that we will deploy at scale. To deliver your products you…
…architecture discussions with Principal Engineers, SDMs, and Principal Scientists, identifying dependencies, scaling factors, boundary conditions, and risks across horizontal distributed systems serving AI inference…
…equivalent - Knowledge of Machine Learning and LLM fundamentals, including transformer architecture, training/inference lifecycles, and optimization techniques Our inclusive culture empowers Amazonians to deliver…
…The Qualcomm Cloud AI team is developing hardware and software solutions for Inference Acceleration. We are hiring LLM Serving Engineers at multiple levels to…
…The Qualcomm Cloud AI team is developing hardware and software solutions for Inference Acceleration. We are hiring an AI Performance Engineers at multiple levels…
…Key Responsibilities · Architect and deliver model optimization strategies that transform PyTorch models for efficient inference on Qualcomm accelerators. · Drive graph capture and deployment using…
…Job Description The Qualcomm Cloud AI team is developing software solutions for Inference Acceleration. We are seeking an ambitious, bright and innovative engineer who…
…Partner with product, architecture, modeling, and engineering to design robust solutions that power our Digital channels Required qualifications, capabilities, and skills BS in Computer…
…We're responsible for guiding products throughout the execution cycle, focusing specifically on analyzing, positioning, packaging, promoting, and tailoring our solutions to our users…
…Building on more than 20 years of electro-optical and infrared systems innovation, Optical Systems delivers solutions to the warfighter for responsive, scalable sensing…
…Building on more than 20 years of electro-optical and infrared systems innovation, Optical Systems delivers solutions to the warfighter for responsive, scalable sensing…
…and secure AI/ML product architectures , including trust boundaries, API boundaries, and data flow through training and inference pipelines Build secure automation through Python…
…technical solutions Apply and help refine architectural standards, patterns, and best practices within the infrastructure software team Partner with senior engineers and architects on…
…teams to translate infrastructure requirements into sound technical solutions Contribute to and help evolve architectural standards, patterns, and best practices across the infrastructure software…
…Familiarity with real-time inference systems and the architectural considerations of serving ML models in low-latency, high-throughput environments. Prior experience in a…
…You'll have the backing of Toast's scale, brand, and resources but you'll be on the team architecting how we operate, scale…
…inference) and APIs, enterprise data grounding via Graph/D365/connectors. Azure Fabric & Azure AI Foundry — One Lake-based unified data management, Lakehouse architecture for…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…