Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Solutions Architect, Inference Deployments”. A match may be a passing mention rather than the job itself. Titles only.
644 roles across 752 listings · show every listing · page 4 of 26
…As a Solution Specialist for Data Services, you are the person who opens that frontier. You identify new market opportunities where data architecture, mobility…
…Applied AI Solutions Architects and Customer Success Specialists to design, build, and deploy AI solutions in customer environments during fixed deployment cycles. You will…
…Design and build deep learning models for computer vision, audio understanding, and multimodal semantic fusion — including architectures that enable joint reasoning across visual, auditory…
…You'll work closely with software engineers to take your models from experimentation through production deployment at scale. If you're excited about applying…
…edge infrastructure, enterprise deployment, and competing platform ecosystems. Cross-Functional Product Collaboration Partner closely with software engineering, hardware engineering, architecture, program management, business development…
…transformer architectures, OCR technologies, and multimodal document understanding models. This role involves managing the full ML lifecycle, from prototyping to production deployment on AWS…
…model deployments and inference Strong DevOps background with complex CI/CD pipelines, infrastructure automation, and deployment strategies Self-directed: you make architectural decisions and…
…inference workloads across hybrid and multi-cloud environments Familiarity with edge computing, high-risk deployment environments, or regulated infrastructure patterns Experience establishing architectural governance…
…augmented systems; Optimizing and scaling AI models for real-time inference, low-latency deployment, and cost-effective performance in cloud environments including AWS, Azure…
…deployment, monitoring, and lifecycle management. - Partner cross-functionally with Data Science, AI, Product, and Engineering to translate ambiguous requirements into robust platform solutions. - Establish…
…models and agents Inference, routing, orchestration, and policy enforcement systems Evaluation, red teaming, and monitoring infrastructure for AI systems Deployment automation, CI/CD, and…
…Push the boundaries of models and agentic architectures by designing novel approaches to catalog understanding, schema inference, where the problem complexity (billions of products…
…implement scalable model architectures optimized for strict latency constraints, including knowledge distillation, quantization, and efficient inference strategies for production deployment. - Lead end-to-end…
…architectures and types of NVIDIA accelerators. Contribute features and code to NVIDIA’s inference libraries, vLLM and SGLang, FlashInfer and LLM software solutions. Work…
…s solutions set the industry standard for performance and availability. Full Software and System Lifecycle: From ideation to architecture, design, development, deployment, operations, and…
…Hands-on experience with model compression, quantization, and real-time inference optimization for production deployments. Previous experience implementing AI solutions in physical settings such…
…10+ years of experience in strategic partnerships, enterprise technology, solution architecture, technical sales, business development, product management, or developer ecosystems. Strong knowledge of enterprise…
…technology, data science, solution architecture, or developer ecosystems. Strong knowledge of machine learning, deep learning, generative AI, real-time inference, data engineering, MLOps, and…
…Partner with Research Account Managers, Solution Architects, Product, Engineering, and Business Development teams to support researcher adoption and long-term engagement. Represent researcher needs…
…What you’ll achieve Lead the architecture, development, and deployment of enterprise scale ML solutions across Dell’s global ecosystem. Drive MLOps standards, build…
…reliable deployment of frontier models across the company. Lead AI Infrastructure Vision: Architect end-to-end training, fine-tuning, and low-latency inference platforms…
…This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This…
…Understanding of the full ML development cycle, from data collection and training to deployment, onboard inference considerations, and data iteration loops. Strong problem-solving…
…Experience with data center architecture and deployment Experience working with ODMs/vendors Familiar with Machine Learning training or inference silicon and systems Work experience…
…Collaborate with silicon, system software, firmware, hardware, and Azure infrastructure teams to deliver scalable networking solutions from concept through datacenter deployment. Participate in architecture…