Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Solutions Architect, Inference Deployments”. A match may be a passing mention rather than the job itself. Titles only.
131 roles across 139 listings · show every listing · page 1 of 6
…Deep expertise in key data science domains (e.g., time-series forecasting, deep learning, causal inference). Extensive practical experience architecting solutions using a broad…
…model serving infrastructure, autoscaling based on inference load, and multi-model deployment patterns - Build internal tools and CLIs that let ML/AI teams deploy…
…AWS is looking for a GenAI/ML Solutions Architect who will be the Subject Matter Expert for helping customers in the United States design…
…We own the complete pipeline for our software, from gathering requirements to development and testing to deployment and availability. We are providing solutions to…
…from architecture design/training to post-training optimization and inference acceleration. You will contribute to our next generation of the Relational Foundation Model. What…
Storage Architect - Engineering Technologist Infrastructure Solutions Group (ISG) builds the products that power infrastructure, solutions, and data management our customers need most. Our teams…
…Risk, and Finance to deliver AI and ML solutions from ideation through production deployment, ensuring solutions are scalable, responsible, and aligned to business needs…
…support development, deployment, and operations at scale. Architects, deploys, and operates secure cloud and container-based environments for training and inference, including GPU-intensive…
…You will own solutions end-to-end, from problem framing and data strategy to production deployment and measurement. You will remain hands-on while…
…You will architect and maintain Airbnb’s end-to-end traffic classification ML systems, balancing high-performance model deployment with rigorous offline data pipelines…
Scale GP is Scale's enterprise Generative AI platform—APIs and infrastructure for knowledge retrieval, inference, evaluation, and intelligent automation. We power mission-critical…
Meta is seeking a Systems Engineer to join our team working on AI/ML initiatives supporting large-scale AI training and inference. Our servers…
…You'll spend your time partnering with VPs and GMs to identify the highest-leverage data science opportunities, architecting novel ML solutions to complex…
…Experience in LLM efficiency research such as efficient attention, inference acceleration, or KV cache compression Experience in on-device AI deployment on mobile or…
…Job Responsibilities Build generative AI, agentic AI, and large language model solutions in Python from proof of concept through production deployment with measurable outcomes…
…scalable ML systems (batch and real-time inference), including data/feature pipelines, model training, evaluation, deployment, monitoring, and drift/performance management Identifies opportunities to…
…industry, focused on developing and supporting innovative ASIC solutions for Meta’s data center applications. Architect, develop, and optimize emulation model builds across Palladium…
…training/inference optimization, integration with cloud-native services, MLOps, etc. Serve as a trusted practitioner for enterprise GenAI solutions, including RAG architectures, agentic systems…
…alignment with company-wide goals - Drive architectural decisions and technical excellence across teams - Ensure robust systems for deployment, monitoring, scalability, and reliability of infrastructure…
…You'll lead technical initiatives at the intersection of data engineering and AI, architecting solutions that process large-scale data across Amazon's global…
…solutions in a secure, stable, and scalable way. As a core technical contributor, you are responsible for maintaining critical data pipelines and architectures across…
…deploy inference workloads. - Provide technical guidance to customers implementing and optimizing our cloud-based AI and ML solutions, including Kubernetes, model deployment, autoscaling, and…
…Founded in 2017 by industry pioneers from Stanford University, we are revolutionizing the rapidly growing AI inference market with our proprietary Dataflow Architecture. Our…
…systems, firmware, architecture, design, validation, product engineering). In this role, you will be developing cutting-edge next-generation RFICs for deployment in space and…
…As a Solution Specialist for Data Services, you are the person who opens that frontier. You identify new market opportunities where data architecture, mobility…