Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Solutions Architect, Inference Deployments”. A match may be a passing mention rather than the job itself. Titles only.
644 roles across 752 listings · show every listing · page 3 of 26
…APIs, tools, profiling, and debugging solutions Performance optimization and benchmarking Validation, quality, and release engineering Partner with silicon architecture and hardware engineering teams to…
…This critical role bridges advanced LLM research and practical deployment, involving the development of model architectures, improving training and inference efficiency, and collaborating on…
…Deep expertise in key data science domains (e.g., time-series forecasting, deep learning, causal inference). Extensive practical experience architecting solutions using a broad…
…model serving infrastructure, autoscaling based on inference load, and multi-model deployment patterns - Build internal tools and CLIs that let ML/AI teams deploy…
…AWS is looking for a GenAI/ML Solutions Architect who will be the Subject Matter Expert for helping customers in the United States design…
…We own the complete pipeline for our software, from gathering requirements to development and testing to deployment and availability. We are providing solutions to…
…from architecture design/training to post-training optimization and inference acceleration. You will contribute to our next generation of the Relational Foundation Model. What…
Storage Architect - Engineering Technologist Infrastructure Solutions Group (ISG) builds the products that power infrastructure, solutions, and data management our customers need most. Our teams…
…support development, deployment, and operations at scale. Architects, deploys, and operates secure cloud and container-based environments for training and inference, including GPU-intensive…
…You will own solutions end-to-end, from problem framing and data strategy to production deployment and measurement. You will remain hands-on while…
…You will architect and maintain Airbnb’s end-to-end traffic classification ML systems, balancing high-performance model deployment with rigorous offline data pipelines…
Scale GP is Scale's enterprise Generative AI platform—APIs and infrastructure for knowledge retrieval, inference, evaluation, and intelligent automation. We power mission-critical…
Meta is seeking a Systems Engineer to join our team working on AI/ML initiatives supporting large-scale AI training and inference. Our servers…
…You'll spend your time partnering with VPs and GMs to identify the highest-leverage data science opportunities, architecting novel ML solutions to complex…
…Experience in LLM efficiency research such as efficient attention, inference acceleration, or KV cache compression Experience in on-device AI deployment on mobile or…
…Job Responsibilities Build generative AI, agentic AI, and large language model solutions in Python from proof of concept through production deployment with measurable outcomes…
…scalable ML systems (batch and real-time inference), including data/feature pipelines, model training, evaluation, deployment, monitoring, and drift/performance management Identifies opportunities to…
…industry, focused on developing and supporting innovative ASIC solutions for Meta’s data center applications. Architect, develop, and optimize emulation model builds across Palladium…
…training/inference optimization, integration with cloud-native services, MLOps, etc. Serve as a trusted practitioner for enterprise GenAI solutions, including RAG architectures, agentic systems…
…alignment with company-wide goals - Drive architectural decisions and technical excellence across teams - Ensure robust systems for deployment, monitoring, scalability, and reliability of infrastructure…
…You'll lead technical initiatives at the intersection of data engineering and AI, architecting solutions that process large-scale data across Amazon's global…
…solutions in a secure, stable, and scalable way. As a core technical contributor, you are responsible for maintaining critical data pipelines and architectures across…
…deploy inference workloads. - Provide technical guidance to customers implementing and optimizing our cloud-based AI and ML solutions, including Kubernetes, model deployment, autoscaling, and…
…Founded in 2017 by industry pioneers from Stanford University, we are revolutionizing the rapidly growing AI inference market with our proprietary Dataflow Architecture. Our…
…systems, firmware, architecture, design, validation, product engineering). In this role, you will be developing cutting-edge next-generation RFICs for deployment in space and…