Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Senior Applied Scientist - AI Evaluation & Quality Systems”. A match may be a passing mention rather than the job itself. Titles only.
98 roles across 108 listings · show every listing · page 1 of 4
…develop the AI agents and agentic capabilities that automate trust decisions, and create the benchmarks and evaluation harnesses that keep decision quality high as…
…data scientists and AI engineers. You'll focus on how practitioners build, evaluate, govern, and ship agentic systems on the Databricks Data + AI Platform…
…business context, evaluate tradeoffs, and align analytical projects with organizational goals. Experience mentoring and guiding junior team members, overseeing project quality, and investigating root…
…end RAG systems for specific product use cases: chunking strategies, embedding model selection, retrieval optimisation, and quality evaluation. - Own the full AI-product lifecycle…
…Familiarity with evaluating agentic or LLM-based systems (e.g., decision-quality measurement, human-in-the-loop calibration) is a plus. Your Location: This…
…to correlation-based detection, automated response, and agentic AI. The team is composed of applied scientists, software engineers, and security engineers working across physical…
…senior scientists to deliver machine learning and generative AI products that carry real degrees of ambiguity, scale, and complexity. - Design, build, and evaluate agentic…
…with applied scientists to turn research prototypes into production capabilities, mentor engineers, lead design reviews, and set the standard for code quality across the…
…Develops and maintains production-grade services, APIs, SDK integrations, and workflows that support model training, serving, evaluation pipelines, and AI application lifecycle management. Partners…
…Experience applying responsible AI practices to model and agent development, including evaluation for safety, reliability, privacy, security, bias, groundedness, and misuse risks. Have publications…
…As a Senior Data Scientist on the AI Research & Reliability team, you will be a driving force in shaping the future of Gong AI…
…training, evaluation, optimization, and production deployment. - Work closely with engineering teams to integrate models into real-time systems, ensuring reliability, uptime, and quality at…
…You'll work closely with senior scientists and engineers, contributing to model development, evaluation, and optimization. Key job responsibilities * Implement machine learning models (classical…
…Define operational standards, evaluation frameworks, and production readiness criteria for AI systems across your organization. Establish mechanisms that ensure AI systems remain healthy across…
…maintaining systems that collect, store, process, govern, and analyze large volumes of data required by analysts, data scientists, and AI/ML applications. It serves…
…experiments and rigorous evaluations, synthesize and communicate results, and deliver well-tested, high-quality code. Contribute to high-impact business applications, reusable assets and…
…around technical system architecture, product design, and software development cycles for their area of ownership. AI Fluency: Ability to understand AI/ML capabilities, limitations…
…a Senior Software Engineer on the ML Infrastructure team, you will design and implement the core backend and infrastructure powering our AI systems. You…
…The goal is that every AI practitioner on SageMaker can produce tailored, high-quality datasets for model training, fine-tuning, and evaluation without needing…
…pipelines, retrieval systems, and evaluation workflows Participate in experimentation and innovation initiatives, such as prototyping new approaches and applying emerging AI techniques to business…
…Experience designing, deploying, and operating AI or LLM-based solutions, including prompt engineering, retrieval-augmented approaches, model tuning, evaluation, data quality, performance monitoring, and…
…Proactively aims to reduce success expectations and improve delivery margin without impact to quality or customer experience and understands impact of decisions on the…
…to define product features, system architecture, and best practices that enable a quality product * Work closely with external customers, scientists, and other engineering teams…
…for large language model applications Conduct applied research by studying scientific articles and current techniques in prompting, fine-tuning, evaluation, and agent design, then…
…trigger-based analytics Apply strong data management and governance practices to ensure analytics are reliable, auditable, and reproducible Implement quality controls for analytic outputs…