Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “ML Engineer - Automated Evaluation and Adversarial Design”. A match may be a passing mention rather than the job itself. Titles only.
25 roles across 29 listings · show every listing
…This role focuses on building and scaling automated evaluation systems and designing adversarial and stress-testing methodologies across multiple AI features. The work requires…
…automated decisions. Expertise in heterogeneous inference platforms supporting LLMs, SLMs, wide & deep models, ensembles, graph models, classical ML models, heuristics, and rules engines. Experience…
…Deep background in AI safety and red-teaming , including hands-on experience with adversarial ML, prompt injection defense strategies, and automated evaluation suites for…
…process controls by technical and social engineering attacks, and preparation of deliverables for senior stakeholders Design and execute testing and simulations – such as penetration…
…Drive engineering excellence through code reviews, design reviews, test strategy, deployment automation, incident analysis, documentation, and AI-assisted development practices using tools such as…
…Sitting at the intersection of applied ML research and engineering, you'll design experiments to improve how we evaluate both model behavior and the…
…We are looking for an exceptional ML Engineer to help us build the next generation of scalable evaluation infrastructure and lead rigorous investigations into…
…The Role As a Staff Applied Machine Learning Engineer focused on Fraud & Abuse, you will design, build, and operate production ML decision systems that…
…research, or AI/ML security. Experience building security tooling and automation (scripts, scanners, detection logic, or evaluation harnesses) that other engineers actually use. Strong…
…You think about scale from the start. - Genuine curiosity about AI security model supply chains, prompt data handling, adversarial ML, and the governance frameworks…
…AI/ML models into existing intelligence platforms • Develop end-to-end solutions including data pipelines, feature engineering, model training, evaluation frameworks, and production monitoring…
…a machine learning engineer for Stripe Capital, you'll be responsible for designing, building, training, evaluating, deploying, and owning ML models in production with…
…architecture, reliability, monitoring, performance, and operational excellence. - Hire and grow an exceptional team across backend, data systems, and applied ML engineering domains as needed…
…Cloud security engineering across one or more major providers; IaC and policy-as-code. Experience or exposure to Cyber operations, Adversarial ML and LLM…
…and agent security, adversarial testing, model evaluation, cyber-defense automation, vulnerability discovery, secure deployment, or autonomous response. Translate research into usable outcomes for engineering…
…automated "exploit-as-code" validator over performing the same manual test twice. You can architect evaluation harnesses and adversarial test suites for ML models…
…fraud tooling and automation, driving enhancements to improve efficiency and scale Product & Model Partnership - Collaborate with Data Science, ML/AI, and Product teams to…
…Our team, part of Apple Services Engineering, is looking for an ML Research Engineer to lead the design and continuous development of automated safety…
…Lead proactive identification of risks, failure modes, and adversarial attack vectors across AI systems - designing structured red-teaming exercises and evaluation frameworks before and…
…establishing SLAs, and instrumenting telemetry, alerting, and feedback loops. - Develop and automate tools for model evaluation, stress testing, backtesting, and adversarial scenario simulation to…
…Transformers - Familiarity with model deployment workflows, evaluation techniques, or MLOps concepts - Understanding of responsible AI development and basic AI safety considerations PREFERRED QUALIFICATIONS - Experience…
…schema design that LLMs actually follow reliably across providers Evaluation & Harness - Own the eval framework end to end: ground truth datasets, automated scoring pipelines…
…AI-First approach to improve vulnerability detection and security automation. - Partner across Product, Security Research, and Engineering to introduce AI capabilities into the broader…
…vulnerability detection and security automation. - Implement features for AI red teaming agents and evaluation tools that help identify risks in LLMs and applied AI…
…Analyze and evaluate the ongoing performance of developed ML systems. Collaborate with multiple partner teams, such as Business, Technology, Product Management, Design, Analytics, and…