Senior AI Scientist

• Posted 3 days ago • Updated 3 days ago
Full Time
On-site
Fitment

Dice Job Match Score™

📋 Comparing job requirements...

Job Details

Skills

  • Generative Artificial Intelligence (AI)
  • Analytical Skill
  • Mentorship
  • Data Science
  • Large Language Models (LLMs)
  • Evaluation
  • Management
  • Python
  • Machine Learning (ML)
  • PyTorch
  • scikit-learn
  • Reasoning
  • Design Of Experiments
  • Testing
  • Financial Services
  • Artificial Intelligence
  • Communication
  • Computer Science
  • Statistics
  • Applied Mathematics
  • Research
  • Finance
  • Collaboration

Summary

About the Role

This role sits at the intersection of AI safety research, LLM evaluation, and applied data science. You will design and run adversarial evaluations against production AI agents, develop quantitative frameworks to measure evaluation quality, and translate findings into actionable insights for application teams and business stakeholders. You will also mentor junior data scientists and contribute to the growth of our AI safety research practice.

Responsibilities
  • Design and execute adversarial evaluations of generative AI and agentic systems, probing for safety failures, policy violations, and unexpected behaviors under realistic conditions
  • Apply rigorous experimental design and statistical methodology to AI safety research questions, including evaluation of model robustness and the reliability of automated evaluation systems
  • Develop and validate quantitative frameworks for assessing LLM outputs and measuring evaluation quality at scale
  • Build and maintain data pipelines to process large-scale evaluation outputs, aggregate metrics, and surface trends for research and stakeholder consumption
  • Engage directly with internal stakeholders including application teams, security analysts, and business leaders to explain findings and translate complex analytical results into clear, actionable recommendations
  • Mentor and develop junior data scientists and analysts on the team
  • Stay current with emerging literature in adversarial ML, LLM safety, and AI evaluation; contribute to the team's evolving research agenda
  • Participate in special projects and perform other duties as assigned

Qualifications
  • Minimum 5 years of experience in data science, applied ML research, or AI evaluation roles
  • Hands-on experience with large language models and agentic AI systems, with working knowledge of LLM behavior, failure modes, and safety evaluation techniques
  • Familiarity with adversarial AI concepts including jailbreaks, prompt injection, and model robustness, whether through direct research or applied work
  • Strong programming skills in Python, including experience with data pipelines, large-scale experimentation, and ML libraries such as PyTorch, Hugging Face, or Scikit-learn
  • Solid foundation in statistical reasoning and experimental design: hypothesis formulation, significance testing, and the ability to identify confounds and methodological flaws in existing analyses
  • Experience in financial services, AI safety, trust and safety, or a regulated industry is a strong plus
  • Clear, confident communication skills with the ability to explain technical findings to both research peers and non-technical stakeholders
  • Bachelor's degree in Computer Science, Statistics, Applied Mathematics, or a related quantitative field; Master's degree or equivalent research experience preferred

Special Factors

Sponsorship
Vanguard is not offering visa sponsorship for this position.

About Vanguard

At Vanguard, we don't just have a mission-we're on a mission.

To work for the long-term financial wellbeing of our clients. To lead through product and services that transform our clients' lives. To learn and develop our skills as individuals and as a team. From Malvern to Melbourne, our mission drives us forward and inspires us to be our best.

How We Work

Vanguard has implemented a hybrid working model for the majority of our crew members, designed to capture the benefits of enhanced flexibility while enabling in-person learning, collaboration, and connection. We believe our mission-driven and highly collaborative culture is a critical enabler to support long-term client outcomes and enrich the employee experience.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 90922487
  • Position Id: 24588761
  • Posted 3 days ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Dallas, Texas

Today

Full-time

Dallas, Texas

Today

Easy Apply

Contract

55

Remote or Washington

Today

Full-time

Remote

Today

Full-time

USD 87,400.00 - 123,400.00 per year

Search all similar jobs