BEFORE APPLYING FOR THE JOB. KINDLY READ THE COMPLETE JOB DESCRIPTION PROPERLY(EVEN THE TERMS OF EMPLOYMENT ALSO), NO 1099, C2C , H1 TRANSFER
Title: AI/ML & Gen AI Testing
Location: Hybrid (1 day/week on-site in Merriweather, MD; occasional travel to Northern VA / DC offices)
Terms of Employment:
W2 Contract, 6-Months (possibile of extension)
Location: Hybrid (1 day/week on-site in Merriweather, MD; occasional travel to Northern VA / DC offices)
Candidate must reside in DMV area only
Role Overview
This role sits at the intersection of software/QA testing and AI/ML it is not a hands-on model-building or deployment position. The engineer will work closely with the lead systems engineer and lead developer supporting two active workstreams:
1. Machine Learning Models supporting existing and new ML initiatives, evaluating fine-tuned and agentic model outputs, and providing iterative feedback to improve model efficiency and accuracy in collaboration with the business team.
2. Generative AI Chatbots (Bridge Console) testing chatbot responses for accuracy and hallucination, crafting and refining prompts, and working with the development team to fine-tune outputs based on findings.
Day-to-day work includes testing across multiple pipeline stages (evaluating intermediate outputs at each processing step, not just final results), analyzing model/chatbot outputs, and partnering with both engineering and business stakeholders to validate results.
Required Qualifications (Mandatory)
3+ years of software engineering / QA testing experience
1 2 years of hands-on experience with Generative AI concepts
3+ years of cloud platform experience (AWS, Azure, or Google Cloud Platform)
Working knowledge of system/software testing methodology
~1 year of experience with BDD, Selenium, and test automation
Understanding of relational and non-relational databases, with experience working with data pipelines (e.g., Oracle, Snowflake)
Hands-on exposure to prompt engineering
Strong analytical skills comfortable interpreting model outputs and providing structured feedback
Experience with Jira and Confluence
Preferred Qualifications
Familiarity with AWS Bedrock and/or SageMaker
Background in healthcare or insurance industry
Exposure to newer AI/ML testing frameworks
Salesforce experience (nice to have, not required)
Interview Process
1. Initial video screen (~30 minutes)
2. In-person interview (typically at Merriweather, location may vary)