Senior Generative AI Engineer (Google Gemini & Vertex AI) - Remote

Remote • Posted 1 hour ago • Updated 1 hour ago
Contract W2
Contract Corp To Corp
6 Months
No Travel Required
Remote
$51 - $55/hr
Fitment

Dice Job Match Score™

👤 Reviewing your profile...

Job Details

Skills

  • AI
  • Python
  • Gemini
  • Vertex AI
  • PyTorch
  • SQL

Summary

DivIHN (pronounced “divine”) is a CMMI ML3-certified Technology and Talent solutions firm. Driven by a unique Purpose, Culture, and Value Delivery Model, we enable meaningful connections between talented professionals and forward-thinking organizations. Since our formation in 2002, organizations across commercial and public sectors have been trusting us to help build their teams with exceptional temporary and permanent talent.

Visit us at to learn more and view our open positions.

 
Please apply or call one of us to learn more

For further inquiries about this opportunity, please contact one of  our Talent Specialists, Justeen at or Ragu
 
Title: Senior Generative AI Engineer (Google Gemini & Vertex AI) - Remote
Duration: 6 Months 
Location: Remote
 
Only W2 candidates are eligible for this position. Third-party or C2C candidates will not be considered.
 
Day to Day Responsibilities
 
Responsibilities:
  • RAG & Prompt Engineering: Craft and refine effective prompts for RAG, grounding, and context tuning to achieve optimal AI performance in product development. 
  • High-Performance API Engineering: Develop asynchronous microservices (FastAPI) using Server-Sent Events (SSE) or WebSockets to stream real-time LLM responses to front-end UIs without backend timeouts. 
  • Vector Database Infrastructure: Design, develop, and implement robust Vector Databases using LLMs and modern retrieval technologies to capture information from diverse engineering sources (PDFs, design docs, regulatory guidelines). 
  • Data Extraction & Structuring Pipelines: Build and optimize pipelines to extract and structure multi-modal data (tables, text, images) from unstructured documents for LLM training, grounding, and runtime query execution. 
  • LLM Fine-Tuning & Training: Fine-tune and train generative AI models using client's engineering data and domain knowledge to create high-accuracy, domain-specific models. 
  • GenAI Application & Tool Development: Design and implement scalable backend APIs (FastAPI/REST) and UI integration interfaces so internal engineers can query knowledge bases and analyze data. 
  • Automated Requirements Generation: Develop backend functionalities to automatically generate technical requirements from design documents, user stories, and system specification files. 
  • Documentation & Knowledge Transfer: Thoroughly document architecture, code, REST endpoints, and model training procedures to enable seamless knowledge transfer to client's internal teams. 
  • Cross-Functional Collaboration: Partner closely with Subject Matter Experts (SMEs), System Engineers, and V&V Test teams to optimize AI-powered workflows. 
  • AI Guardrails, MLOps & Cost Governance: Implement hallucination checks, PII masking, and guardrails (e.g., NeMo Guardrails) for medical device context. Track token usage, latency, and costs using LangSmith or Vertex AI monitoring.
Must-Have Qualifications :
  •  Relevant Experience & STEM Foundation: 4+ years of professional software/ML engineering experience, with a dedicated AI/ML focus in the last 1–2 years.
  •  Google Gemini / Vertex AI (Non-negotiable): Hands-on experience with the Gemini model family and Vertex AI, including deployment, grounding, and integration into production AI services.  Hands-on experience with containerization (Docker) and deploying services via Cloud Run or GKE (Kubernetes).
  • Languages & AI Libraries: Proficiency in Python and modern ML/AI frameworks (PyTorch, LangChain, LangSmith) for building autonomous LLM agents, tools, and RAG pipelines.
  • Agent Building & Tool Calling: Proven experience building AI/LLM agents and tool-calling systems in Python against unstructured, multi-source data.
  • Context Engineering & RAG: Expertise in RAG pipelines, prompt engineering, context tuning, grounding, and Vector Databases (e.g., Milvus, Postgres/Pgvector). Clear understanding of advanced RAG architecture including Hybrid Search (Vector + Keyword), Re-ranking models, and semantic caching.
  • Unstructured Data Handling (Non-negotiable): Demonstrated ability to ingest, clean, extract, and structure text, tables, and images from unstructured documents (PDFs, design docs, regulatory files) for LLM training and usage.
Required Skills :
 
1.    Full Stack Engineering — Building applications for AI-powered services (Backend APIs + Front-End Integration).
2.    Generative AI & LLM Platforms — 2+ years building RAG pipelines & LLM apps on Google Gemini / Vertex AI using Python, LangChain, and LangSmith.
3.    Agent Building & Unstructured Data — Building AI agents/tool-calling systems in Python and handling unstructured, multi-source data extraction (PDFs, docs).
 
Preferred Skills :
 
1.    Direct experience architecting and serving custom REST APIs.
2.    Experience in regulated / compliance-sensitive domains (e.g., healthcare / medical device guidelines like FDA, ISO 13485).
3.    Experience integrating multiple LLM APIs beyond a single provider (OpenAI, Bedrock, Claude, or similar).
 
Additional Preferred Skills 
 
•    Direct experience designing and deploying high-throughput REST APIs (e.g., FastAPI/Flask).
•    Familiarity with medical device development regulations and compliance (e.g., FDA guidelines, ISO 13485).
•    Experience integrating multiple LLM APIs beyond a single vendor (e.g., OpenAI, AWS Bedrock, Anthropic Claude).
•    Front-end development/integration experience for UI design (e.g., Streamlit, Gradio, React/Next.js integration).
 
Education Requirements:
  • Minimum Bachelors in Software/Computer/IT/Systems/Biomedical Engineering + 3 years
Required Testing:
  • Technical evaluation of Python proficiency, RAG architecture concepts, and API/Agent design.
Software Skills Required:
  • Languages: Python (AsyncIO, OOP), SQL.
  • AI & Agent Frameworks: PyTorch, LangChain, LangSmith, Vertex AI SDK.
  • API & Web Frameworks: FastAPI, Flask, REST APIs, Server-Sent Events (SSE).
  • Databases & Search: Milvus, Pgvector, Qdrant, Redis (caching).
  • Front-End Integration: Streamlit, Gradio, basic React/Next.js.
  • Testing & Guardrails: pytest, JUnit, LangSmith evaluation, NeMo Guardrails.
  • DevOps & Cloud: Docker, Google Cloud Platform (Vertex AI, Cloud Run, GKE), Git.
Required Certifications:
  • Professional certifications specific to AI/ML (e.g., Certified AI Professional / CAIP, Google Cloud ML Engineer) considered a plus.
Interview 
  • Number of Interviews: 1,
  • Web Conference (Zoom/ TEAMs) 

About us:
DivIHN, the ''IT Asset Performance Services'' organization, provides Professional Consulting, Custom Projects, and Professional Resource Augmentation services to clients in the Mid-West and beyond. The strategic characteristics of the organization are Standardization, Specialization, and Collaboration.

DivIHN is an equal opportunity employer. DivIHN does not and shall not discriminate against any employee or qualified applicant on the basis of race, color, religion (creed), gender, gender expression, age, national origin (ancestry), disability, marital status, sexual orientation, or military status.

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10109463
  • Position Id: 11671-3720-1786985591
  • Posted 1 hour ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote

8d ago

Easy Apply

Contract

Depends on Experience

Remote

11d ago

Easy Apply

Contract

85 - 90

Remote

5d ago

Easy Apply

Contract, Third Party

$93.34 - $93.34

Remote or Eden Prairie, Minnesota

Today

Full-time

USD 120,100.00 - 214,500.00 per year

Search all similar jobs