Senior GenAI Engineer

Remote • Posted 1 day ago • Updated 10 hours ago
Full Time
Remote
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Generative Artificial Intelligence (AI)
  • Underwriting
  • Customer Service
  • Vector Databases
  • Semantic Search
  • Software Development Methodology
  • Development Testing
  • Production Support
  • Collaboration
  • Software Development
  • FOCUS
  • Prompt Engineering
  • Microsoft Certified Professional
  • Servers
  • Workflow
  • Orchestration
  • LangChain
  • LlamaIndex
  • Database
  • LangSmith
  • Docker
  • Continuous Integration
  • Continuous Delivery
  • Cloud Computing
  • Amazon Web Services
  • Microsoft Azure
  • Google Cloud
  • Google Cloud Platform
  • Software Engineering
  • Testing
  • Code Review
  • Version Control
  • Git
  • Insurance
  • Financial Services
  • Evaluation
  • Kubernetes
  • Terraform
  • Privacy
  • Regulatory Compliance
  • Python
  • Message Queues
  • React.js
  • Artificial Intelligence
  • English

Summary

Project description

We are seeking a hands-on GenAI Engineer to join a team developing production-grade, AI-powered solutions for a major US insurance provider. The role involves designing and delivering end-to-end LLM-based applications, including scalable backend APIs, agentic workflows, deployment, monitoring, and observability. You will support our client in applying generative AI to key insurance processes such as underwriting, claims, and customer service.

Responsibilities

Design, develop and maintain scalable backend services and APIs using Python and FastAPI.

Build and integrate LLM-powered features using OpenAI GPT and Anthropic Claude models.

Develop and maintain MCP (Model Context Protocol) servers using FastMCP to expose enterprise tools and data to AI agents.

Create and extend Skills and Plugins that enhance LLM capabilities for business-specific workflows.

Design and implement Retrieval-Augmented Generation (RAG) pipelines, including document ingestion, chunking, embedding and retrieval strategies.

Build agentic solutions using LLM orchestration frameworks such as LangChain, LangGraph or similar.

Work with vector databases to support semantic search and knowledge retrieval.

Implement observability and diagnostics for LLM applications: tracing, logging, evaluation, token and cost tracking, latency and quality monitoring.

Own the full application lifecycle: development, testing, deployment and production support.

Use AI-assisted development tools (Claude, Codex) to accelerate delivery while maintaining code quality.

Collaborate with client stakeholders, architects and business analysts to translate requirements into working solutions.

Ensure solutions meet enterprise standards for security, data privacy and responsible AI use.

Skills

Must have

4+ years of professional software development experience with a strong focus on Python.

Solid experience building production REST APIs with FastAPI (or a comparable framework).

Hands-on experience integrating LLMs (OpenAI GPT, Anthropic Claude) into real applications, including prompt engineering, tool/function calling and structured outputs.

Practical experience with MCP servers (FastMCP preferred) and LLM tool ecosystems (Skills, Plugins).

Proven experience designing and implementing RAG pipelines.

Experience building agentic workflows with orchestration frameworks such as LangChain, LangGraph, LlamaIndex or similar.

Hands-on experience with at least one vector database (e.g., Pinecone, Weaviate, Qdrant, Chroma, pgvector, Azure AI Search).

Experience with observability and diagnostics for LLM systems (e.g., LangSmith, Langfuse, Arize Phoenix, OpenTelemetry).

Experience deploying and running applications in production: Docker, CI/CD, and at least one major cloud platform (AWS, Azure or Google Cloud Platform).

Daily, confident use of modern development tools including VS Code and AI coding assistants such as Claude and Codex.

Strong understanding of software engineering best practices: clean code, testing, code review, version control (Git).

Upper-Intermediate (B2) or higher English, with the ability to communicate directly with US-based stakeholders.

Nice to have

Experience in the insurance or broader financial services domain.

Knowledge of LLM evaluation techniques (automated evals, LLM-as-judge, guardrails).

Experience with Kubernetes and infrastructure-as-code (Terraform, Bicep).

Familiarity with data privacy and compliance requirements for handling PII in regulated industries.

Experience with asynchronous Python, message queues or event-driven architectures.

Frontend experience (React, Streamlit) for building internal AI tools and demos.

Other

Languages

English: B2 Upper Intermediate

Seniority

Senior
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10335530
  • Position Id: 391bc316c98f6dcb89d4716ef23a4ffd
  • Posted 1 day ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Remote or Chicago, Illinois

•

Today

Easy Apply

Full-time, Part-time, Contract, Third Party

Remote

•

Today

Full-time

Remote

•

Today

Full-time

Remote or Columbus, Ohio

•

Today

Full-time

USD 110,300.00 - 183,800.00 per year

Search all similar jobs