AI Foundation Model Engineer

Hybrid in Jersey City, NJ, US • Posted 22 hours ago • Updated 22 hours ago
Contract W2
12 Months
No Travel Required
Hybrid
$75 - $75/hr
Company Branding Image
Fitment

Dice Job Match Score™

🤯 Applying directly to the forehead...

Job Details

Skills

  • LLM
  • GenAI
  • RAG
  • LangChain
  • LlamaIndex
  • Hugging Face
  • PyTorch
  • AWS
  • Bedrock
  • SageMaker
  • OpenSearch
  • Kubernetes
  • Docker
  • MLOps
  • LLMOps

Summary

We are partnering with one of our Global Consulting client to fill a below position... 

 

Job Description:

Role: AI Foundation Model Engineer

Location: Jersey City, NJ – 4 Days onsite

Duration: 12 Months CTH

Interview Mode: Face to Face (Final round)

Salary/Rate: $75/hr. W2

 

Core keywords:

·       LLM, GenAI, RAG, embeddings, vector database, LangChain, LlamaIndex, Hugging Face, PyTorch, AWS Bedrock, SageMaker, OpenSearch, Kubernetes, Docker, Terraform, CI/CD, MLOps, LLMOps, model serving

 

Role purpose

·       Design, build, deploy, and optimize enterprise-grade AI systems powered by foundation models, LLMs, retrieval-augmented generation, and agentic workflows. The role converts AI concepts into secure, scalable, observable, and supportable production systems on the enterprise AI-ready platform (AIRP), which is currently AWS-hosted while following a cloud-agnostic architecture blueprint.

 

Client-specific emphasis

·       Hands-on AWS AI and cloud engineering is a major asset because AIRP currently runs on AWS.

·       Candidates should be comfortable working with Terraform/IaC and CI/CD teams to move AI services and infrastructure through controlled deployment pipelines.

·       Experience should map to business AI use cases such as KYC, credit underwriting, pitch book generation, Banker 360, Customer 360, deal library intelligence, financial crime quality, and sanctions screening.

 

Primary ownership

·       Production LLM applications, RAG pipelines, AI services, and model-serving integrations for AIRP.

·       End-to-end LLMOps/MLOps lifecycle from experimentation to deployment, monitoring, evaluation, rollback, and continuous improvement.

·       Reusable AI service components, APIs, prompts, retrieval logic, and observability patterns that can be federated across multiple business use cases.

 

Key responsibilities

·       Design and implement LLM-powered applications such as knowledge assistants, document intelligence solutions, workflow agents, summarization tools, and decision-support systems.

·       Build RAG pipelines using embeddings, chunking strategies, vector databases, semantic retrieval, reranking, response grounding, and citation patterns.

·       Integrate AI capabilities with AWS-hosted platform components, including model APIs, model gateways, data services, container platforms, and enterprise authentication patterns.

·       Collaborate with cloud engineering teams on Terraform modules, IaC templates, environment promotion, CI/CD pipelines, release controls, and rollback procedures.

·       Adapt and optimize models using LoRA, PEFT, instruction tuning, distillation, transfer learning, quantization, and domain adaptation techniques where appropriate.

·       Optimize inference workloads for latency, throughput, token efficiency, cost, reliability, and user experience.

·       Implement model and application observability, including prompt logs, retrieval quality, hallucination indicators, drift signals, feedback loops, cost telemetry, and service health.

·       Embed security, privacy, Responsible AI, and model risk controls into AI application design and delivery.

·       Create production documentation, runbooks, release notes, test evidence, and audit-ready implementation records.

 

Must-have candidate profile

·       7+ years in AI/ML engineering, platform engineering, software engineering, or applied machine learning.

·       Hands-on experience with LLMs, transformers, embeddings, RAG, semantic search, and GenAI application patterns.

·       Strong Python engineering skills with PyTorch, TensorFlow, Hugging Face, LangChain, LlamaIndex, Semantic Kernel, or equivalent frameworks.

·       Experience deploying production AI services using APIs, containers, Kubernetes, CI/CD, cloud-native services, and monitoring platforms.

·       Practical exposure to AWS AI/cloud services or comparable cloud-native AI deployment experience, with ability to ramp quickly on AWS-hosted AIRP patterns.

·       Working knowledge of Terraform/IaC, DevOps pipelines, release management, model evaluation, inference optimization, and secure data handling.

·       Preferred experience

·       Banking, risk, compliance, financial crime, operations, or enterprise technology background.

·       Experience with AWS Bedrock, SageMaker, OpenSearch, Kendra, Lambda, EKS/ECS, Azure OpenAI, Vertex AI, Databricks, vLLM, Triton, MLflow, Kubeflow, or model gateways.

·       Exposure to cloud-agnostic application patterns, reusable IaC modules, model risk, AI governance, audit controls, AI cost governance, and private or open-source LLM deployments.

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91130450
  • Position Id: 9029839
  • Posted 22 hours ago

Company Info

About Echo IT Solutions, Inc.

At ECHO IT SOLUTIONS we offer comprehensive IT solutions and consultancy services designed specifically for organizations across the world. We value our employees as our greatest strength; genuinely care about their growth and continuous support.

We believe happy employees working together can achieve greatness and help contribute to the growing IT industry.

Consistently, we have grown into a prominent and trusted player in the technology industry. We’ve been repeatedly recognized as one of the most prompt ServiceProvider. We strive to improve our customers business for the better and help ease the burden. We take pride in approaching every customer commitment as a novel & unique relationship.

Contact the job poster
Deva Perni

Deva Perni

Recruiter @ Echo IT Solutions, Inc.
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Hybrid in Jersey City, New Jersey

Today

Easy Apply

Contract

73 - 73

Secaucus, New Jersey

Today

Easy Apply

Contract

Depends on Experience

Hybrid in Jersey City, New Jersey

Today

Easy Apply

Contract

80 - 80

Irving, Texas

7d ago

Easy Apply

Full-time

140,000 - 140,000

Search all similar jobs