Senior LLM Engineer – Generative AI- 10+ yrs- New York, United States- Onsite

New York, NY, US • Posted 9 hours ago • Updated 9 hours ago
Contract Corp To Corp
Contract W2
Contract Independent
6 Months
No Travel Required
On-site
Depends on Experience
Company Branding Image
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • AI/ML. You will design
  • develop
  • fine-tune
  • and deploy state-of-the-art Large Language Models (LLMs) and Generative AI applications directly within our on-premises
  • air-gapped enterprise infrastructure.In this role
  • you will lead the end-to-end lifecycle of local GenAI solutions—from self-hosted model serving and custom prompt engineering to fine-tuning open-weight models (e.g.
  • Llama 3
  • Mistral
  • Qwen) while ensuring strict enterprise data privacy
  • security
  • and low latency.Required Qualifications & Technical SkillsExperience: 7+ years of overall software development experience
  • with 4+ years of hands-on experience in Machine Learning
  • Deep Learning
  • and AI.Python Mastery: Expert-level Python skills and deep familiarity with core AI ecosystems: PyTorch
  • TensorFlow
  • Hugging Face (transformers
  • peft
  • datasets
  • accelerate)
  • spaCy
  • and Scikit-Learn.Self-Hosted / Open-Source LLMs: Hands-on experience working with open-weight foundation models (Llama
  • Gemma
  • DeepSeek
  • Qwen) and local serving engines (vLLM
  • Ollama
  • TensorRT-LLM
  • Triton).On-Prem Infrastructure & Orchestration: Solid understanding of Linux
  • Docker/Kubernetes (OpenShift
  • Rancher
  • microK8s)
  • local GPU orchestration
  • and CUDA driver configurations.Deployments: Proven track record of deploying at least one end-to-end GenAI application in a production environment.Education & Core Competencies: Bachelor’s or Master’s degree in Computer Science
  • Data Science
  • AI
  • or a related quantitative field. Strong problem-solving
  • analytical
  • and cross-functional communication skills.

Summary

Job Description :

Role OverviewWe are seeking an experienced LLM Engineer with 7+ years of software engineering experience, including 4+ years dedicated to AI/ML. You will design, develop, fine-tune, and deploy state-of-the-art Large Language Models (LLMs) and Generative AI applications directly within our on-premises, air-gapped enterprise infrastructure.In this role, you will lead the end-to-end lifecycle of local GenAI solutions—from self-hosted model serving and custom prompt engineering to fine-tuning open-weight models (e.g., Llama 3, Mistral, Qwen) while ensuring strict enterprise data privacy, security, and low latency.Required Qualifications & Technical SkillsExperience: 7+ years of overall software development experience, with 4+ years of hands-on experience in Machine Learning, Deep Learning, and AI.Python Mastery: Expert-level Python skills and deep familiarity with core AI ecosystems: PyTorch, TensorFlow, Hugging Face (transformers, peft, datasets, accelerate), spaCy, and Scikit-Learn.Self-Hosted / Open-Source LLMs: Hands-on experience working with open-weight foundation models (Llama, Mistral, Gemma, DeepSeek, Qwen) and local serving engines (vLLM, Ollama, TensorRT-LLM, Triton).On-Prem Infrastructure & Orchestration: Solid understanding of Linux, Docker/Kubernetes (OpenShift, Rancher, microK8s), local GPU orchestration, and CUDA driver configurations.Deployments: Proven track record of deploying at least one end-to-end GenAI application in a production environment.Education & Core Competencies: Bachelor’s or Master’s degree in Computer Science, Data Science, AI, or a related quantitative field. Strong problem-solving, analytical, and cross-functional communication skills.

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91143549
  • Position Id: 5000-31204-1790689060
  • Posted 9 hours ago

Company Info

About iMedhas Consulting Services

Welcome to iMedhas Consulting Services. We are an IT consulting and services enterprises with precision expertise in Digital Transformations, Big data and Analytics. Through our expert team, we provide greater adaptabilities to disseminate with new technologies to increase the overall productivity, operational efficiency, and productivity of a particular business process of an organization.

The goal is the achievement of higher customer satisfaction and providing lucrative returns on investment to the business. On the other hand, we deliver content management expertise, ERP (SAP) systems integration, EAI, and Information Management services.

About_Company_One
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

New York, New York

•

Yesterday

Easy Apply

Contract, Third Party

Depends on Experience

New York, New York

•

Yesterday

Easy Apply

Third Party, Contract

Depends on Experience

Reading, Massachusetts

•

Today

Easy Apply

Contract, Third Party

Depends on Experience

Search all similar jobs