Experience: 10+ years is must
Must be willing to work 4 days onsite in Jersey City, NJ
Job Summary :
Seeking a senior AI Foundation Model Engineer with 10+ years of experience in AI/ML and software engineering to build, fine-tune, optimize, and deploy enterprise-scale Foundation Models/LLMs and Generative AI solutions.
Key Requirements:
10+ years of experience in AI/ML, Machine Learning Engineering, or Software Engineering.
Strong Python and hands-on PyTorch/TensorFlow experience.
Expertise in LLMs, Transformers, Foundation Models, NLP, and Generative AI.
Experience with model training, fine-tuning, LoRA/QLoRA, RLHF/DPO, and model evaluation.
Strong knowledge of GPU/CUDA, distributed training, and model optimization.
Experience with Hugging Face, Docker, Kubernetes, and cloud platforms (AWS/Azure/Google Cloud Platform).
Experience with MLOps/LLMOps, model deployment, monitoring, and inference optimization.
Knowledge of RAG, embeddings, vector databases, and LLM inference/serving is a plus.
Strong communication and problem-solving skills.
Technical Skills:
AI/ML: LLMs, Foundation Models, Generative AI, NLP, Transformers, Deep Learning, RAG, Embeddings, Fine-Tuning, RLHF/DPO, PEFT, LoRA/QLoRA
Frameworks: PyTorch, TensorFlow, Hugging Face, DeepSpeed, Megatron-LM, Ray
Model Serving: vLLM, TensorRT-LLM, Triton Inference Server, FastAPI
Cloud & Infrastructure: AWS, Azure, Google Cloud Platform, Kubernetes, Docker, NVIDIA GPCUDA
MLOps: MLflow, Kubeflow, CI/CD, Model Monitoring, LLMOps
Data: Python, SQL, Spark, Kafka, Databricks, Vector Databases