Haystack
← Back to Jobs
Other
SC

AI Foundation Model Engineer

Source Code Technologies LLCJersey City, NJ🇺🇸United StatesPosted 31 Aug 2026

Quick Overview

Seniority
Mid Senior
Work mode
On Site
Location
Jersey City, NJ, United States
Posted
Yesterday
DockerFastAPISQLAWSMLOpsMLflowMachine LearningNLPAzureCUDADatabricksDeep LearningGenerative AIGoogle CloudHugging FaceKafkaKubernetesLLMPyTorchPythonTensorFlow

Job Description

Experience: 10+ years is must

Must be willing to work 4 days onsite in Jersey City, NJ

 

Job Summary :

Seeking a senior AI Foundation Model Engineer with 10+ years of experience in AI/ML and software engineering to build, fine-tune, optimize, and deploy enterprise-scale Foundation Models/LLMs and Generative AI solutions.

Key Requirements:
10+ years of experience in AI/ML, Machine Learning Engineering, or Software Engineering.
Strong Python and hands-on PyTorch/TensorFlow experience.
Expertise in LLMs, Transformers, Foundation Models, NLP, and Generative AI.
Experience with model training, fine-tuning, LoRA/QLoRA, RLHF/DPO, and model evaluation.
Strong knowledge of GPU/CUDA, distributed training, and model optimization.
Experience with Hugging Face, Docker, Kubernetes, and cloud platforms (AWS/Azure/Google Cloud Platform).
Experience with MLOps/LLMOps, model deployment, monitoring, and inference optimization.
Knowledge of RAG, embeddings, vector databases, and LLM inference/serving is a plus.
Strong communication and problem-solving skills.

 

Technical Skills:

AI/ML: LLMs, Foundation Models, Generative AI, NLP, Transformers, Deep Learning, RAG, Embeddings, Fine-Tuning, RLHF/DPO, PEFT, LoRA/QLoRA

Frameworks: PyTorch, TensorFlow, Hugging Face, DeepSpeed, Megatron-LM, Ray

Model Serving: vLLM, TensorRT-LLM, Triton Inference Server, FastAPI

Cloud & Infrastructure: AWS, Azure, Google Cloud Platform, Kubernetes, Docker, NVIDIA GPCUDA

MLOps: MLflow, Kubeflow, CI/CD, Model Monitoring, LLMOps

Data: Python, SQL, Spark, Kafka, Databricks, Vector Databases

 

Similar jobs