Why This Role Stands Out
This onsite role at iMedhas Consulting Services offers a unique opportunity to lead the development and deployment of cutting-edge Generative AI solutions within a secure, on-premises environment, perfect for experienced LLM Engineers passionate about deep technical challenges and enterprise-level AI innovation. You'll gain invaluable experience fine-tuning open-weight models and optimizing local serving engines, contributing directly to impactful AI advancements.
Quick Overview
Job Description
Role OverviewWe are seeking an experienced LLM Engineer with 7+ years of software engineering experience, including 4+ years dedicated to AI/ML. You will design, develop, fine-tune, and deploy state-of-the-art Large Language Models (LLMs) and Generative AI applications directly within our on-premises, air-gapped enterprise infrastructure.In this role, you will lead the end-to-end lifecycle of local GenAI solutions—from self-hosted model serving and custom prompt engineering to fine-tuning open-weight models (e.g., Llama 3, Mistral, Qwen) while ensuring strict enterprise data privacy, security, and low latency.Required Qualifications & Technical SkillsExperience: 7+ years of overall software development experience, with 4+ years of hands-on experience in Machine Learning, Deep Learning, and AI.Python Mastery: Expert-level Python skills and deep familiarity with core AI ecosystems: PyTorch, TensorFlow, Hugging Face (transformers, peft, datasets, accelerate), spaCy, and Scikit-Learn.Self-Hosted / Open-Source LLMs: Hands-on experience working with open-weight foundation models (Llama, Mistral, Gemma, DeepSeek, Qwen) and local serving engines (vLLM, Ollama, TensorRT-LLM, Triton).On-Prem Infrastructure & Orchestration: Solid understanding of Linux, Docker/Kubernetes (OpenShift, Rancher, microK8s), local GPU orchestration, and CUDA driver configurations.Deployments: Proven track record of deploying at least one end-to-end GenAI application in a production environment.Education & Core Competencies: Bachelor’s or Master’s degree in Computer Science, Data Science, AI, or a related quantitative field. Strong problem-solving, analytical, and cross-functional communication skills.
Similar jobs
- MC
AI Engineer || Atlanta, GA
NewMcKinsol Consulting Inc
Atlanta, GA🇺🇸On-site17 hours agoJavaScriptLLMREST+1Technology - EP
Enterprise AI Architect
NewEmpower Professionals
United States🇺🇸Remote17 hours agoAgileAzureStakeholder ManagementTechnology - AI
AI Engineer
NewAsterism IT Solutions
Dallas, TX🇺🇸Hybrid17 hours agoNLPGPTHIPAA+4Technology - TE
Gen AI Engineer - Columbus, OH, Chicago, IL, Minneapolis, MN, Detroit, MI, Madison, WI.
TechniPros, LLC
Columbus, OH🇺🇸Hybrid1 week agoDockerFastAPIAzure+5Technology - TE
Gen AI Engineer - Kansas City, KS, Denver, CO, Phoenix, AZ, St. Louis, MO, Nashville, TN.
TechniPros, LLC
Kansas City, KS🇺🇸Hybrid1 week agoDockerFastAPIMicroservices+10Technology - TE
Agentic AI Engineer - Nashville, TN, Kansas City, KS, Denver, CO, Phoenix, AZ, St. Louis, MO.
NewTechniPros, LLC
Nashville, TN🇺🇸Hybrid17 hours agoSQLAWSScrum+8Technology