Haystack
← Back to Jobs
Technology
SI

Gen AI Engineer

Source InfotechCharlotte, NC🇺🇸United StatesPosted 11 Sept 2026

Quick Overview

Seniority
Mid Senior
Work mode
Hybrid
Location
Charlotte, NC, United States
Posted
17 hours ago
AWSMLOpsMachine LearningNLPAzureGPTGenerative AIGoogle CloudHugging FaceLLMPyTorchPythonTensorFlow

Job Description

Gen AI Engineer

Charlotte, NC - Locals or Nearby preffered 

 

The Role:

Responsibilities:

Prompt Development & Engineering:

  • Design and optimize effective prompts for large language models to improve output quality for various use cases.
  • Fine-tune prompt strategies for specific applications, including chatbots, content generation, and automated customer service.
  • Strong expereince with Devin AI and Claude Code
  • Test and iterate on different prompt approaches to ensure alignment with project goals.

Large Language Model (LLM) & Retrieval-Augmented Generation (RAG):

  • Develop and fine-tune large language models like GPT, Gemini, Langchain, and Llama for specific business needs.
  • Implement RAG techniques to improve model outputs by integrating external knowledge from retrieval systems.
  • Leverage GraphRAG to enhance complex knowledge retrieval and graph-based data representation in AI models.
  • Stay updated on advancements in AI/LLM technologies and recommend new tools or models to enhance the AI stack.

Collaboration & Communication:

  • Work closely with product managers, software developers, and other stakeholders to align AI capabilities with business objectives.
  • Communicate technical concepts and model behavior to non-technical team members in a clear and concise manner.
  • Provide documentation and training to users and developers on utilizing AI models effectively.

Deployment & Monitoring:

  • Deploy AI models into production environments using cloud services or on-premise infrastructures.
  • Continuously monitor model performance, scaling solutions as needed, and ensuring models meet security and compliance standards.
  • Troubleshoot and optimize models for speed, accuracy, and scalability in production systems.

Required Qualifications:

  • Bachelor’s or Master’s degree in Computer Science, Data Science, Machine Learning, or a related field.
  • Strong understanding of machine learning concepts, natural language processing (NLP), and generative AI.
  • Experience with prompt development and fine-tuning large language models like GPT, Gemini, Langchain, and Llama.
  • Proficiency in programming languages such as Python, with experience in AI/ML libraries (e.g., TensorFlow, PyTorch, Hugging Face).
  • Knowledge of MLOps tools for model deployment and monitoring.
  • Experience working with cloud platforms (e.g., AWS, Google Cloud Platform, Azure) for model training and deployment.

Preferred Qualifications:

  • Prior experience with Langchain for integrating LLMs into applications.
  • Familiarity with tools and techniques for AI model interpretability and responsible AI practices.
  • Strong analytical and problem-solving skills with attention to detail.
  • Ability to work in a fast-paced, collaborative environment.

Preferred, but not required:

  • 6+ years in Gen AI Python development.
  • Proven track record of implementing AI solutions in production environments.
  • Previous experience in working on projects with a strong focus on LangChain and multi-agent systems.

 

Similar jobs