Haystack
← Back to Jobs
Remote
Technology

GEN AI or AI Architect

Quantum World Technologies Inc.United States🇺🇸United StatesPosted 20 Jul 2026

Quick Overview

Work Type
Remote
Level
Mid Senior

Job Description

Job Title: Gen AI or AI Architect
Location: Remote Position
Job Type: Full-Time Position

Job Summary

We are seeking a talented Gen AI Engineer to join our AI engineering team and build enterprise-grade Generative AI solutions. The ideal candidate will have hands-on experience developing applications using Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), prompt engineering, and modern AI frameworks such as LangChain or LangGraph. You will work closely with data scientists, software engineers, and business stakeholders to design, develop, and deploy scalable AI-powered applications.

Key Responsibilities
  • Design, develop, and deploy Generative AI applications using Large Language Models (LLMs).
  • Build AI solutions leveraging OpenAI, Azure OpenAI, Anthropic Claude, Gemini, or Llama models.
  • Develop Retrieval-Augmented Generation (RAG) pipelines using vector databases.
  • Create intelligent AI agents using LangChain, LangGraph, or LlamaIndex.
  • Design and optimize prompts to improve LLM performance and response quality.
  • Integrate AI services with enterprise applications through REST APIs and microservices.
  • Deploy AI applications on AWS, Azure, or Google Cloud Platform using cloud-native services.
  • Work with structured and unstructured data to build intelligent search and knowledge retrieval systems.
  • Monitor, evaluate, and optimize AI model performance, latency, and cost.
  • Collaborate with cross-functional teams to translate business requirements into AI solutions.
  • Follow AI security, governance, and responsible AI best practices.
Required Skills
  • Bachelor's or Master's degree in Computer Science, Engineering, Data Science, or a related field.
  • 5+ years of software development experience with at least 2+ years in Generative AI.
  • Strong programming skills in Python.
  • Hands-on experience with OpenAI, Azure OpenAI, Anthropic Claude, Gemini, or Llama.
  • Experience with LangChain, LangGraph, LlamaIndex, or Semantic Kernel.
  • Strong understanding of Large Language Models (LLMs) and Prompt Engineering.
  • Experience implementing Retrieval-Augmented Generation (RAG) architectures.
  • Knowledge of vector databases such as Pinecone, Chroma, FAISS, Weaviate, or Milvus.
  • Experience developing REST APIs using FastAPI or Flask.
  • Experience with cloud platforms (AWS, Azure, or Google Cloud Platform).
  • Familiarity with Git, Docker, Kubernetes, and CI/CD pipelines.
Preferred Skills
  • Experience with AI agents and multi-agent systems.
  • Knowledge of fine-tuning and model evaluation techniques.
  • Experience with Hugging Face Transformers and open-source LLMs.
  • Familiarity with ML frameworks such as PyTorch or TensorFlow.
  • Experience with monitoring AI applications using LangSmith, MLflow, or similar tools.
  • Exposure to MLOps practices and AI governance.
Nice to Have
  • Experience building enterprise AI chatbots or AI assistants.
  • Knowledge of AI security, responsible AI, and compliance standards.
  • Experience working in Agile/Scrum environments.
  • Relevant cloud or AI certifications (AWS, Azure AI, Google Cloud AI).

Skills

Docker
FastAPI
Flask
Microservices
AWS
MLOps
MLflow
Scrum
Agile
Azure
Generative AI
Git
Google Cloud
Hugging Face
Kubernetes
LLM
PyTorch
Python
REST
TensorFlow

Similar jobs