Quick Overview
Job Description
We are seeking an AI Engineer to design and deploy production-grade AI solutions across LLM applications, RAG pipelines, retrieval systems, and scalable ML services.
This role blends model orchestration, search relevance, MLOps, and cross-functional collaboration to deliver reliable, high-impact AI products.
Responsibilities
Build and optimize LLM, RAG, and multi-agent workflows for production use cases.
Design retrieval systems using vector search, reranking, and metadata filtering.
Fine-tune, evaluate, and benchmark models to improve output quality and business impact.
Deploy AI/ML services with strong standards for latency, scalability, and observability.
Develop data, inference, and experimentation pipelines in partnership with product and engineering teams.
Required Qualifications
Strong Python and SQL skills with solid machine learning and NLP fundamentals.
Experience with LangChain, LangGraph, LlamaIndex, prompt engineering, and retrieval optimization.
Familiarity with vector databases and search tools such as Pinecone, OpenSearch, FAISS, or pgvector.
Experience with Docker, Kubernetes, MLflow, CI/CD, and cloud platforms such as Azure.
Similar jobs
- IN
Computer Vision / AI Engineer II
NewIntone Networks Inc.
Coppell, TX🇺🇸HybridYesterdayAWSMLOpsMachine Learning+7Technology - VI
Gen AI Engineer- (Exp -3+ Years )-Full Time-Hybrid
NewVisionary Innovative Technology Solutions
Tampa, FL🇺🇸HybridYesterdayFastAPIFlaskAWS+5Technology - AG
AI Developer - ServiceNow
NewAdvent Global Solutions, Inc.
United States🇺🇸HybridYesterdayTechnology - CY
Data Scientist/Applied AI Engineer (3)
NewCybersearch, Ltd.
San Francisco, CA🇺🇸$125 - $140/hrHybridYesterdayMachine LearningLLMPyTorch+1Technology - NG
Senior AI Engineer (only W2)
NewNGTalentTech Group LLC
United States🇺🇸RemoteYesterdayGenerative AILLMPython+2Technology - HT
AI Engineer
Horizontal Talent
Chicago, IL🇺🇸Hybrid4 weeks agoFastAPIMongoDBNeo4j+4Technology