AI Architect Generative AI, LLMs & AWS
Quick Overview
Job Description
Apexon is a digital-first technology services firm specializing in accelerating business transformation and delivering human-centric digital experiences. We have been meeting customers wherever they are in the digital lifecycle and helping them outperform their competition through speed and innovation. Apexon brings together distinct core competencies in AI, analytics, app development, cloud, commerce, CX, data, DevOps, IoT, mobile, quality engineering and UX, and our deep expertise in BFSI, healthcare, and life sciences to help businesses capitalize on the unlimited opportunities digital offers. Our reputation is built on a comprehensive suite of engineering services, a dedication to solving clients toughest technology problems, and a commitment to continuous improvement. Backed by Goldman Sachs Asset Management and Everstone Capital, Apexon now has a global presence of 15 offices (and 10 delivery centers) across four continents.
We enable #HumanFirstDIGITAL
Job Description: AI Architect Generative AI, LLMs & AWS
Role Summary
We are looking for a hands-on AI Architect to lead the design, development, enhancement, and deployment of an enterprise AI platform powered by Large Language Models (LLMs). The ideal candidate will combine strong software engineering skills with deep expertise in Generative AI, Agentic AI, and AWS cloud services to build scalable, secure, and production-ready AI applications.
This role requires active involvement in solution architecture, coding, cloud deployment, performance optimization, and mentoring engineering teams.
Key Responsibilities
- Design, develop, and enhance an enterprise AI platform integrating multiple Large Language Models (LLMs).
- Architect and implement scalable AI solutions using modern Agentic AI frameworks and Retrieval-Augmented Generation (RAG) architectures.
- Build intelligent AI agents capable of reasoning, planning, tool execution, and workflow orchestration.
- Develop secure, scalable, and highly available cloud-native AI applications on AWS.
- Design and implement REST APIs and microservices to expose AI capabilities to enterprise applications.
- Integrate AI services with enterprise systems, databases, APIs, and third-party platforms.
- Deploy, monitor, and optimize AI workloads on AWS ensuring performance, scalability, security, and cost efficiency.
- Work closely with product owners and engineering teams to translate business requirements into technical solutions.
- Drive architecture reviews, code quality, CI/CD automation, and engineering best practices.
- Evaluate emerging LLMs, AI frameworks, and cloud services to continuously improve the AI platform.
- Mentor developers and provide technical leadership across AI initiatives.
Required Technical Skills
Generative AI
- Strong hands-on experience with OpenAI GPT, Anthropic Claude, Llama, Mistral, Amazon Nova, or similar foundation models.
- Experience building enterprise-grade LLM applications.
- Expertise in Retrieval-Augmented Generation (RAG).
- Prompt engineering, prompt optimization, embeddings, semantic search, and model evaluation.
- Experience integrating multiple LLM providers and managing model orchestration.
Agentic AI
Hands-on experience with one or more of:
- LangChain
- LangGraph
- CrewAI
- Microsoft Semantic Kernel
- AutoGen
- Amazon Bedrock Agents
Experience developing:
- Multi-agent workflows
- Tool calling
- Function calling
- Memory management
- Planning and reasoning agents
AWS Cloud
Strong hands-on experience with:
- Amazon Bedrock
- Amazon SageMaker
- AWS Lambda
- ECS/EKS
- API Gateway
- Step Functions
- Amazon S3
- DynamoDB
- Amazon OpenSearch
- Amazon Aurora
- CloudWatch
- IAM
- VPC
- EventBridge
- Secrets Manager
Experience with Infrastructure as Code (Terraform, AWS CDK, or CloudFormation) is highly desirable.
Programming
- Python (mandatory)
- FastAPI / Flask
- REST APIs
- Microservices
- Docker
- Kubernetes
- Git
- CI/CD pipelines
Databases & Search
Experience with:
- PostgreSQL
- DynamoDB
- OpenSearch
- Pinecone
- Weaviate
- Chroma
- FAISS
- Milvus
Preferred Qualifications
- Experience building AI products from concept to production.
- Strong understanding of LLMOps, MLOps, observability, and AI monitoring.
- Experience implementing AI guardrails, responsible AI, and enterprise security controls.
- Familiarity with event-driven and serverless architectures.
- Knowledge of authentication, authorization, and API security.
Our Commitment to Diversity & Inclusion:
Did you know that Apexon has been Certified by Great Place To Work , the global authority on workplace culture, in each of the three regions in which it operates: USA (for the fourth time in 2023), India (seven consecutive certifications as of 2023), and the UK.Apexon is committed to being an equal opportunity employer and promoting diversity in the workplace. We take affirmative action to ensure equal employment opportunity for all qualified individuals. Apexon strictly prohibits discrimination and harassment of any kind and provides equal employment opportunities to employees and applicants without regard to gender, race, color, ethnicity or national origin, age, disability, religion, sexual orientation, gender identity or expression, veteran status, or any other applicable characteristics protected by law. You can read about our Job Applicant Privacy policy here Job Applicant Privacy Policy (apexon.com)
Our Commitment to Environment:
Actively contribute to Apexon's commitment to environmental responsibility by following sustainable practices and supporting ESG initiatives.
Our Perks and Benefits:
Our benefits and rewards program has been thoughtfully designed to recognize your skills and contributions, elevate your learning/upskilling experience and provide care and support for you and your loved ones. As an Apexon Associate, you get continuous skill-based development, opportunities for career advancement, and access to comprehensive health and well-being benefits and assistance.
We also offer:
o Health Insurance with Dental & Vision
o 401K Plan
o Life Insurance, STD & LTD
o Paid Vacations & Holidays
o Paid Parental Leave
o FSA Dependent & Limited Purpose care
Skills
Similar jobs
AI Engineer
StratEdge It consulting INC · Bellevue, United States
10 minutes ago$150k/yrAI Engineer (Austin, TX, or Fort Mill, SC )
SANS · Austin, United States
11 minutes agoSenior AI Engineer - W2 role
Wise Skulls Corp. · East Hartford, United States
59 minutes agoNICE CXone Cognigy Aentic AI Developer
Ztek Consulting · United States
1 hour agoAI Engineer |Gen AI (healthcare Domain)_Hybrid@Denver(CO)
BURGEON IT SERVICES LLC · Denver, United States
1 hour agoGenerative AI / Agentic AI Architect
AIT Global, Inc. · United States
2 hours ago