Haystack
← Back to Jobs
Technology
ES

Cloud Infrastructure Architect

eCom Solutions, Inc.Atlanta, GA🇺🇸United StatesPosted 16 Sept 2026

Why This Role Stands Out

This hybrid Cloud Infrastructure Architect role offers a fantastic opportunity to leverage your expertise in cloud platforms and enterprise resilience, driving innovative AI integrations within a reputable company. You'll thrive if you have a strong background in Kubernetes, IaC, and observability, especially within regulated environments, so seize this chance to advance your career!

Quick Overview

Seniority
Mid Senior
Work mode
Hybrid
Location
Atlanta, GA, United States
Posted
17 hours ago
DockerAWSSplunkAnsibleAzureGoogle CloudGrafanaHelmKubernetesPrometheusTerraform

Job Description

Job Profile: Application Architect – AI Integration
Location: Atlanta, GA
Duration: 11 Months

The ideal candidate will have a strong background in cloud/platform architecture and enterprise resilience, with hands-on experience evaluating highly available infrastructure, DR/BCP readiness, Kubernetes environments, IaC practices, and observability platforms. Experience working within financial services or other regulated environments is highly desirable.

 

Required Skills

We are seeking an experienced Application Architect – AI Integration with strong expertise in cloud infrastructure, platform resilience, disaster recovery, observability, Kubernetes, and Infrastructure as Code.

Key Responsibilities

  • Assess AWS, Azure, Google Cloud Platform, on-premises, and vendor-hosted infrastructure against 8 Non-Functional Requirement (NFR) resilience domains, with a focus on identifying platform-layer gaps.
  • Evaluate platform architecture across networking, compute, storage, and resilience configurations, including:
    • Load balancers
    • Auto-scaling
    • Multi-AZ / Multi-Region architecture
    • Failover routing
    • High-availability configurations
  • Review and validate Disaster Recovery (DR) and Business Continuity Planning (BCP) documentation.
  • Validate:
    • RTO/RPO targets
    • Failover test evidence
    • DR testing results
    • Chaos engineering results
    • Application resilience for assigned applications
  • Assess observability maturity, including:
    • Alerting coverage
    • Distributed tracing
    • Log aggregation
    • SLO/SLA definitions
    • Incident detection and response latency
  • Evaluate container orchestration and Infrastructure as Code (IaC) practices, including:
    • Kubernetes cluster health and resilience
    • Terraform state management
    • Helm releases
    • Cluster-level availability and resilience
  • Identify infrastructure and platform risks, gaps, and areas requiring remediation.
  • B(Background Check) is required.

Nice-to-Have Skills

  • 10–15 years of overall IT / Infrastructure / Platform Engineering experience
  • 7+ years of Cloud Infrastructure Architecture experience with AWS, Azure, or Google Cloud Platform
  • 5+ years of Disaster Recovery / Business Continuity Planning design, testing, and validation in enterprise environments
  • 3+ years of experience within banking or financial services / regulated environments
  • 5+ years of Kubernetes, Docker, and container-native architecture
  • 3+ years of Infrastructure as Code experience with:
    • Terraform
    • Helm
    • Ansible
  • 3+ years of experience with observability and monitoring tools such as:
    • Dynatrace
    • Splunk
    • Prometheus
    • Grafana

 

Amit Mehra

Delivery Manager

Office: X 1018

Direct:

E-mail   :

Website :

Location : 2120 Crown Center Drive, Suite # 400,

               Charlotte, NC - 28227

 

Similar jobs