Quick Overview
Job Description
About the Role
This is a high-ownership infrastructure role at a fast-growing AI-powered consumer intelligence platform serving enterprise retail brands. You'll architect and maintain the cloud systems that underpin the product, working closely with backend and ML engineers to keep the platform secure, scalable, and highly reliable.
What You'll Do
Architect and build secure, scalable cloud infrastructure on AWS using Infrastructure-as-Code (Terraform/Terragrunt).
Design, implement, and maintain robust CI/CD pipelines, converting manual processes into automated, repeatable workflows.
Own production reliability — set up observability stacks, define SLI/SLOs, and lead incident response.
Champion GitOps workflows across all infrastructure deployments.
Build compliance-ready infrastructure (SOC 2, GDPR) with strong IAM practices and secrets management.
Optimize cloud costs while sustaining performance and reliability at scale.
Create developer tooling and documentation that accelerates engineering workflows across the team.
Mentor team members on DevOps best practices and help shape infrastructure strategy.
What We're Looking For
5+ years of hands-on infrastructure engineering experience with end-to-end ownership — not limited to a single domain.
Expert-level, production-active Kubernetes experience (EKS strongly preferred) within the last 12 months.
Proven experience building and managing AWS infrastructure for B2B SaaS products.
Deep hands-on expertise with Terraform and/or Terragrunt.
Experience designing and maintaining CI/CD pipelines end-to-end.
Broad infrastructure scope covering compute, networking, storage, and observability.
Familiarity with monitoring, logging, and observability platforms (e.g., Prometheus, ELK, Datadog).
Background at VC-backed startups or mid-sized companies — not exclusively large enterprises.
No visa sponsorship available; candidates must be independently authorized to work.
Compensation & Benefits
Salary: $160,000 – $200,000 USD annually. Visa sponsorship is not available for this role.
Location
On-site in New York, NY.
Similar jobs
- ST
Site Reliability Engineer - W2 Contract
NewSDVS Technologies LLC
Jersey City, NJ🇺🇸Hybrid22 hours agoSpinnakerSplunkAnsible+6Technology - CS
DevOps Engineer
NewCynet Systems
Toronto, ON🇺🇸Hybrid22 hours agoDockerSQLSQL Server+11Technology - CT
MLOps / Platform Engineer
NewConch Technologies
Charlotte, NC🇺🇸On-site22 hours agoDockerMicroservicesMongoDB+17Technology - SE
Staff Site Reliability Engineer, Government
NewSentinelOne
United States - Remote🇺🇸RemoteYesterdayGCPRubyAWS+13Technology - IN
Senior Observability Platform Engineer Remote(EST Hours) w2 Position
NewIntone Networks Inc.
United States🇺🇸Remote22 hours agoDatadogTechnology - WH
Google Cloud Platform DevOps AI Engineer
NewWhiztek Corp
Schaumburg, IL🇺🇸On-site22 hours agoOAuthGoogle CloudGoogle WorkspaceTechnology