Quick Overview
Job Description
Who We Are & Why Join Us
Avathon is the leading Industrial AI autonomy platform, helping customers across heavy industries -- energy, mining, manufacturing, aerospace, defense, and logistics -- accelerate the journey toward autonomous operations. Our platform is built on a Computational Knowledge Graph foundation that contextualizes and connects operational data across siloed systems, bringing together time series, structured, unstructured, and machine vision data to power AI-driven applications in asset performance management, supply chain intelligence, visual AI, and global trade management. With capabilities spanning digital twins, normal behavior modeling, natural language processing, and computer vision, Avathon delivers real-time predictive intelligence and agentic decision-making at industrial scale.
Cutting-Edge AI Innovation -- Join a team at the forefront of AI, developing groundbreaking solutions that shape the future. High-Growth Environment -- Thrive in a fast-scaling startup where agility, collaboration, and rapid professional growth are the norm. Meaningful Impact -- Work on AI-driven projects that drive real change across industries and improve lives.
Learn more at: avathon.com
Position Summary:
We are seeking an experienced DevOps Engineer to join our dynamic team operating in an agile development environment. The ideal candidate will have a solid foundation in DevOps across building software pipelines, creating repeatable deployments, and experience with various infrastructure components in the cloud. As a Sr. DevOps Engineer you will be working closely with team members across domains with deep technical skills and passion for AI.
You Will:
- Design and continuously improve the infrastructure for cloud-based services and client interfaces
- Manage the day-to-day operations of our build, testing, and continuous integration environment
- Work with internal IT, Software Engineers, Cloud Architects, and fellow DevOps Engineers to build & maintain dev & prod environments
- Implement best practices for always-up, always-available services
- Proactively communicate project & task status to project stakeholders
- Occasional on-call support and customer meetings may include irregular hours as needed
You'll Have Skills:
- Kubernetes Architecture & Operations: 2+ years of experience managing production-grade clusters (GKE preferred), including expertise in K8s internals, Helm chart creation, Service Mesh (Istio), and autoscaling strategies (HPA/VPA).
- GitOps & CI/CD: Proven ability to design container-native pipelines and manage cluster state using GitOps methodologies (ArgoCD) alongside standard CI tools.
- Infrastructure as Code (IaC): Experience on Terraform for provisioning cloud resources and bootstrapping clusters.
- Observability & Reliability: Hands-on experience on full-stack observability for microservices using Prometheus, Grafana, ELK/Loki, and distributed tracing to ensure system reliability and rapid incident response.
- Core Engineering Foundation: 2+ years in DevOps/Linux Administration with strong scripting proficiency (Python, Bash, or Go) and a track record of automating workflows in fast-paced cloud environments.
**A PLUS if you provide the link to your GitHub Website
Avathon is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, pregnancy, genetic information, disability, status as a protected veteran, or any other protected category under applicable federal, state, and local laws.
Avathon is committed to providing reasonable accommodations throughout the recruiting process. If you need a reasonable accommodation, please contact us to discuss how we can assist you.
Similar jobs
- KA
Senior DevOps Engineer (India)
NewAuto ApplyKarat
Bengaluru🇮🇳Remote20 hours agoDockerAWSNew Relic+10Technology - CL
Site Reliability Engineer
NewAuto ApplyCharger Logistics Inc
India🇮🇳Remote2 days agoDockerGCPMicroservices+23Technology - ZE
Lead Site Reliability Engineer - Platform Engineering / SRE
NewAuto ApplyZenoti
Hyderabad🇮🇳Hybrid2 days agoDockerMicroservicesTeamCity+28Technology - OK
Staff SRE for K8s Platform Team (AWS, Kubernetes, Platform Creation, Helm, Karpenter, Istio)
NewAuto ApplyOkta
Bengaluru🇮🇳Hybrid2 days agoDockerMicroservicesSpinnaker+18Technology - CO
DevOps Engineer-II
NewAuto ApplyCommerceIQ
Bengaluru🇮🇳Hybrid2 days agoDockerGCPRuby+22Technology - BG
DevOps (CI/CX)
NewAuto ApplyBosch Group
hosur road bangalore🇮🇳On-site2 days agoSQLSeleniumSonarQube+9Technology - NI
Specialist, Application Support - GCP Site Reliability Engineering
NewAuto ApplyNielsenIQ
Chennai, TN🇮🇳On-site2 days agoGCPSOAPShell+11Technology - CG
IT engineer Data & Analytics DevOps
NewAuto ApplyContinental Group Sector ContiTech
Bangalore, Karnataka🇮🇳Hybrid2 days agoScalaMachine LearningAzure+5Technology - EU
Senior DevOps Engineer
Auto ApplyEurofins
Bengaluru, KA🇮🇳On-site5 months agoMFASonarQubeActive Directory+9Technology - LY
Platform Engineer - Backend
Auto ApplyLyric
Chennai🇮🇳Remote3 days agoLESSPythonSchedulingTechnology - OK
Staff Site Reliability Engineer
Auto ApplyOkta
Bengaluru🇮🇳Hybrid5 days agoGCPMySQLSQL+23Technology - ME
AI Security & Platform Engineer
Auto ApplyMetaforms
Bengaluru🇮🇳On-site3 days agoOAuthSOC 2Assembly+3Technology