Why This Role Stands Out
This remote SRE Leader role offers significant growth opportunities by leveraging your AI experience to build robust and scalable applications. You'll thrive here if you're passionate about driving innovation and collaborating within a dynamic tech environment. Apply to shape the future of technology with a leading company!
Quick Overview
Seniority
Mid Senior
Work mode
Hybrid
Location
United States
Posted
3 weeks ago
MicroservicesAWSSplunkAgileDatadogJavaJenkins
Job Description
Job Title: SRE Leader with AI Experience
Location- Remote
Job Description-
Responsibilities-
- Collaborate with cross-functional teams to design, develop, test and deploy scalable, reliable, and high-performance applications.
- Implement best practices for application design, testing, deployment, monitoring, and maintenance to ensure optimal performance and availability.
- Develop automation tools and scripts to streamline processes and improve efficiency in application lifecycle management.
- Conduct thorough analysis of application performance metrics and system health to identify areas for optimization and enhancement.
- Proactively identify and address potential issues and bottlenecks in the application architecture to prevent downtime and service disruptions.
- Stay updated on emerging technologies and industry trends to drive continuous improvement and innovation in application design and development practices.
- Familiarity with Agile methodologies and DevOps practices
Requirement-
(Experience, Qualification, Knowledge & Skills)
- 10 + years of IT experience with Java Web/Enterprise projects
- Good Understanding of Runtime Support and Application Development lifecycle
- Good understanding about the Concepts / Design for reliability (i.e. Automation, Scaling, Auto Recovery, Redundancy, Availability)
- Should have a good understanding of App/Infra capacity planning, SLA/SLOs
- Strong proficiency in programming languages such as Java, sprint boot, microservices and Java Frameworks
- Solid understanding of AWS cloud computing platforms, AWS Certification preferred.
- Proficiency in implementing and maintaining CI/CD pipelines (preferred Cloudbees / Jenkins)
- Excellent understanding on APM/ observability tools like NewRelic/ Datadog/ Dynatrace, Splunk etc.
- Hands on with PCF - App-Pilot & Presto
- Should have skills to doing POCs and reflect POC to enterprise scale
- Excellent problem-solving skills and attention to detail.
- Good understanding of AI technologies
- Knowledge and experience on the Gen AI tools and technology like GHCP, claud, AWS bedrock
- focuses on reliability, scalability, and performance of large-scale
- Strong communication and collaboration skills
Similar jobs
- LS
Senior Engineer - Site Reliability Engineering
NewLondon Stock Exchange Group
Allen, TX🇺🇸Hybrid16 hours agoAWSELKActive Directory+9Technology - LS
Technical Lead - Site Reliability Engineering
NewLondon Stock Exchange Group
Allen, TX🇺🇸Hybrid16 hours agoAWSELKAzure+8Technology - JT
Senior Cloud Platform Engineer
NewJaven Technologies, Inc
Farmington Hills, MI🇺🇸On-site16 hours agoAWSPythonTerraform+1Technology - CD
Sr. Site Reliability Engineer
NewCosmic-I LLC DBA Northern Base
United States🇺🇸Remote16 hours agoDockerRubyAWS+15Technology - 2C
Cloud Platform Engineer
New22nd Century Technologies, Inc.
Washington, DC🇺🇸Hybrid16 hours agoAWSAzurePowerShell+1Technology - Y-
Lead Site Reliability Engineer
NewYoh - A Day & Zimmerman Company
Dallas, TX🇺🇸Remote16 hours agoNode.jsGitHub ActionsGoogle Cloud+3Technology