Why This Role Stands Out
This hybrid SRE role at TSQ Systems Inc offers a fantastic opportunity to enhance your skills in automation, cloud infrastructure, and system reliability, contributing directly to the performance of critical applications. You'll thrive here if you're a proactive problem-solver with a passion for building robust and scalable systems, and we encourage you to apply to join their innovative team.
Quick Overview
Job Description
SRE
Introduction
As a Site Reliability Engineer (SRE), you will play a crucial role in ensuring the high availability, scalability, and performance of applications and infrastructure in production environments. You will work closely with development and DevOps teams to automate deployment processes, monitor systems, and troubleshoot production issues.
Responsibilities
- Ensure high availability, scalability, and performance of applications and infrastructure in production environments.
- Monitor systems using tools like Prometheus, Grafana, Datadog, or CloudWatch and respond to incidents.
- Automate deployment, monitoring, and operational tasks using CI/CD pipelines and scripting (Python, Bash, etc.).
- Troubleshoot production issues, perform root cause analysis, and implement preventive measures.
- Collaborate with development and DevOps teams to improve system reliability and release processes.
- Manage cloud infrastructure (AWS, Azure, or Google Cloud Platform) and implement best practices for security and performance.
Requirements
Required Skills:
- Experience with monitoring tools such as Grafana.
- Proficiency in scripting languages like Python, Bash, etc.
- Strong problem-solving and troubleshooting skills.
- Knowledge of CI/CD pipelines.
- Experience with cloud platforms like AWS, Azure, or Google Cloud Platform.
Preferred Skills:
- Certifications in relevant technologies.
- Experience with containerization technologies like Docker, Kubernetes.
- Knowledge of networking and security best practices.
- Experience with infrastructure as code tools like Terraform.
- Excellent communication and collaboration skills.
Similar jobs
- BL
Sr. Software Engineer, Site Reliability
NewBloomerang
Remote🇺🇸Remote2 hours ago401kNode.jsPHP+9Technology - BF
Senior DevOps Engineer
NewBeyond Finance
Chicago🇺🇸4 hours agoDockerAWSLinear+11Technology - PG
Cloud Platform Engineer Kubernetes, Onsite - 69995
NewPRIMUS Global Services Inc.
Plano, TX🇺🇸On-site18 hours agoShellAWSNginx+9Technology - TA
DevOps Software Engineer, TS/SCI with Poly with Security Clearance
NewTalentOps
Annapolis Junction, MD🇺🇸$160k - $230k/yrRemote18 hours agoDockerMariaDBMySQL+13Technology - NI
Sr DevOps Engineer
NewNisum
Johns Creek, GA🇺🇸Hybrid18 hours agoEncryptionScrumAgile+9Technology - OC
Software Engineer (DevOps) - Hybrid - FS Poly with Security Clearance
Our client
Annapolis Junction, MD🇺🇸Remote1 week agoDockerShellAnsible+5Technology