Haystack
← Back to Jobs
Technology
TC

Cloud Engineer

Trinite Consulting Group LLCFarmington Hills, MI🇺🇸United StatesPosted Sep 28, 2026

Quick Overview

Seniority
Mid Senior
Work mode
On Site
Location
Farmington Hills, MI, United States
Posted
3 days ago
AWSAgileAzureBashDatadogPagerDutyPowerShellPython

Job Description

Cloud Operations Analyst II
Hybrid - 3 days a week

Farmington Hills MI (Detroit suburb) 


Summary Statement

The Cloud Operations Analyst II is responsible for operating and enhancing the performance, availability, and reliability of cloud and on-premises infrastructure. This individual consults on more complex observability scenarios, streamlining server and batch operations, and contributing to the efficient management of data center resources.  Consults with cross-functional teams to identify opportunities for process improvements, implement best practices, and support critical business operations. 

 

Essential Tasks/Major Duties

·         Develop, implement, and maintain observability tools to monitor cloud and on-premises systems.  

·         Create dashboards, alerts, and reports to track system health, performance, and availability.  

·         Proactively leverage observability tools and identify opportunities.  

·         Analyze metrics and logs to identify trends, prevent potential issues, and optimize system performance.  

·         Collaborate with FinOps teams to monitor resource utilization and ensure cost-effective operations across cloud environments. 

·         Support the lifecycle of cloud and on-premises servers, including provisioning, patching, configuration, and decommissioning. 

·         Troubleshoot and resolve server-related issues, ensuring minimal downtime. Implement and enforce server security policies and compliance requirements. 

·         Schedule, monitor, and manage batch processes to ensure timely execution of critical tasks. 

·         Identify and resolve batch failures or delays, coordinating with relevant teams to ensure smooth operations.  

·         Optimize batch jobs for improved performance and resource utilization. 

·         Manage on-site and remote data center operations, ensuring proper functioning of hardware, power, cooling, and network infrastructure.  

·         Coordinate with vendors and service providers for hardware maintenance, replacements, and upgrades.  

·         Maintain accurate inventory of data center assets and ensure compliance with organizational standards. 

·         Participate in on-call rotations to address system incidents and outages promptly.  

·         Conduct root cause analysis and implement solutions to prevent recurrence of issues. Document and communicate incident resolution processes to relevant stakeholders. 

·         Work closely with cross-functional teams, including DevOps, Networking, and Application Development, to implement and maintain system integrations.  

·         Maintain and create comprehensive documentation for configurations, processes, and incident resolutions.  

·         Provide training and support to team members and other departments. 

 

Knowledge, Skills & Abilities

·         Bachelor’s degree in computer science, Information Technology, or a related field, or equivalent experience. 

·         3 years of experience working with monitoring and observability tools (e.g., Datadog, PagerDuty). 

·         Certified Datadog Fundamentals or equivalent experience required. 

·         Certified PagerDuty Administrator or equivalent experience required. 

·         3 years of experience in cloud operations or server management roles. 

·         Certified AWS SysOps Administrator or equivalent experience required. 

·         3 years of progressive server administration experience (Windows, Linux). 

·         3 years of experience in designing, implementing, and managing IT workload automation solutions to optimize scheduling, orchestration, and execution of enterprise workflows across on-prem and cloud environments.

·         Experience leveraging artificial intelligence to drive innovation and solve complex problems. Demonstrated ability to utilize AI-driven solutions that optimize processes, enhance decision-making, or create transformative business outcomes.

·         3 years working with cloud platforms (AWS, Azure, OCI).

·         Certified AWS Cloud Practitioner or equivalent experience required.

·         Strong experience with data center infrastructure and knowledge of best practices.  

·         Proficiency in scripting and automation tools (Python, Bash, PowerShell). 

·         Strong understanding of networking and identity management in cloud environments. 

·         Working knowledge of security best practices and compliance standards.

·         Working knowledge of agile methodologies. 

·         Excellent troubleshooting, problem-solving, and communication skills. 

Similar jobs