Haystack
← Back to Jobs
Technology

Sr. Cloud Platform Systems Engineer (DevOps/SRE) | Mississauga, ON - Hybrid

Software Guidance & AssistanceMississauga, ON🇺🇸United StatesPosted 21 Jul 2026

Quick Overview

Work Type
Hybrid
Level
Mid Senior

Job Description

Software Guidance & Assistance, Inc., (SGA), is searching for a DevOps Engineer for a Right to Hire assignment with one of our premier Financial Services clients in Mississauga, ON.

Responsibilities :
Seeking a highly skilled Cloud Platform Systems Engineer to join our global Cloud Infrastructure and development team responsible for the stability, reliability, and performance of firm's mission-critical Enterprise Risk Technology (ERT). These Applications under ERT serve as the operational backbone for the firm's Risk businesses, including Compliance Risk, Operation Risk, Retail Risk, and Regulatory Pillars.

This is a pivotal role within a high-performing team dedicated to building a world-class Dev-Ops, infrastructure, and architecture group. Candidate will work with the latest technologies, AI-driven automation, and Site Reliability Engineering (SRE) principles to enhance our client experience, digitize services, and ensure our platforms can scale for future growth. In this global capacity, candidates will collaborate extensively with Application Development, Infrastructure setup and performance management, Product rollout using CI/CD pipeline, Business, and Operations teams to guarantee seamless service delivery.
  • Containerization and Orchestration: Deep expertise in deploying, managing, and scaling applications using container technologies, including Kubernetes and OpenShift.
  • Service Mesh: Strong knowledge of service mesh concepts and hands-on experience with Istio for ingress and egress traffic management.
  • Security Proficiency: Solid understanding of security concepts including Single Sign-On (SSO, SAML Secured SSO, MFA), SSG, COIN authentication, TLS/SSL certificate setup, and authorization protocols (AD/LDAP).
  • WebSphere Application Server: Experience in managing WebSphere environments, including support activities like JNDI, connection pool, and JMS configurations, as well as heap and session management.
  • Big Data Technologies: Experience supporting data platforms and ingestion pipelines using technologies such as Spark, Hive, and Starburst.
  • Cloud Platforms: Experience with cloud migration projects and supporting applications on cloud platforms such as AWS or IBM Cloud.
  • Incident and Problem Management: Lead the resolution of critical production incidents, including prioritization, timely escalation, and clear communication to all stakeholders. Conduct comprehensive post-mortems and root cause analyses to ensure permanent solutions and prevent recurrence.
  • Technical and Business Support: Serve as a key point of contact for technical and business support for firm users. Address daily queries and issues, and proactively drive improvements in stability, efficiency, and risk management.
  • Change and Release Management: Support the planning and execution of all system changes, including application releases, infrastructure maintenance, and continuity of business (COB) tests, ensuring production stability is maintained throughout the process.
  • System Observability and Monitoring: Maintain and optimize the production monitoring and observability estate. Implement new features and leverage advanced analytics to enhance system visibility, proactive alerting, and predictive fault detection.
  • Automation and Efficiency: Identify opportunities to reduce operational toil and mitigate risk by developing and implementing automation script, tools and processes. Familiarity with Infrastructure Architecture and automation IaC (infrastructure as code).
  • Collaboration and Stability: Partner closely with development teams to recommend and implement architectural and code improvements that enhance application stability, performance, and recoverability.
Required Skills:
  • Bachelor's/University Degree or equivalent professional experience.
  • 5 years of experience in an infrastructure role managing enterprise-level, business-critical applications.
  • Prior direct experience with cloud migration projects and supporting applications on cloud platforms such as AWS or IBM Cloud.
  • Strong expertise in the deployment, management, and scalability of applications using container technologies, including Kubernetes and OpenShift.
  • Proven expertise in service mesh architecture with hands-on Istio experience for ingress and egress traffic control.
  • Comprehensive knowledge of security concepts encompassing Single Sign-On (SSO, SAML-based SSO, MFA), SSG, COIN authentication, TLS/SSL certificate set up, and directory-based authorization protocols (AD/LDAP).
  • Direct experience in administering WebSphere environments, supporting JNDI, connection pool, and JMS configurations, as well as performing heap and session management.
  • Experience supporting data platforms and ingestion pipelines leveraging technologies such as Spark, Hive, and Starburst.
  • Advanced proficiency in Unix/Linux environments, including complex shell and Python scripting.
  • Demonstrated experience with enterprise monitoring tools (e.g., ITRS Geneos, AppDynamics) and log aggregation platforms (e.g., Splunk, ELK).
  • Hands-on experience with containerization and orchestration platforms, specifically OpenShift and/or Kubernetes.
  • Strong database skills, with experience in both relational (Oracle, MSSQL) and NoSQL (MongoDB) databases.
  • In-depth knowledge of enterprise messaging solutions such as Tibco EMS, IBM MQ, or Kafka.
  • Solid understanding of distributed application architecture, including networks, load balancers, storage, and authentication protocols (AD/LDAP).
  • Experience working with and troubleshooting REST APIs.
  • Exceptional written and verbal communication skills, with the ability to articulate complex technical issues to both technical and non-technical audiences.
Preferred Skills:
  • Proficiency in a high-level programming language (e.g., Python, Java) for developing automation solutions.
  • Experience with modern observability platforms and standards, such as Prometheus, and Grafana.
  • Familiarity with prompt engineering techniques for interacting with Generative AI and Large Language Models (LLMs).
SGA is a technology and resource solutions provider driven to stand out. We are a women-owned business. Our mission: to solve big IT problems with a more personal, boutique approach. Each year, we match consultants like you to more than 1,000 engagements. When we say let's work better together, we mean it. You'll join a diverse team built on these core values: customer service, employee development, and quality and integrity in everything we do. Be yourself, love what you do and find your passion at work. Please find us at .

SGA is an Equal Opportunity Employer and does not discriminate on the basis of Race, Color, Sex, Sexual Orientation, Gender Identity, Religion, National Origin, Disability, Veteran Status, Age, Marital Status, Pregnancy, Genetic Information, or Other Legally Protected Status. We are committed to providing access, equal opportunity, and reasonable accommodation for individuals with disabilities in employment, and our services, programs, and activities. Please visit our company to request an accommodation or assistance regarding our policy.

#LI-VV1

Skills

MongoDB
Oracle
Shell
AWS
ELK
MFA
SAML
SSO
Service Mesh
Splunk
Generative AI
Grafana
Hive
Istio
Java
Kafka
Kubernetes
LDAP
Prometheus
Python
REST

Similar jobs