← Back to Jobs
Technology
Sr. Cloud Platform Systems Engineer (DevOps/SRE) | Mississauga, ON - Hybrid
Software Guidance & AssistanceMississauga, ON🇺🇸United StatesPosted 21 Jul 2026
Quick Overview
Work Type
Hybrid
Level
Mid Senior
Job Description
Software Guidance & Assistance, Inc., (SGA), is searching for a DevOps Engineer for a Right to Hire assignment with one of our premier Financial Services clients in Mississauga, ON.
Responsibilities :
Seeking a highly skilled Cloud Platform Systems Engineer to join our global Cloud Infrastructure and development team responsible for the stability, reliability, and performance of firm's mission-critical Enterprise Risk Technology (ERT). These Applications under ERT serve as the operational backbone for the firm's Risk businesses, including Compliance Risk, Operation Risk, Retail Risk, and Regulatory Pillars.
This is a pivotal role within a high-performing team dedicated to building a world-class Dev-Ops, infrastructure, and architecture group. Candidate will work with the latest technologies, AI-driven automation, and Site Reliability Engineering (SRE) principles to enhance our client experience, digitize services, and ensure our platforms can scale for future growth. In this global capacity, candidates will collaborate extensively with Application Development, Infrastructure setup and performance management, Product rollout using CI/CD pipeline, Business, and Operations teams to guarantee seamless service delivery.
SGA is an Equal Opportunity Employer and does not discriminate on the basis of Race, Color, Sex, Sexual Orientation, Gender Identity, Religion, National Origin, Disability, Veteran Status, Age, Marital Status, Pregnancy, Genetic Information, or Other Legally Protected Status. We are committed to providing access, equal opportunity, and reasonable accommodation for individuals with disabilities in employment, and our services, programs, and activities. Please visit our company to request an accommodation or assistance regarding our policy.
#LI-VV1
Responsibilities :
Seeking a highly skilled Cloud Platform Systems Engineer to join our global Cloud Infrastructure and development team responsible for the stability, reliability, and performance of firm's mission-critical Enterprise Risk Technology (ERT). These Applications under ERT serve as the operational backbone for the firm's Risk businesses, including Compliance Risk, Operation Risk, Retail Risk, and Regulatory Pillars.
This is a pivotal role within a high-performing team dedicated to building a world-class Dev-Ops, infrastructure, and architecture group. Candidate will work with the latest technologies, AI-driven automation, and Site Reliability Engineering (SRE) principles to enhance our client experience, digitize services, and ensure our platforms can scale for future growth. In this global capacity, candidates will collaborate extensively with Application Development, Infrastructure setup and performance management, Product rollout using CI/CD pipeline, Business, and Operations teams to guarantee seamless service delivery.
- Containerization and Orchestration: Deep expertise in deploying, managing, and scaling applications using container technologies, including Kubernetes and OpenShift.
- Service Mesh: Strong knowledge of service mesh concepts and hands-on experience with Istio for ingress and egress traffic management.
- Security Proficiency: Solid understanding of security concepts including Single Sign-On (SSO, SAML Secured SSO, MFA), SSG, COIN authentication, TLS/SSL certificate setup, and authorization protocols (AD/LDAP).
- WebSphere Application Server: Experience in managing WebSphere environments, including support activities like JNDI, connection pool, and JMS configurations, as well as heap and session management.
- Big Data Technologies: Experience supporting data platforms and ingestion pipelines using technologies such as Spark, Hive, and Starburst.
- Cloud Platforms: Experience with cloud migration projects and supporting applications on cloud platforms such as AWS or IBM Cloud.
- Incident and Problem Management: Lead the resolution of critical production incidents, including prioritization, timely escalation, and clear communication to all stakeholders. Conduct comprehensive post-mortems and root cause analyses to ensure permanent solutions and prevent recurrence.
- Technical and Business Support: Serve as a key point of contact for technical and business support for firm users. Address daily queries and issues, and proactively drive improvements in stability, efficiency, and risk management.
- Change and Release Management: Support the planning and execution of all system changes, including application releases, infrastructure maintenance, and continuity of business (COB) tests, ensuring production stability is maintained throughout the process.
- System Observability and Monitoring: Maintain and optimize the production monitoring and observability estate. Implement new features and leverage advanced analytics to enhance system visibility, proactive alerting, and predictive fault detection.
- Automation and Efficiency: Identify opportunities to reduce operational toil and mitigate risk by developing and implementing automation script, tools and processes. Familiarity with Infrastructure Architecture and automation IaC (infrastructure as code).
- Collaboration and Stability: Partner closely with development teams to recommend and implement architectural and code improvements that enhance application stability, performance, and recoverability.
- Bachelor's/University Degree or equivalent professional experience.
- 5 years of experience in an infrastructure role managing enterprise-level, business-critical applications.
- Prior direct experience with cloud migration projects and supporting applications on cloud platforms such as AWS or IBM Cloud.
- Strong expertise in the deployment, management, and scalability of applications using container technologies, including Kubernetes and OpenShift.
- Proven expertise in service mesh architecture with hands-on Istio experience for ingress and egress traffic control.
- Comprehensive knowledge of security concepts encompassing Single Sign-On (SSO, SAML-based SSO, MFA), SSG, COIN authentication, TLS/SSL certificate set up, and directory-based authorization protocols (AD/LDAP).
- Direct experience in administering WebSphere environments, supporting JNDI, connection pool, and JMS configurations, as well as performing heap and session management.
- Experience supporting data platforms and ingestion pipelines leveraging technologies such as Spark, Hive, and Starburst.
- Advanced proficiency in Unix/Linux environments, including complex shell and Python scripting.
- Demonstrated experience with enterprise monitoring tools (e.g., ITRS Geneos, AppDynamics) and log aggregation platforms (e.g., Splunk, ELK).
- Hands-on experience with containerization and orchestration platforms, specifically OpenShift and/or Kubernetes.
- Strong database skills, with experience in both relational (Oracle, MSSQL) and NoSQL (MongoDB) databases.
- In-depth knowledge of enterprise messaging solutions such as Tibco EMS, IBM MQ, or Kafka.
- Solid understanding of distributed application architecture, including networks, load balancers, storage, and authentication protocols (AD/LDAP).
- Experience working with and troubleshooting REST APIs.
- Exceptional written and verbal communication skills, with the ability to articulate complex technical issues to both technical and non-technical audiences.
- Proficiency in a high-level programming language (e.g., Python, Java) for developing automation solutions.
- Experience with modern observability platforms and standards, such as Prometheus, and Grafana.
- Familiarity with prompt engineering techniques for interacting with Generative AI and Large Language Models (LLMs).
SGA is an Equal Opportunity Employer and does not discriminate on the basis of Race, Color, Sex, Sexual Orientation, Gender Identity, Religion, National Origin, Disability, Veteran Status, Age, Marital Status, Pregnancy, Genetic Information, or Other Legally Protected Status. We are committed to providing access, equal opportunity, and reasonable accommodation for individuals with disabilities in employment, and our services, programs, and activities. Please visit our company to request an accommodation or assistance regarding our policy.
#LI-VV1
Skills
MongoDB
Oracle
Shell
AWS
ELK
MFA
SAML
SSO
Service Mesh
Splunk
Generative AI
Grafana
Hive
Istio
Java
Kafka
Kubernetes
LDAP
Prometheus
Python
REST
Similar jobs
Senior Databricks Platform Engineer
Electronic Consulting Services, Inc (ECS Federal) · Arlington, United States
55 minutes ago$140k - $165k/yrData Platform Engineer - Private Cloud / Kubernetes
Motion Recruitment Partners, LLC · Charlotte, United States
55 minutes agoDevOps Engineer
Robert Half · Durham, United States
56 minutes agoDevOps Engineer / Azure / Linux / Kubernetes
Motion Recruitment Partners, LLC · Chelmsford, United States
1 hour agoSenior Devops AI tools Engineer
Celestica International LP · United States
1 hour ago$101k - $150k/yrDevOps Engineer
Robert Half · Newton, United States
1 hour ago