Quick Overview
Seniority
Mid Senior
Employment type
Full Time
Work mode
Remote
Location
Australia
DockerGCPOpenStackShellAWSELKAnsibleAzureGrafanaKubernetesPrometheusPythonRESTTerraform
Job Description
Upsun is the software factory for AI-human workflows. It is built for today's hybrid teams, where AI agents write and test code and humans focus on solving the problems that really matter. Developers, DevOps engineers, and platform teams use Upsun to build, ship, and scale confidently without wrestling with backend infrastructure. We give you your time back. You get:
Upsunners are a remote, global workforce, and we thrive in a multicultural team. We are committed to open source and an open, welcoming environment. Our team spans the globe and the experience spectrum.
What's our commonality, our cultural fabric? A curious spirit and a thirst for knowledge; an eagerness for innovative ideas and cultures. We believe we can build anything together in an environment that frees you to do your best work.
Our values:
We make a positive impact.
We aim for the stars.
We care for each other.
Impact of a Senior Site Reliability Engineer As a Senior Site Reliability Engineer at Upsun, you will lead the evolution of our cloud application platform from traditional cloud operations into a proactive, automation-driven SRE model. You will own critical engineering workstreams that enhance system reliability, scalability, and operational efficiency across multi-cloud environments. Partnering closely with engineering, product, and platform teams, you will embed reliability and performance into every stage of the software delivery lifecycle. In this role, you will anticipate architectural bottlenecks, drive infrastructure-as-code practices, and establish robust observability standards that ensure long-term system stability and uptime for our global users.
What to expect
This role includes on-call hours: One week every 4-5 weeks, 02:00 AM - 10:00 AM UTC (or 02:00 - 10:00 UTC). Weekend shift included in the one-week of on-call.
How we hire We know that a great hire won't meet every requirement that we've outlined. If you can see yourself elevating the team, we want to hear your story. Few of us would be here had we not taken a chance.
You can expect 4 interviews on Google Meet to follow the order below. Should you successfully move through the entire process you will have the opportunity to meet with a variety of Upsunners. Our goal is to ensure you can make the most informed decision on whether this role, and our culture aligns with what you're looking for in your future working environment.
What we offer
Voluntary Self-Identification For government reporting purposes, we ask candidates to respond to the below self-identification survey.Completion of the form is entirely voluntary. Whatever your decision, it will not be considered in the hiringprocess or thereafter. Any information that you do provide will be recorded and maintained in a confidential file.
As set forth in Upsun (via Remote Woman)'s Equal Employment Opportunity policy,we do not discriminate on the basis of any protected group status under any applicable law.
If you believe you belong to any of the categories of protected veterans listed below, please indicate by making the appropriate selection.As a government contractor subject to the Vietnam Era Veterans Readjustment Assistance Act (VEVRAA), we request this information in order to measurethe effectiveness of the outreach and positive recruitment efforts we undertake pursuant to VEVRAA. Classification of protected categoriesis as follows:
A "disabled veteran" is one of the following: a veteran of the U.S. military, ground, naval or air service who is entitled to compensation (or who but for the receipt of military retired pay would be entitled to compensation) under laws administered by the Secretary of Veterans Affairs; or a person who was discharged or released from active duty because of a service-connected disability.
A "recently separated veteran" means any veteran during the three-year period beginning on the date of such veteran's discharge or release from active duty in the U.S. military, ground, naval, or air service.
. click apply for full job details
- Predictable performance, even at scale
- Secure, compliant environments by default
- Real-time observability and profiling built in
- Cloning, configuration, and provisioning in seconds
- AI-ready features that plug directly into your stack
Upsunners are a remote, global workforce, and we thrive in a multicultural team. We are committed to open source and an open, welcoming environment. Our team spans the globe and the experience spectrum.
What's our commonality, our cultural fabric? A curious spirit and a thirst for knowledge; an eagerness for innovative ideas and cultures. We believe we can build anything together in an environment that frees you to do your best work.
Our values:
We make a positive impact.
We aim for the stars.
We care for each other.
Impact of a Senior Site Reliability Engineer As a Senior Site Reliability Engineer at Upsun, you will lead the evolution of our cloud application platform from traditional cloud operations into a proactive, automation-driven SRE model. You will own critical engineering workstreams that enhance system reliability, scalability, and operational efficiency across multi-cloud environments. Partnering closely with engineering, product, and platform teams, you will embed reliability and performance into every stage of the software delivery lifecycle. In this role, you will anticipate architectural bottlenecks, drive infrastructure-as-code practices, and establish robust observability standards that ensure long-term system stability and uptime for our global users.
What to expect
- Drive reliability & observability strategy: Architect and elevate system monitoring, alerting, and logging using Prometheus, Grafana, and ELK Stack, establishing actionable SLIs/SLOs aligned with core business metrics.
- Automate infrastructure & workflows: Eliminate operational toil by designing and implementing resilient, automated solutions using IaC tools like Terraform and Ansible across AWS, GCP, and Azure.
- Scale CI/CD & delivery pipelines: Optimize pipeline architectures for fast, secure, and zero-downtime releases, ensuring infrastructure resilience during high-volume deployment cycles.
- Lead incident response & post-mortems: Guide high-priority incident triage, drive blameless post-mortem analysis, and implement preventative measures to continuously improve system resiliency.
- Cross-functional leadership: Partner with product and software engineering teams to incorporate SRE best practices into product roadmaps.
- Champion technical innovation: Proactively identify performance bottlenecks and evaluate emerging technologies (e.g., eBPF, container orchestration) to optimize platform stability and performance.
- Time distribution: Follow a 4-week rotation balancing engineering and operations to focus on reliability, automation, and scalability through hands-on troubleshooting and engineering innovation.
- Senior SRE & Cloud Expertise: 5+ years of experience in Site Reliability Engineering, Cloud Operations, or DevOps, with proven experience owning reliability for production platforms at scale.
- Software Engineering & Tooling: Strong proficiency in Go or Python to build custom automation tools, custom controllers, or SRE platform components (beyond basic shell scripting).
- Deep Linux Internals Proficiency: Advanced hands-on knowledge of Linux operating system internals, kernel parameters, networking protocols, performance profiling, and system troubleshooting.
- Infrastructure as Code & Cloud Platforms: Deep expertise with cloud providers (AWS, GCP, Azure, or Openstack) with custom tooling built around cloud SDKs, and declarative infrastructure tools (e.g., Terraform) to manage distributed systems.
- Autonomous Ownership & Systems Thinking: Proven ability to anticipate operational risks, make architectural trade-offs, and lead technical infrastructure initiatives with minimal guidance.
- Collaborative Communication: Outstanding cross-functional communication skills with a track record of building alignment, and fostering an inclusive engineering culture.
- Experience with custom-built orchestration, edge, storage, and operational tooling in a dynamic environment.
- Experience with Docker and production Kubernetes cluster management or containerized deployment architectures.
- Familiarity with Platform-as-a-Service (PaaS) architectures or developer-facing cloud platforms.
This role includes on-call hours: One week every 4-5 weeks, 02:00 AM - 10:00 AM UTC (or 02:00 - 10:00 UTC). Weekend shift included in the one-week of on-call.
How we hire We know that a great hire won't meet every requirement that we've outlined. If you can see yourself elevating the team, we want to hear your story. Few of us would be here had we not taken a chance.
You can expect 4 interviews on Google Meet to follow the order below. Should you successfully move through the entire process you will have the opportunity to meet with a variety of Upsunners. Our goal is to ensure you can make the most informed decision on whether this role, and our culture aligns with what you're looking for in your future working environment.
- 45 Minutes with Talent Acquisition
- 60 Minutes with Hiring Manager
- 60 Minutes with Team (ICs)
- 60 Minutes with Senior Director, SRE
What we offer
- A product you can believe in - Join us in transforming how businesses build and manage web applications, driven making a positive impact as a proud B Corp.
- An Award-Winning Workplace - We've been recognized by Forbes' Top 30 Companies for Remote Jobs and France's Best Workplaces for Women.
- A culture that values your voice - Join a flexible, open, and inclusive work environment where your voice is encouraged, and your ideas shape our growth and evolution.
- A global team - Collaborate with colleagues from diverse backgrounds across the world, embracing different perspectives
- Benefits and perks - Make the most of what matters to you
- Company stock options
Voluntary Self-Identification For government reporting purposes, we ask candidates to respond to the below self-identification survey.Completion of the form is entirely voluntary. Whatever your decision, it will not be considered in the hiringprocess or thereafter. Any information that you do provide will be recorded and maintained in a confidential file.
As set forth in Upsun (via Remote Woman)'s Equal Employment Opportunity policy,we do not discriminate on the basis of any protected group status under any applicable law.
If you believe you belong to any of the categories of protected veterans listed below, please indicate by making the appropriate selection.As a government contractor subject to the Vietnam Era Veterans Readjustment Assistance Act (VEVRAA), we request this information in order to measurethe effectiveness of the outreach and positive recruitment efforts we undertake pursuant to VEVRAA. Classification of protected categoriesis as follows:
A "disabled veteran" is one of the following: a veteran of the U.S. military, ground, naval or air service who is entitled to compensation (or who but for the receipt of military retired pay would be entitled to compensation) under laws administered by the Secretary of Veterans Affairs; or a person who was discharged or released from active duty because of a service-connected disability.
A "recently separated veteran" means any veteran during the three-year period beginning on the date of such veteran's discharge or release from active duty in the U.S. military, ground, naval, or air service.
. click apply for full job details
Similar jobs
- AC
DevOps Manager
NewAccenture PLC
Melbourne, Victoria🇦🇺Hybrid1 hour agoPackerSeleniumSonarQube+5Technology - CO
Data Platform Engineer - MongoDB
NewCognizant
Sydney🇦🇺Hybrid1 hour agoMongoDBTechnology - CG
Senior MarTech Platform Engineer
NewCoStar Group, Inc.
Sydney🇦🇺Hybrid1 hour agoSnowflakeC#Java+1Technology - ZO
Cyber Security Platform Engineer: RunZero Specialist
NewZoho
Sydney🇦🇺Hybrid1 hour agoTechnology - ZO
Security Platform Engineer
NewZoho
Sydney🇦🇺On-site1 hour agoMFASSODNS+1Technology - LA
Principal Cloud & DevOps Engineer (Azure, NV1) - Remote
NewLAB3
Canberra, Australian Capital Territory🇦🇺Remote1 hour agoAzureIoTTerraformTechnology - TR
Platform Engineer Akuna Capital Sydney, Australia 3 hours ago
NewTradermath
Sydney🇦🇺Hybrid1 hour agoDynamoDBGCPAWS+11Technology - 8A
DevOps Manager
New8200 Accenture Australia P/L Company
Melbourne, Victoria🇦🇺Hybrid1 hour agoPackerSeleniumSonarQube+5Technology - SM
Principal AI Platform Engineer
NewSmartRecruiters, Inc.
Sydney🇦🇺On-site1 hour agoGCPAWSMLOps+9Technology - TO
Senior Site Reliability Engineer (Core)
NewTravelCenters of America
Melbourne, Victoria🇦🇺Hybrid1 hour agoAWSTechnology - NP
Security-Clearance DevOps Engineer Cloud, CI/CD & IaC
NewNixil Pty
Sydney🇦🇺Hybrid1 hour agoAgileAnsibleCloudFormation+2Technology - VL
Senior MarTech Platform Engineer
NewVisual Lease
Sydney🇦🇺Hybrid1 hour agoSnowflakeC#Java+1Technology