Senior Azure Messaging Platform Engineer
Quick Overview
Job Description
Title - Senior Azure Messaging Platform Engineer
Location: Remote (PST Preferred)
Duration: 12-Month Contract
Overview
We are seeking a Senior Azure Messaging Platform Engineer to design, implement, and operationally support highly available, mission-critical Azure Service Bus messaging platforms.
This is a hands-on senior individual contributor role focused specifically on Azure Service Bus Premium, enterprise messaging, Geo-Replication, High Availability, Disaster Recovery, failover/failback, automation, and production resiliency.
The successful candidate must be capable of independently owning work from solution design and engineering through production implementation, DR validation, documentation, and operational handoff.
This is not a general Azure infrastructure role. We are specifically looking for an engineer with deep experience designing and operating enterprise messaging platforms using Azure Service Bus.
Key Responsibilities
Azure Service Bus & Messaging Engineering
- Design, build, configure, and support Azure Service Bus Premium environments supporting mission-critical enterprise workloads.
- Design and implement Azure Service Bus Geo-Replication and multi-region resiliency across Azure regions.
- Engineer highly available messaging solutions that meet defined RPO/RTO objectives while minimizing message loss during regional or service disruptions.
- Design and support Azure Service Bus:
- Queues
- Topics
- Subscriptions
- Dead Letter Queues (DLQ)
- Message replay and recovery
- Retry strategies
- Duplicate detection and duplicate-message handling
- Sessions and message ordering
- Idempotent message-processing patterns
- Message expiration and operational recovery
- Troubleshoot messaging delivery, latency, connectivity, sequencing, and application-integration issues across distributed systems.
- Establish messaging architecture standards, resiliency patterns, and operational best practices.
High Availability, Disaster Recovery & Geo-Replication
- Design and implement HA/DR architecture for Azure Service Bus across Azure regions.
- Develop and maintain detailed DR runbooks covering failover, failback, validation, recovery, and escalation procedures.
- Plan and execute scheduled DR exercises and production resiliency testing.
- Execute and validate Azure Service Bus regional failover and failback procedures.
- Verify message flow, application connectivity, namespace configuration, subscriptions, and downstream processing following failover.
- Validate solutions against defined RTO and RPO requirements.
- Document DR test results, findings, gaps, risks, and recommended remediation.
- Drive identified remediation through implementation and retesting.
- Ensure DR capabilities are operationally repeatable and supportable rather than existing only as architecture documentation.
Infrastructure as Code & Automation
- Develop and maintain Azure Service Bus infrastructure using Terraform.
- Enhance existing infrastructure-as-code frameworks and create reusable Terraform modules.
- Automate deployment, configuration, validation, failover/failback, monitoring, and operational processes.
- Develop automation using:
- Python
- PowerShell
- Azure CLI
- Reduce manual operational dependencies through repeatable automation and documented procedures.
Reliability, Monitoring & Production Operations
- Provide senior-level engineering and L3 production support for business-critical messaging services.
- Implement and maintain monitoring, alerting, dashboards, and observability for Azure Service Bus.
- Monitor key messaging indicators including:
- Queue and subscription depth
- Dead-letter activity
- Message processing failures
- Retry patterns
- Throughput
- Latency
- Availability
- Capacity and resource utilization
- Configure and support observability solutions including Dynatrace and Azure-native monitoring capabilities.
- Troubleshoot complex production messaging and distributed-system incidents.
- Participate in root-cause analysis and implement permanent corrective actions.
Platform Ownership & Delivery
- Independently own technical work from requirements and architecture through implementation and production readiness.
- Produce technical designs, implementation plans, operational documentation, and support procedures.
- Coordinate changes with application, cloud, networking, security, and middleware engineering teams.
- Drive solutions through testing, change management, production implementation, validation, and operational handoff.
- Mentor engineers and provide technical leadership around enterprise messaging and resiliency practices.
Required Qualifications
- 7+ years of engineering experience supporting cloud, middleware, integration, or enterprise messaging platforms.
- Deep hands-on experience with Azure Service Bus, including production implementation and operations.
- Strong experience with Azure Service Bus Premium.
- Demonstrated experience implementing Geo-Replication and multi-region Azure Service Bus architectures.
- Hands-on experience designing and executing:
- High Availability
- Disaster Recovery
- Regional failover
- Failback
- DR testing
- Strong understanding of RPO/RTO, resiliency, and business-continuity requirements.
- Deep understanding of enterprise messaging concepts including:
- Queues, topics, and subscriptions
- DLQ management and message replay
- Retry strategies
- Duplicate handling
- Sessions and message ordering
- Idempotency
- Message delivery guarantees
- Messaging monitoring and observability
- Advanced Terraform experience.
- Strong automation and scripting experience using Python, PowerShell, and/or Azure CLI.
- Experience supporting mission-critical production environments and handling senior/L3 escalations.
- Strong troubleshooting skills across distributed applications, messaging platforms, networking, and cloud infrastructure.
- Demonstrated ability to independently deliver solutions from design through production implementation and operational handoff.
Preferred Qualifications
- Experience with IBM MQ or another enterprise messaging platform.
- IBM MQ experience with HA/DR, message recovery, queue management, or enterprise integration is strongly preferred.
- Previous experience working within middleware, integration, or messaging platform engineering teams.
- Experience with Azure integration services and enterprise application integration patterns.
- Experience building automated DR validation or operational automation.
- Experience with Dynatrace or equivalent enterprise observability platforms.
- Experience working with highly available, 24×7 mission-critical systems.
- Airline or travel industry experience is a plus, but not required.
Technical Environment
- Azure Service Bus Premium
- Azure Service Bus Geo-Replication
- Microsoft Azure
- Azure Regions / Multi-Region Architecture
- Terraform
- Python
- PowerShell
- Azure CLI
- Dynatrace
- Azure Monitor
- Enterprise Messaging
- Queues / Topics / Subscriptions
- DLQ / Replay / Retry
- Sessions / Ordering
- Duplicate Handling / Idempotency
- High Availability (HA)
- Disaster Recovery (DR)
- RPO / RTO
- Failover / Failback
- IBM MQ — Preferred
- TIBCO — Plus
Candidate Profile
The ideal candidate is a senior messaging engineer first and an Azure engineer second. We are looking for someone who has actually built, operated, failed over, recovered, and troubleshot Azure Service Bus in production, and who can take ownership of the platform without requiring significant day-to-day technical direction.
- For information on benefits, equal opportunity employment, and location-specific applicant notices, click
At SPECTRAFORCE, we are committed to maintaining a workplace that ensures fair compensation and wage transparency in adherence with all applicable state and local laws.This position''''s pay range is $70.00/hr - $75.00/hr.
Skills
Similar jobs
DevOps Engineer [AQ-12975]
Aquent Talent · Phoenix, United States
37 minutes agoSenior Cloud Architect/Azure/DevOps with Security Clearance
Oteemo Inc · College Park, United States
1 hour agoLead/Senior DevOps Engineer (AI OPS AUTOMATION)
Openmind Technologies · United States
1 hour agoPrincipal DevOps Engineering Manager
DTCC · Jersey City, United States
1 hour agoDevSecOps Engineer - RedHat Openshift with Security Clearance
Woodside Staffing Solutions & Consulting · Tullahoma, United States
1 hour ago$100k - $110k/yrSite Reliability Engineer 2
Oracle Corporation · Nashville, United States
1 hour ago$69.8k - $148.3k/yr