Quick Overview
Seniority
Mid Senior
Work mode
Hybrid
Location
Washington, DC, United States
Posted
16 hours ago
AWSAnsibleAzurePowerShellPython
Job Description
Job Title: Cloud Systems Developer / Engineer
Location: Washington, DC
Duration: Full Time
Position Summary:
- Designs, develops, and maintains automation scripts and engineering solutions for cloud and platform operations to improve efficiency and reduce manual effort.
- The role works in an enterprise hosting environment and requires practical capability in cloud and platform solution development; scripting, automation, and infrastructure-as-code (e.g., powershell, python, ansible).
Responsibilities:
- supports provisioning, integration, optimization, and onboarding of resources.
- performs Tier 2/3 troubleshooting and root cause analysis.
- executes changes via ITSM processes.
- ensures accurate configuration, documentation, and compliance. Supporting Deliverables: D-TEC-02. D-TEC-01 and D-TEC-07. D-TEC-03 and D-TEC-04. D-ITSM-01, D ITSM-02, D-RE-01 and D-RE-02.
Core Skills:
- Cloud and platform solution development
- scripting, automation, and infrastructure-as-code (e.g., PowerShell, Python, Ansible)
- cloud platform management (AWS/Azure)
- virtualization and system integration
- performance optimization and monitoring
- and ITIL-aligned ITSM processes. Roles and Responsibilities: Designs, develops, and maintains automation scripts and engineering solutions for cloud and platform operations to improve efficiency and reduce manual effort
- supports provisioning, integration, optimization, and onboarding of resources
- performs Tier 2/3 troubleshooting and root cause analysis
Required Skills:
- Provide Tier 2/3 operational support, troubleshoot complex issues, perform root cause analysis, and drive issues through resolution.
- Use ITIL-aligned ITSM processes for incidents, changes, configuration tracking, approvals, and operational documentation.
- Keep system/configuration records, technical documentation, status information, and implementation details accurate and current.
- Support onboarding, implementation, and rollout activities while coordinating dependencies with engineering and operations teams.
- Monitor service/system performance, identify trends or risks, and take or recommend corrective action before issues affect service delivery.
Required Experience:
- Systems and services remain stable, supportable, documented, and compliant with operational standards.
- Incidents and technical issues are diagnosed efficiently, root causes are addressed, and recurring problems are reduced.
- Changes, patches, provisioning, maintenance, and onboarding activities are executed accurately through approved processes.
Deliverables:
- Provide accurate operational/technical inputs, maintain supporting records, and complete assigned sections or artifacts early enough for internal quality review and Government submission.
- D-TEC-02 Automation Scripts and Documentation: Develop and maintain automation code/scripts and supporting documentation in a Government-approved repository. Timing: As developed.
- D-TEC-01 Engineering Design Document / Implementation Plan: Document architecture/design, test plans, and implementation steps for new services or significant modifications and support Change Management approval. Timing: As required.
- D-TEC-07 Operational Rollout Plan: Plan operational integration of new technology, including SOPs, training, monitoring integration, cutover steps, and operational readiness. Timing: As required for new technology/onboarding.
- D-TEC-03 Monthly System Performance & Availability Report: Provide system health, availability, SLA uptime, CPU/memory/storage trends, capacity forecasts, and service interruption summaries. Timing: Monthly, as agreed upon award.
- D-TEC-04 Root Cause Analysis (RCA) Report: Document the incident timeline, root cause, business impact, corrective actions, and recurrence-prevention measures for P1/P2 or major service-impacting incidents. Timing: Within 72 hours of incident resolution.
- D-ITSM-01 CMDB Accuracy Report: Report CMDB audit results, data-accuracy metrics, discrepancies, and remediation actions needed to maintain accurate configuration-item inventory. Timing: First day of each quarter.
- D-RE-01 Quarterly Operational Improvement Roadmap: Identify the top five chronic operational issues/toil and define systematic improvements to eliminate them. Timing: First day of each quarter.
- D-RE-02 Service Level Objective (SLO) Report: Track compliance with PaaS SLAs and internal SLOs, including latency, throughput, and error rates. Timing: First day of each month.
Candidate Profile:
- Hands-on experience performing the core responsibilities and using the technologies/processes identified above in an enterprise IT environment.
- Able to work independently on assigned activities while coordinating effectively across engineering, operations, service management, and program teams.
- Strong troubleshooting, documentation, communication, and follow-through skills; comfortable working in structured change and compliance environments.
Similar jobs
- NG
Principal/Sr. Principal Cloud Engineer with Security Clearance
Northrop Grumman
Rome, NY🇺🇸$103.6k - $155.4k/yrHybrid4 weeks agoDockerExpressMicroservices+14Technology - BA
AWS Cloud Engineer, Senior with Security Clearance
Booz Allen Hamilton
Rome, NY🇺🇸$99k - $225k/yrOn-site6 weeks agoAWSAgileBash+7Technology - JT
Fortinet Network Engineer
NewJBS Technologies
United States🇺🇸Hybrid16 hours agoAzureDNSTerraformTechnology - BR
Senior AI/ML Cloud Engineer - Agentic AI
NewBrooksource
MO🇺🇸$65 - $80/hrHybrid16 hours agoAWSGoogle CloudJava+3Technology - PI
Sr Azure Cloud Engineer
NewPegasys Information Technologies
Sunrise, FL🇺🇸On-site16 hours agoAzurePythonTechnology - VD
Cloud Infrastructure Architect - Identity Management ; SSL
VDart, Inc.
United States🇺🇸Hybrid1 week agoAWSOAuthSAML+7Technology