Haystack
← Back to Jobs
Technology

Senior Cloud Engineer - Observability Platform

Digipulse Technologies, IncDallas, TX🇺🇸United StatesPosted 20 Jul 2026

Quick Overview

Work Type
Hybrid
Level
Mid Senior

Job Description

About Digipulse  

Digipulse Technologies, Inc (DTI) is a software solutions company offering focused IT services to fortune 1000 clients. We provide services in enterprise application development, system integration & support services for Insurance, Financial, Hospitality, Telecommunication, Pharmaceutical & Banking sectors.

 

About the Role

 The Expertise and Skills You Bring

  • ·       5+ years of hands-on experience deploying and/or supporting highly distributed multi-tiered systems at scale.
  • ·       Deep understanding of distributed systems and telemetry architecture.
  • ·       Experience choosing/modeling the right technique for the job (e.g., anomaly detection, ranking/recommendation, NLP), and knowing when a heuristic beats a model
  • ·       Experience operating observability pipelines in Kubernetes or similar orchestration environments.
  • ·       Hands-on experience in 2 or more languages (Python, Java, Go etc.).
  • ·       Hands-on experience with Open Telemetry (OTEL) or any opensource Observability implementations.
  • ·       5+ years of experience leading projects and designing, analyzing, and troubleshooting distributed systems
  • ·       Create and maintain Grafana dashboards, visualizations, and alerts for real-time operational insights
  • ·       Hands-on experience designing and building scalable and resilient applications in the cloud.
  • ·       Good knowledge on networking concepts such as DNS, Load Balancers, routers, Linux etc.
  • ·       Experience on containerization technologies such as Docker, Kubernetes etc.
  • ·       Experience in deploying applications in AWS platforms such as EC2 & EKS or equivalent platforms from any other cloud provider such as Google Cloud Platform, Azure etc
  • ·       Experience delivering software using engineering best practices and principles.
  • ·       Extensive knowledge of infrastructure as code (Terraform, CFT, CDK, etc.).
  • ·       Hands-on experience with continuous integration and continuous delivery/deployment using ALM tools such as Jenkins.

 

Bonus skills

  • ·       Proven experience delivering LLM/agent features to production (prompting, tooling, evals, safety/guardrails)
  • ·       Experience working on Logging, Metrics, and Tracing frameworks and understanding of Logging and Metrics Data Models.
  • ·       You have proven ability to use AI coding tools in day-to-day workflows and validate, critique, and refine AI-generated output
  • ·       Experience operating and/or using modern observability tools at scale (Prometheus, Grafana, ELK, Jaeger, Open Telemetry, fluentD, fluentBit, Open Tracing, Datadog, Splunk, etc…)

If you are intrested in this role, please apply to this job or drop a mail with profile @

Skills

Docker
AWS
ELK
NLP
Splunk
Azure
CDK
DNS
Datadog
Google Cloud
Grafana
Java
Jenkins
Kubernetes
LLM
Prometheus
Python
Terraform

Similar jobs