Why This Role Stands Out
This Embedded Platform Engineer role at Balin Technologies offers significant opportunities for professional growth and skill development within a reputable tech company. You'll thrive here if you are a proactive, collaborative individual eager to contribute to robust system reliability and enjoy a dynamic, customer-focused environment. Apply now to join their innovative team!
Quick Overview
Seniority
Mid Senior
Work mode
Hybrid
Location
San Jose, CA, United States
Posted
2 weeks ago
TCP/IPAnsibleGitLab CIKubernetesPythonTerraform
Job Description
Role 1 — Core Platform Engineer
The engineering first line of defense: incident response, triage, reliability, and automation across the full infrastructure stack.
Day-to-day:
- Runs incident response drills, post-mortems, and root cause analysis; learns from past incidents to prevent recurrence.
- Starts the day reviewing overnight alerts and system performance metrics, triaging anomalies.
- Participates in team stand-ups on projects, incidents, and daily priorities.
- Automates routine processes, analyzes system logs, and builds tools to strengthen monitoring.
- Works alongside software engineers advising on resilient-code best practices and reviewing changes pre-deployment.
- Maintains high SLIs/SLOs; documents work and shares insights with a customer-centric mindset.
Must have:
- Architecture, design patterns, reliability, and scaling of new and existing systems.
- Incident command experience — driving RCA, coordinating cross-functional teams, ensuring corrective-action follow-through.
- Observability built from the ground up — defining SLOs/SLIs, closing monitoring gaps, alerting strategies that catch failures before customers do.
- Linux kernel internals — scheduler, memory allocation, driver subsystems.
- High-quality code in at least one language (Python, Go, or similar).
- System-level debugging — kdump, kernel panic analysis.
- IaC (Ansible, Terraform, Kubernetes) and CI/CD (GitLab CI, AWX, etc.) for bare-metal or cloud infrastructure.
- TCP/IP and network programming.
- Distributed storage systems — object, block, and/or file storage paradigms.
- Strong communication skills.
Similar jobs
- MA
Senior Site Reliability Engineer
NewMastercard
O Fallon, MO🇺🇸$96k - $163k/yrOn-site12 hours agoRubyChefC+++5Technology - OK
Senior Manager, Site Reliability Engineering - Infrastructure Platform
NewOkta
Bellevue, Washington🇺🇸$232k - $319k/yrOn-site2 hours agoAWSMachine LearningNginx+5Technology - OK
Senior Manager, Site Reliability Engineering - Infrastructure Platform
NewOkta
San Francisco, California🇺🇸$232k - $319k/yrOn-site2 hours agoAWSMachine LearningNginx+5Technology - OK
Senior Manager, Site Reliability Engineering - Infrastructure Platform
NewOkta
Chicago, Illinois🇺🇸$232k - $319k/yrOn-site2 hours agoAWSMachine LearningNginx+5Technology - OK
Senior Manager, Site Reliability Engineering - Infrastructure Platform
NewOkta
Washington, Washington DC🇺🇸$232k - $319k/yrOn-site2 hours agoAWSMachine LearningNginx+5Technology - NO
DevOps Engineer
NewNetwork Objects Inc.
Union City, NJ🇺🇸Hybrid12 hours agoDockerScrumAgile+5Technology