Why This Role Stands Out
This role offers a unique opportunity to shape the infrastructure for groundbreaking AI robotics, with significant potential for skill development in cutting-edge cloud and MLOps technologies. You'll thrive here if you're a seasoned DevOps engineer passionate about building robust systems and driving innovation in a rapidly growing, well-funded company. We encourage you to apply and contribute to the future of robotics!
Quick Overview
Job Description
At Rhoda AI, we’re building the next generation of generalist intelligent robots. We own the full robotics stack from high-performance hardware and robot systems to the infrastructure and state-of-the-art foundation world models that control our robots. Our robots are designed to be generalists capable of operating in complex, real-world environments and handling long-tail edge cases, made possible by our cutting edge research and end-to-end system design. We've raised over $450M and are investing aggressively in model research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality.
Responsibility:
Maintain and improve our CI/CD pipeline: instrument and monitor our CI jobs, and iteratively improve on it.
Manage and maintain infrastructure around build artifacts and dependencies storage.
Build and improve on our cloud and GPU cluster, and create essential devops infrastructure to support our daily research and software development effort.
Improve our MLOps training and deployment path: checkpoint registry/DB, model artifact promotion, and rollout to inference/robot endpoints.
Improve our infrastructure observability.
Harden security & multi-tenancy
Qualifications
5+ years in DevOps / SRE / platform / infra, with ownership of production systems.
Proficient in a programming language such as Python, C++, Rust, Go
Strong knowledge of software engineering best practices and design patterns
Experience with docker and containerized environments
Experience with software build systems for cloud infrastructure and embedded systems.
Experience with procedural CI and CD, and build artifact delivery (OTA update) system.
Comfortable with operating in Linux environment
Self-starter mentality — comfortable with ambiguity, able to prioritize independently, and willing to jump in wherever needed
Effective communication skills; able to work cross-functionally in a fast-moving team
Preferred Qualifications
Knowledge of industrial communication protocols (EtherCAT, Modbus, gRPC, etc.)
Experience with ML Ops
Understanding of networking fundamentals (TCP/IP, DNS, firewalls) and security best practices for embedded IoT devices
Knowledge with kubernetes and cloud system orchestration
Similar jobs
- MA
Site Reliability Engineer I
NewMastercard
O Fallon, Missouri🇺🇸Hybrid1 hour agoDockerGCPAWS+9Technology - SO
Senior Infrastructure Automation Engineer
NewSystem One
United States🇺🇸$80/hrHybridYesterdayAnsibleBashDatadog+5Technology - BA
AI DevOps Engineer
NewBooz Allen Hamilton
Ashburn, VA🇺🇸$77.6k - $176k/yrOn-siteYesterdayAgileKubernetesTechnology - DS
Senior Cloud Platform Engineer VMware Cloud Foundation
NewDella Solutions and Services Inc.
Irving, TX🇺🇸$58/hrOn-siteYesterdayAWSTCP/IPAnsible+9Technology - AT
AWS Devops Engineer
NewAcadia Technologies, Inc.
Dallas, TX🇺🇸HybridYesterdayAWSTechnology - SS
Site Reliability Engineer (SRE) Security Infrastructure
Simple Solutions
Cary, NC🇺🇸On-site4 days agoPythonTechnology