Site Reliability Engineer, Robotics at Hadrian Automation
- Company: Hadrian Automation
- Location: Los Angeles, CA
- Job type: full time
- Workplace: onsite
- Posted: 2026-05-27
Job description
Hadrian - Manufacturing the Future Hadrian is building autonomous factories to reindustrialize America. By combining AI, advanced software, robotics, and full-stack manufacturing, we help aerospace and defense companies build rockets, satellites, aircraft, ships, and other mission-critical systems up to 10x faster and at significantly lower cost. Following our $1.37B Series D at a $7.87B valuation, Hadrian is rapidly expanding our manufacturing footprint, launching new capabilities across welding, casting, forging, electronics, additive manufacturing, and more, while scaling our Factory-as-a-Service platform to transform how critical products are built. Backed by leading investors including JPMorgan Chase, Valor Equity Partners, Andreessen Horowitz, Founders Fund, 137 Ventures, Lux Capital, T. Rowe Price, and Morgan Stanley, we’re building the future of American manufacturing—and looking for exceptional people to help make it happen. If you’re ready to take on the most challenging and rewarding work of your career while helping create American manufacturing jobs for generations to come, you’re exactly who we’re looking for. The Role What You’ll Do • Own the reliability of our robotics systems, from PLCs through ROS2/middleware to Kubernetes. • Build interfaces to our observability system to ingest telemetry from our controls and robotics systems. Leverage solutions such as Prometheus, Telegraf, OpenTelemetry, and Datadog. • Write code frameworks and tools to support our controls and robotics systems, including diagnostic tools, shared libraries for telemetry data, and automated remediation. • Partner with controls, robotics, and platform engineering teams to bake reliability in early. Review designs, develop SLOs and SLIs, introduce reliability release gates, and push for telemetry contracts to develop production-grade services. What We’re Looking For • Ownership. Someone who has owned the reliability of a production system where downtime had physical or operational consequences (manufacturing line, autonomous vehicle, lab automation, network operations) • Systems Thinker. Focused on understanding the relationship among various systems to design sustainable solutions, not one-time fixes. • Problem Solver. Solving complex puzzles excites and motivates you to find an efficient solution. • T-Shaped Skill Set. Comfortable with bare metal Kubernetes, networking, GitOps workflows, and Infrastructure as Code (IaC). Also skilled in programming in TypeScript, Python, Golang, or C++. • Strong Communication. You can run a war room, write a post-mortem, and explain a reliability tradeoff to a stakeholder. What Will Set You Apart • You've built automated or self-healing remediation at scale. We want systems that remove humans from the loop. • Background in edge/on-prem infrastructure . You’ve run Kubernetes at the edge (k3s, k0s, k0smotron), managing on-prem clusters, time-series at the edge, or air-gapped deployments. A deep understanding of Linux operating system fundamentals such as cgroups, sockets, and system tuning, is a big plus. • Deep understanding of shipping and storing telemetry data at scale. Experience with Kafka/MQTT/RabbitMQ is a plus. • Direct robotics experience . ROS/ROS2, OPC UA, EtherCAT, motion controllers, or fleet management for autonomous systems • An individual who is self-directed and can deliver with high velocity. Compensation For this role, the target salary range is $164,000 - $270,000 (actual range may vary based on experience). This is the lowest to highest salary we reasonably and in good faith believe we would pay for this role at the time of this posting. We may ultimately pay more or less than the posted range, and the range may be modified in the future. An employee's pay position within the salary range will be based on several factors, including, but not limited to, relevant education, qualifications, certifications, experience, skills, geographic location, performance, and business or organizational
Site Reliability Engineer, Robotics on JobPost.