Primary record

Site Reliability Engineer

Skydio Indexed employerSan Mateo, California, United States · US Remote
Source-hosted applyChecked 4h agoFull-Time
Apply at Skydio

Skydio receives this application through Ashby. Babu Careers does not claim delivery.

Workplace

remote

Employment

Full-Time

Published

Aug 13, 2026

Closes

No date supplied

The role

Skydio is the leading US drone company and the world leader in autonomous flight, the key technology for the future of drones and aerial mobility. The Skydio team combines deep expertise in artificial intelligence, best-in-class hardware and software product development, operational excellence, and customer obsession to empower a broader, more diverse audience of drone users, from utility inspectors https://www.skydio.com/solutions/energy-and-utilities to first responders https://www.skydio.com/solutions/public-safety, soldiers in battlefield scenarios https://www.skydio.com/solutions/national-security/tactical-isr, and beyond https://www.skydio.com/solutions. About the role: We are looking for a hands-on Site Reliability Engineer to build, operate, and scale the cloud infrastructure that powers our products. This role is focused on owning production infrastructure, including Kubernetes, AWS, infrastructure as code, CI/CD, observability, networking, and reliability. You don't need to be an expert in every area, but you should have strong Kubernetes and cloud fundamentals with meaningful depth in at least one infrastructure domain. Our technology helps save lives. You’ll play a critical role in keeping the infrastructure behind it reliable, scalable, and available when it matters most. How you'll make an impact: - Build, operate, and troubleshoot production Kubernetes/EKS clusters. - Perform Kubernetes upgrades, node rollouts, and cluster maintenance. - Build and manage AWS infrastructure including VPCs, networking, subnets, load balancers, IAM, EKS, databases, and storage. - Define and maintain infrastructure using Terraform. - Build and operate CI/CD and deployment infrastructure. - Troubleshoot production issues across Kubernetes, AWS, Linux, networking, and databases. - Build monitoring, alerting, and observability for critical infrastructure. - Participate in on-call rotations and respond to production incidents. - Identify and solve infrastructure scaling and r

Requirements

Department: R&D; Team: Software