Senior Site Reliability Engineer (SRE)
Job Description
Key Skills Required
Master these to land this role
Want to know if you're a match for this job?
LeoLabs is seeking a skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate will bridge the gap between development and operations, ensuring that our systems are scalable, reliable, and efficient. You will be responsible for automating processes, monitoring system performance, and resolving incidents to enhance our service reliability.
Key Responsibilities:
- System Reliability: Design, implement, and maintain scalable and reliable systems.
- Monitoring and Incident Response: Set up monitoring tools and create incident response plans to quickly identify and resolve issues, as well as implementing preventative measures.
- Automation: Develop and maintain scripts and automation tools for deployment, monitoring, and system health checks.
- Capacity Planning: Analyze system capacity and performance metrics to forecast future needs and implement scaling solutions.
- Collaboration: Work closely with development teams to enhance product reliability and streamline the deployment process.
- Documentation: Create and maintain documentation for system architecture, processes, and incident reports.
- On-Call Support: Participate in on-call rotations to provide 24/7 support for critical systems.
Security: Implement and enforce security best practices across all systems, ensuring compliance with industry standards.
Qualifications:
- Education: Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent work experience.
- Experience: 5+ years of experience in a Site Reliability Engineering, DevOps, or related role.
- Technical Skills:
- Proficiency in scripting or programming language (e.g., Python, Go)
- Experience with cloud services (AWS, Azure)
- Proficiency with containerization (Docker, Kubernetes, ECS)
- Proficiency in configuration management tools (Terraform, Atlantis, Terragrunt)
- Familiarity with CI/CD tools (GitHub Actions, AWS CodeBuild, CircleCI)
- Experience with monitoring tools (Grafana, Datadog)
- Familiarity with database technologies (RDS, Aurora, PostgreSQL)
- Experience with large-scale distributed systems and microservices architecture.
- Problem-Solving: Strong analytical and problem-solving skills with the ability to troubleshoot complex systems.
- Communication: Excellent verbal and written communication skills, with the ability to collaborate effectively across teams.
- Ability to obtain a U.S. personnel security clearance.
Preferred qualifications
- Active TS/SCI clearance
What Success Looks Like:
Within 1 month, you’ll:
- Complete onboarding to understand our business, vision, and team structure.
- Get familiar with LeoLabs' engineering stack, security posture, and key initiatives.
- Gain an understanding about how your role fits into LeoLabs broader organization.
Within 3 months, you’ll:
- Independently deploy infrastructure changes using Infrastructure as Code.
- Identify key reliability risks and recommend improvements.
- Improve dashboards, alerts, and operational runbooks.
Within 6 months, you’ll:
- Improve deployment pipelines, infrastructure provisioning, or self-service capabilities.
- Optimize infrastructure utilization and cloud costs without compromising reliability.
- Drive automation that reduces operational toil and improves deployment reliability.
Within 12 months, you’ll:
- Lead cross-functional initiatives to improve availability, scalability, and operational efficiency.
- Be a key advisor for site reliability in new product developments and platform evolution.
- Mentor junior engineers and foster the development culture.
How would you rate this job post?
See what other professionals think about this role.
Similar Opportunities
More Openings at LeoLabs
Explore Top Companies in this Space
Clickhouse
Database Technology
Vannevarlabs
Artificial Intelligence, Data Science, Scientific Software
Cribl
Software / Data Observability
Foodics
FoodTech & Restaurant Management / POS & F&B Operations SaaS / Inventory Management & Analytics / B2B Enterprise Software
LeoLabs
View Company ProfileLeoLabs is a pioneering SpaceTech enterprise that is fundamentally mapping the new frontier of Low Earth Orbit (LEO). Founded in 2016 and headquartered in Menlo Park, California, the company provides critical Space Situational Awareness (SSA) and Space Domain Awareness (SDA) to a rapidly commercializing space industry. Under the hood, LeoLabs operates a massive, proprietary global network of state-of-the-art phased-array radars capable of continuously tracking thousands of satellites and dangerous orbital space debris in real-time. Their primary target audience spans commercial satellite operators, national defense agencies, and aerospace regulators who desperately need precise, high-frequency tracking data to execute collision avoidance maneuvers and protect billion-dollar orbital assets. What sets LeoLabs apart in the deep tech and aerospace sector is its ability to commercialize and scale orbital mapping—transforming space tracking from a historically slow, government-monopolized capability into an agile, highly accessible, cloud-based data platform that effectively serves as the real-time air traffic control for the stars.
Safety First
- Never pay for a job application.
- Do not share sensitive bank info.
- Verify the client before starting work.

