Senior Site Reliability Engineer (SRE)
New
L
LeoLabsSpace/Aerospace
Location: RemoteFull-TimeSenior
SalaryThe estimated base salary is $192,000, with additional compensation opportunities to include bonus and equity.
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years
- Required Skills
- AWSDockerPostgreSQLPythonKubernetesAzureGoCI/CDTerraformDatadog
Requirements
- Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent work experience.
- 5+ years of experience in a Site Reliability Engineering, DevOps, or related role.
- Proficiency in scripting or programming language (e.g., Python, Go).
- Experience with cloud services (AWS, Azure).
- Proficiency with containerization (Docker, Kubernetes, ECS).
- Proficiency in configuration management tools (Terraform, Atlantis, Terragrunt).
- Familiarity with CI/CD tools (GitHub Actions, AWS CodeBuild, CircleCI).
- Experience with monitoring tools (Grafana, Datadog).
- Familiarity with database technologies (RDS, Aurora, PostgreSQL).
- Experience with large-scale distributed systems and microservices architecture.
- Ability to obtain a U.S. personnel security clearance.
Responsibilities
- Design, implement, and maintain scalable and reliable systems.
- Set up monitoring tools and create incident response plans to identify and resolve issues.
- Develop and maintain scripts and automation tools for deployment, monitoring, and system health checks.
- Analyze system capacity and performance metrics to forecast future needs and implement scaling solutions.
- Work closely with development teams to enhance product reliability and streamline the deployment process.
- Create and maintain documentation for system architecture, processes, and incident reports.
- Participate in on-call rotations to provide 24/7 support for critical systems.
- Implement and enforce security best practices across all systems.
View Full Description & ApplyYou'll be redirected to the employer's site