Site Reliability Engineer - AWS
New
J
JobgetherCloud Computing
Based in the United StatesFull-TimeMiddle
Salary$110,000–$140,000
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years
- Required Skills
- AWSDockerPythonSQLKubernetesCI/CDTerraform
Requirements
- Bachelor's degree or equivalent combination of education and professional experience.
- 5+ years of relevant industry or technical experience in DevOps, cloud infrastructure, or SRE.
- Extensive experience with AWS cloud infrastructure and production environments.
- Strong proficiency with Kubernetes, Docker, and Amazon EKS.
- Hands-on experience building CI/CD pipelines using GitOps methodologies (ArgoCD or FluxCD).
- Strong knowledge of Azure DevOps for CI/CD and infrastructure-as-code workflows.
- Proficiency in Terraform and infrastructure-as-code best practices.
- Advanced scripting and automation skills using Python, Bash, or PowerShell.
- Experience with AWS CloudWatch for monitoring and observability.
- Strong SQL and database administration skills, particularly with RDS.
- Solid understanding of HTTP concepts, APIs, microservices, and network management.
Responsibilities
- Support and continuously improve production applications running in AWS, ensuring high availability, reliability, and performance.
- Monitor system health, availability, latency, metrics, and logs to proactively address reliability risks.
- Participate in incident response, root cause analysis, and post-incident improvement activities.
- Engineer and optimize AWS environments using services like EKS, RDS, and Lambda.
- Automate recurring operational processes using Python, Bash, and Terraform.
- Manage containerized environments using Kubernetes, Docker, and Amazon EKS with GitOps approaches.
- Partner with development teams to identify performance bottlenecks and reliability risks.
View Full Description & ApplyYou'll be redirected to the employer's site