Site Reliability Engineer - AWS

New
J
JobgetherCloud Computing
Based in the United StatesFull-TimeMiddle
Salary$110,000–$140,000
Apply NowOpens the employer's application page

Job Details

Experience
5+ years
Required Skills
AWSDockerPythonSQLKubernetesCI/CDTerraform

Requirements

  • Bachelor's degree or equivalent combination of education and professional experience.
  • 5+ years of relevant industry or technical experience in DevOps, cloud infrastructure, or SRE.
  • Extensive experience with AWS cloud infrastructure and production environments.
  • Strong proficiency with Kubernetes, Docker, and Amazon EKS.
  • Hands-on experience building CI/CD pipelines using GitOps methodologies (ArgoCD or FluxCD).
  • Strong knowledge of Azure DevOps for CI/CD and infrastructure-as-code workflows.
  • Proficiency in Terraform and infrastructure-as-code best practices.
  • Advanced scripting and automation skills using Python, Bash, or PowerShell.
  • Experience with AWS CloudWatch for monitoring and observability.
  • Strong SQL and database administration skills, particularly with RDS.
  • Solid understanding of HTTP concepts, APIs, microservices, and network management.

Responsibilities

  • Support and continuously improve production applications running in AWS, ensuring high availability, reliability, and performance.
  • Monitor system health, availability, latency, metrics, and logs to proactively address reliability risks.
  • Participate in incident response, root cause analysis, and post-incident improvement activities.
  • Engineer and optimize AWS environments using services like EKS, RDS, and Lambda.
  • Automate recurring operational processes using Python, Bash, and Terraform.
  • Manage containerized environments using Kubernetes, Docker, and Amazon EKS with GitOps approaches.
  • Partner with development teams to identify performance bottlenecks and reliability risks.
View Full Description & ApplyYou'll be redirected to the employer's site
$110,000–$140,000
Apply Now