Software Development Engineer II, Cloud Platform
New
J
JobgetherCloud Infrastructure
USFull-TimeMiddle
Salary157,675 - 213,325 USD per year
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years
- Required Skills
- AWSNode.jsPythonKubernetesGoCI/CDTerraform
Requirements
- 5+ years of experience managing AWS infrastructure using infrastructure-as-code tools such as Terraform, Terragrunt, Atlantis, CDK, or similar technologies.
- 4+ years of experience managing containerized workloads at scale using Kubernetes, EKS, ECS, or related platforms.
- Strong experience building and maintaining CI/CD systems in distributed engineering organizations using GitHub Actions or comparable tools.
- Advanced knowledge of Kubernetes, ArgoCD, Istio, and cloud-native deployment practices.
- Proven ability to design secure, reliable, and cost-efficient AWS solutions using services such as EKS, ECS, EC2, Lambda, Fargate, CloudFront, IAM, Route53, and DynamoDB.
- Proficiency in at least one programming language such as Python, Node.js, or Go.
- Experience implementing observability solutions using platforms such as Datadog, CloudWatch, or similar monitoring tools.
- Knowledge of incident response practices, resilience engineering, and blameless post-mortem processes.
- Strong documentation, communication, and knowledge-sharing skills with a willingness to mentor others.
Responsibilities
- Build, maintain, and improve cloud-native infrastructure and deployment platforms following GitOps principles.
- Manage and expand AWS resources through infrastructure-as-code frameworks such as Terraform and Terragrunt.
- Design and promote Kubernetes-based deployments for new services, supporting scalable and reliable application delivery.
- Lead migration initiatives from legacy deployment approaches toward modern platforms such as EKS and ArgoCD.
- Develop and maintain centralized CI/CD frameworks using tools such as GitHub Actions and automated deployment solutions.
- Configure and improve observability platforms to support monitoring, alerting, analytics, and operational visibility.
- Promote reliability engineering practices through testing, incident response, monitoring, and continuous improvement.
View Full Description & ApplyYou'll be redirected to the employer's site