Sr. Site Reliability Engineer
New
P
PayNearMeFinTech Payments
RemoteFull-TimeSenior
Salary$180,000 - $200,000 USD
Apply NowOpens the employer's application page
Job Details
- Experience
- +3 years
- Required Skills
- AWSDockerPythonBashKubernetesGoTerraformAnsibleDatadog
Requirements
- 3+ years of experience in SRE, DevOps, or a related role.
- Proficiency with cloud platforms such as AWS, GCP, or Azure (EC2, RDS, VPCs).
- Strong experience with Kubernetes and Docker.
- Expert in Infrastructure as Code using Terraform.
- Proficient with configuration management tools such as Ansible, Puppet, or Chef.
- Extensive experience with monitoring and observability tools like Datadog, Prometheus, Grafana, ELK stack, or Splunk.
- Proven ability to define, monitor, and maintain SLOs and SLAs.
- Strong scripting skills in Python, Bash, or Go.
- Experience with GitLab CI or similar CI/CD tools.
- Experience supporting production environments running Go or Ruby/Rails applications.
- Strong organizational, documentation, and collaborative problem-solving skills.
Responsibilities
- Design, implement, and maintain scalable and resilient infrastructure using Terraform.
- Deploy, manage, and optimize Kubernetes clusters and containerized applications.
- Develop and maintain monitoring and observability solutions using Datadog.
- Define, monitor, and maintain Service Level Objectives (SLOs) and SLAs.
- Respond to incidents, perform root cause analysis, and participate in post-incident reviews.
- Develop automation scripts using Python, Bash, or Go and maintain CI/CD pipelines.
- Implement security best practices and ensure compliance.
- Participate in an on-call rotation to address production issues.
View Full Description & ApplyYou'll be redirected to the employer's site