Senior Site Reliability Engineer

New
T
TailorCareHealthTech
United States, all US time zones (ET, CT, MT, and PT)Full-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page

Job Details

Experience
5+ years
Required Skills
AWSPythonTypeScriptGoGrafanaCI/CDTerraformDatadog

Requirements

  • 5+ years in Software Engineering, SRE, DevOps, or Platform Engineering.
  • Deep hands-on AWS expertise including VPC, IAM, ECS/EKS, Lambda, RDS, and S3.
  • Production-grade Terraform experience at scale including modules, state management, and multi-environment setups.
  • Strong programming skills in Python, Go, TypeScript, or similar languages.
  • Hands-on experience with modern observability stacks such as Datadog, CloudWatch, or Grafana.
  • Practical understanding of implementing SLOs, SLIs, and error budgets.
  • Experience maintaining and standardizing CI/CD pipelines and tracking DORA metrics.
  • Ability and willingness to travel up to 10% for onsite meetings and company events.
  • Experience operating in a HIPAA, HITRUST, or SOC 2 Type II regulated environment is highly desired.

Responsibilities

  • Implement the standardization of our AWS footprint using Terraform.
  • Build and maintain universal CI/CD pipelines that remove friction from the SDLC.
  • Implement observability improvements across AWS and third-party integrations.
  • Monitor and support SLOs, SLIs, and error budgets for key services.
  • Act as a bridge between SRE and software/data engineering.
  • Stand up the on-call rotation and contribute to shaping core on-call hours.
  • Lead production incidents and drive blameless post-incident reviews.
  • Implement and maintain infrastructure controls to align with HIPAA and HITRUST requirements.
View Full Description & ApplyYou'll be redirected to the employer's site
View details
Apply Now