Senior Site Reliability Engineer
New
T
TailorCareHealthTech
United States, all US time zones (ET, CT, MT, and PT)Full-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years
- Required Skills
- AWSPythonTypeScriptGoGrafanaCI/CDTerraformDatadog
Requirements
- 5+ years in Software Engineering, SRE, DevOps, or Platform Engineering.
- Deep hands-on AWS expertise including VPC, IAM, ECS/EKS, Lambda, RDS, and S3.
- Production-grade Terraform experience at scale including modules, state management, and multi-environment setups.
- Strong programming skills in Python, Go, TypeScript, or similar languages.
- Hands-on experience with modern observability stacks such as Datadog, CloudWatch, or Grafana.
- Practical understanding of implementing SLOs, SLIs, and error budgets.
- Experience maintaining and standardizing CI/CD pipelines and tracking DORA metrics.
- Ability and willingness to travel up to 10% for onsite meetings and company events.
- Experience operating in a HIPAA, HITRUST, or SOC 2 Type II regulated environment is highly desired.
Responsibilities
- Implement the standardization of our AWS footprint using Terraform.
- Build and maintain universal CI/CD pipelines that remove friction from the SDLC.
- Implement observability improvements across AWS and third-party integrations.
- Monitor and support SLOs, SLIs, and error budgets for key services.
- Act as a bridge between SRE and software/data engineering.
- Stand up the on-call rotation and contribute to shaping core on-call hours.
- Lead production incidents and drive blameless post-incident reviews.
- Implement and maintain infrastructure controls to align with HIPAA and HITRUST requirements.
View Full Description & ApplyYou'll be redirected to the employer's site