Sr Site Reliability Engineer

New
B
BeyondTrustCybersecurity SaaS
Remote United States | Remote CanadaFull-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page

Job Details

Required Skills
AWSKubernetesTerraformGitLabDatadog

Requirements

  • Work with AWS cloud resources, including S3, EC2, and RDS.
  • Work with Kubernetes clusters in EKS.
  • Work with the Istio service mesh.
  • Use infrastructure-as-code tools such as Terraform or AWS CDK.
  • Work with continuous build tools such as GitLab.
  • Work with continuous delivery tools such as ArgoCD.
  • Use Datadog.
  • Participate in an on-call rotation for platform availability incidents.

Responsibilities

  • Design long-term technical solutions and cross-team mechanisms to achieve reliability goals.
  • Define a roadmap for engineering teams to use automated, self-service, scalable, efficient, observable, and reliable infrastructure services.
  • Align and help drive execution of the Platform Infrastructure team’s roadmap.
  • Collaborate with SREs and senior engineers across engineering organizations on best practices.
  • Provide technical guidance and feedback during engineering design reviews for teams onboarding to Platform Infrastructure.
  • Automate work to reduce toil.
  • Build monitoring and alerting for Platform Infrastructure.
  • Participate in an on-call rotation to respond to incidents that impact platform availability.
View Full Description & ApplyYou'll be redirected to the employer's site
View details
Apply Now