Sr Site Reliability Engineer
New
B
BeyondTrustCybersecurity SaaS
Remote United States | Remote CanadaFull-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Required Skills
- AWSKubernetesTerraformGitLabDatadog
Requirements
- Work with AWS cloud resources, including S3, EC2, and RDS.
- Work with Kubernetes clusters in EKS.
- Work with the Istio service mesh.
- Use infrastructure-as-code tools such as Terraform or AWS CDK.
- Work with continuous build tools such as GitLab.
- Work with continuous delivery tools such as ArgoCD.
- Use Datadog.
- Participate in an on-call rotation for platform availability incidents.
Responsibilities
- Design long-term technical solutions and cross-team mechanisms to achieve reliability goals.
- Define a roadmap for engineering teams to use automated, self-service, scalable, efficient, observable, and reliable infrastructure services.
- Align and help drive execution of the Platform Infrastructure team’s roadmap.
- Collaborate with SREs and senior engineers across engineering organizations on best practices.
- Provide technical guidance and feedback during engineering design reviews for teams onboarding to Platform Infrastructure.
- Automate work to reduce toil.
- Build monitoring and alerting for Platform Infrastructure.
- Participate in an on-call rotation to respond to incidents that impact platform availability.
View Full Description & ApplyYou'll be redirected to the employer's site