Senior Site Reliability Engineer
New
B
Branch AppFinTech
REMOTE within the United States of AmericaFull-TimeSenior
Salary$175-185k
Apply NowOpens the employer's application page
Job Details
- Experience
- 3+ years
- Required Skills
- DockerJavaKubernetesSpring BootGoTerraform
Requirements
- Bachelor's degree in an appropriate engineering discipline or equivalent experience
- 3+ years experience in site reliability engineering
- Strong hands-on experience building and operating Java / Spring Boot services in production
- Experience with Terraform
- Experience with Go
- Experience with Java
- Experience with Gradle
- Experience with Docker
- Experience with OpenTelemetry
- Experience with Kubernetes
Responsibilities
- Partner with Developers to produce high-performing and robust services through rigorous testing and release procedures
- Design infrastructure, monitoring, processes, and standards for systems and applications
- Support services through design, development, load testing, and launch phases
- Develop, measure, and monitor key performance and service level indicators including availability, latency, and overall system health
- Define and establish SLIs, SLOs, and error budgets with service owners, and drive adoption across platform teams
- Profile and optimize platform performance, resilience, and efficiency, including latency, throughput, and capacity planning under load
- Participate in incident response and root cause analysis
- Remediate tasks and develop preventative and automated measures to meet SLAs/SLOs/SLIs
- Manage monitoring services utilized by applications
View Full Description & ApplyYou'll be redirected to the employer's site