Senior Site Reliability Engineer
New
P
PrizePicksDaily Fantasy Sports
While we prefer candidates based in Atlanta, we are open to qualified applicants from anywhere in the U.S. and are willing to consider remote candidates.Full-TimeSenior
Salary$120,000 to $175,000
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years
- Required Skills
- AWSPythonGCPKubernetesRubyAzureGoGrafanaTerraform
Requirements
- 5+ years of experience as a reliability-focused engineer in a fast-paced, rapidly growing, enterprise environment.
- Deep understanding of cloud computing such as AWS, Azure, and/or GCP.
- Proficiency with Infrastructure as Code tools such as Terraform or Crossplane.
- Experience developing applications in languages such as Python, Ruby, or Go.
- Experience deploying and supporting applications in Kubernetes at scale.
- Proficiency implementing monitoring in tools like Grafana, New Relic, or Datadog.
- Experience debugging live, critical production issues.
- Familiarity with reliability principles, such as resilient systems, application and supply chain security, and SLO governance.
- Ability to work cross-functionally with diverse engineering teams.
Responsibilities
- Design, implement, maintain, and monitor reliable production systems at scale.
- Lead incident response, mitigate production issues, and conduct post mortem analysis.
- Proactively monitor performance, analyze system failures, identify bottlenecks, and propose solutions.
- Create and support observability/monitoring tools and vendor integrations.
- Drive the growth of a reliability culture, promoting cross-functional collaboration towards improving system reliability, scalability, resilience, and security.
- Train and mentor other engineers.
View Full Description & ApplyYou'll be redirected to the employer's site