Senior Site Reliability Engineer

New
P
PrizePicksDaily Fantasy Sports
While we prefer candidates based in Atlanta, we are open to qualified applicants from anywhere in the U.S. and are willing to consider remote candidates.Full-TimeSenior
Salary$120,000 to $175,000
Apply NowOpens the employer's application page

Job Details

Experience
5+ years
Required Skills
AWSPythonGCPKubernetesRubyAzureGoGrafanaTerraform

Requirements

  • 5+ years of experience as a reliability-focused engineer in a fast-paced, rapidly growing, enterprise environment.
  • Deep understanding of cloud computing such as AWS, Azure, and/or GCP.
  • Proficiency with Infrastructure as Code tools such as Terraform or Crossplane.
  • Experience developing applications in languages such as Python, Ruby, or Go.
  • Experience deploying and supporting applications in Kubernetes at scale.
  • Proficiency implementing monitoring in tools like Grafana, New Relic, or Datadog.
  • Experience debugging live, critical production issues.
  • Familiarity with reliability principles, such as resilient systems, application and supply chain security, and SLO governance.
  • Ability to work cross-functionally with diverse engineering teams.

Responsibilities

  • Design, implement, maintain, and monitor reliable production systems at scale.
  • Lead incident response, mitigate production issues, and conduct post mortem analysis.
  • Proactively monitor performance, analyze system failures, identify bottlenecks, and propose solutions.
  • Create and support observability/monitoring tools and vendor integrations.
  • Drive the growth of a reliability culture, promoting cross-functional collaboration towards improving system reliability, scalability, resilience, and security.
  • Train and mentor other engineers.
View Full Description & ApplyYou'll be redirected to the employer's site
$120,000 to $175,000
Apply Now