Senior Site Reliability Engineer

New
M
Multi Media, LLCLive Streaming
United StatesFull-TimeSenior
Salary$169,000 - $215,000 USD
Apply NowOpens the employer's application page

Job Details

Required Skills
DockerPythonBashKubernetesGoLinuxTerraformAnsibleNetworking

Requirements

  • STEM degree and/or relevant experience as a Site Reliability Engineer, Devops Engineer, or SWE
  • Proficiency in Python or Golang; experience with other compiled or high-level languages like C, C#, C++, Java, or Rust is accepted
  • Experience running Web applications at scale
  • Experience with Web application concepts and frameworks (e.g., ORM, MVC architecture, Django, Flask, Laravel)
  • Proficiency with Linux administration, Bash shell, and strong knowledge of Linux internals
  • Strong networking knowledge (e.g., routing, switching, TCP stack) for metal and cloud environments
  • Experience in database administration and configuration
  • Experience with DevOps tools such as Terraform, Ansible, Docker, Kubernetes, ArgoCD, or Helm
  • Willingness to participate in on-call rotation and respond to monitoring and alerting

Responsibilities

  • Analyze system performance using APM and distributed telemetry data to identify sources of instability
  • Improve scalability, reliability, and performance through software enhancements and patching
  • Develop tools and automation to streamline the DevOps pipeline
  • Design and manage infrastructure in both data center metal environments and in the public cloud
  • Conduct predictive failure analysis and disaster planning
  • Administer and configure databases and key-value stores with a focus on uptime and performance
  • Analyze complex systems to identify operational surprises and minimize downtime
  • Participate in incident response and produce postmortem reports
  • Collaborate with other engineering teams
View Full Description & ApplyYou'll be redirected to the employer's site
$169,000 - $215,000 USD
Apply Now