Senior Site Reliability Engineer
New
M
Multi Media, LLCLive Streaming
United StatesFull-TimeSenior
Salary$169,000 - $215,000 USD
Apply NowOpens the employer's application page
Job Details
- Required Skills
- DockerPythonBashKubernetesGoLinuxTerraformAnsibleNetworking
Requirements
- STEM degree and/or relevant experience as a Site Reliability Engineer, Devops Engineer, or SWE
- Proficiency in Python or Golang; experience with other compiled or high-level languages like C, C#, C++, Java, or Rust is accepted
- Experience running Web applications at scale
- Experience with Web application concepts and frameworks (e.g., ORM, MVC architecture, Django, Flask, Laravel)
- Proficiency with Linux administration, Bash shell, and strong knowledge of Linux internals
- Strong networking knowledge (e.g., routing, switching, TCP stack) for metal and cloud environments
- Experience in database administration and configuration
- Experience with DevOps tools such as Terraform, Ansible, Docker, Kubernetes, ArgoCD, or Helm
- Willingness to participate in on-call rotation and respond to monitoring and alerting
Responsibilities
- Analyze system performance using APM and distributed telemetry data to identify sources of instability
- Improve scalability, reliability, and performance through software enhancements and patching
- Develop tools and automation to streamline the DevOps pipeline
- Design and manage infrastructure in both data center metal environments and in the public cloud
- Conduct predictive failure analysis and disaster planning
- Administer and configure databases and key-value stores with a focus on uptime and performance
- Analyze complex systems to identify operational surprises and minimize downtime
- Participate in incident response and produce postmortem reports
- Collaborate with other engineering teams
View Full Description & ApplyYou'll be redirected to the employer's site