Site Reliability Engineer II
New
M
MrsoolDelivery Platform
IndiaFull-TimeMiddle
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years
- Required Skills
- AWSDockerPythonGCPJavaKubernetesRubyAzureGoTerraform
Requirements
- Bachelor’s degree in Computer Engineering, Computer Science, or related field.
- 5+ years of experience in a similar role, preferably in a high-traffic, high-availability environment.
- Proficiency in at least one programming language (e.g., Python, Ruby, Java, Go).
- Strong understanding of cloud infrastructure (AWS, GCP, Azure).
- Experience with containerization technologies such as Kubernetes and Docker.
- Experience with one or more automation and configuration management tools (e.g., Chef, Ansible, Puppet, Terraform).
- Familiarity with monitoring and alerting tools (e.g., Prometheus, Grafana, Nagios).
- Strong communication and interpersonal skills for cross-functional collaboration.
- Ability to navigate ambiguity and thrive in a fast-paced environment.
- Solid grasp of computer science fundamentals, distributed systems, and networks.
Responsibilities
- Collaborate with development teams to design and implement scalable infrastructure.
- Design and implement automated deployment and testing pipelines.
- Develop and maintain monitoring and alerting systems to proactively identify and address issues.
- Troubleshoot and escalate production incidents to minimize downtime.
- Continuously improve infrastructure and processes to optimize scalability and efficiency.
- Participate and take ownership for on-call rotations to ensure 24/7 support.
- Perform routine maintenance and upgrades to keep systems updated.
- Contribute to security posture improvements and industry compliance efforts.
- Mentor and coach junior engineers.
View Full Description & ApplyYou'll be redirected to the employer's site