Site Reliability Engineer II

New
M
MrsoolDelivery Platform
IndiaFull-TimeMiddle
Salary not disclosed
Apply NowOpens the employer's application page

Job Details

Experience
5+ years
Required Skills
AWSDockerPythonGCPJavaKubernetesRubyAzureGoTerraform

Requirements

  • Bachelor’s degree in Computer Engineering, Computer Science, or related field.
  • 5+ years of experience in a similar role, preferably in a high-traffic, high-availability environment.
  • Proficiency in at least one programming language (e.g., Python, Ruby, Java, Go).
  • Strong understanding of cloud infrastructure (AWS, GCP, Azure).
  • Experience with containerization technologies such as Kubernetes and Docker.
  • Experience with one or more automation and configuration management tools (e.g., Chef, Ansible, Puppet, Terraform).
  • Familiarity with monitoring and alerting tools (e.g., Prometheus, Grafana, Nagios).
  • Strong communication and interpersonal skills for cross-functional collaboration.
  • Ability to navigate ambiguity and thrive in a fast-paced environment.
  • Solid grasp of computer science fundamentals, distributed systems, and networks.

Responsibilities

  • Collaborate with development teams to design and implement scalable infrastructure.
  • Design and implement automated deployment and testing pipelines.
  • Develop and maintain monitoring and alerting systems to proactively identify and address issues.
  • Troubleshoot and escalate production incidents to minimize downtime.
  • Continuously improve infrastructure and processes to optimize scalability and efficiency.
  • Participate and take ownership for on-call rotations to ensure 24/7 support.
  • Perform routine maintenance and upgrades to keep systems updated.
  • Contribute to security posture improvements and industry compliance efforts.
  • Mentor and coach junior engineers.
View Full Description & ApplyYou'll be redirected to the employer's site
View details
Apply Now