Site Reliability Engineer, Tech Lead
L
LoadsmartLogistics Technology
Anywhere in Brazil - RemoteContractLead
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Languages
- Fluent in English (both written and spoken)
- Experience
- 1-3 years leading Reliability Work; Over 5 years in Cloud Computing, SRE/DevOps
- Required Skills
- AWSDockerPythonBashKubernetesCI/CDLinuxDevOpsTerraform
Requirements
- 1-3 years leading Reliability Work across multiple engineering squads.
- Over 5 years of experience in Cloud Computing, SRE/DevOps.
- Proven experience collaborating with internal stakeholders across multiple engineering squads.
- Strong project management skills with a demonstrated ability to delegate and mentor team members.
- Fluent in English communication (both written and spoken).
- Strong understanding of software engineering principles and systems internals.
- In-depth knowledge of modern networking and operating systems.
- Proficiency in AWS, cloud environments, containers, Kubernetes, Docker, and CI/CD pipelines.
- Familiarity with automation tools and provisioners like Terraform, Ansible, or Chef.
- Solid troubleshooting and system engineering experience in UNIX/Linux production environments.
- Experience with monitoring, alerting, and incident management.
- Proficiency in automating tasks with scripting languages like Python or Bash.
Responsibilities
- Design, deploy, and operate critical systems while balancing reliability, cost, and agility.
- Drive reliability projects in collaboration with engineering teams.
- Perform troubleshooting and root-cause analysis of system operation issues.
- Take accountability for the platform's Service Level Agreements and Objectives.
- Provide infrastructure support during off-hours as needed.
- Take ownership of software infrastructure projects.
- Provide and receive constructive feedback through code and specification reviews.
View Full Description & ApplyYou'll be redirected to the employer's site