Senior Site Reliability Engineer - Cloud Platform
New
J
JobgetherCloud Infrastructure
CanadaFull-TimeSenior
SalaryCAD $107,000–$161,000 per year
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years
- Required Skills
- AWSPythonCI/CDLinuxCloudFormation
Requirements
- 5+ years of experience in Site Reliability Engineering, Platform Engineering, or Infrastructure Engineering.
- Strong Python development skills for automation and infrastructure tooling.
- Hands-on experience with Infrastructure as Code (AWS CloudFormation and/or AWS CDK).
- Strong working knowledge of core AWS services (IAM, VPC, EC2, Lambda, managed databases).
- Solid Linux systems administration and production troubleshooting skills.
- Experience with incident management and operational excellence.
- Experience with multi-account AWS environments or AWS Organizations.
- Familiarity with CI/CD pipelines, GitOps, policy-as-code, and cloud governance.
- Knowledge of AWS networking (Transit Gateway, IPAM) and FinOps practices.
- Familiarity with AI-assisted engineering workflows.
Responsibilities
- Operate and scale production AWS infrastructure, taking ownership of the reliability and health of services.
- Design, develop, and maintain cloud platform capabilities using Python, CloudFormation, and AWS CDK.
- Drive cloud cost optimization initiatives to improve infrastructure efficiency.
- Strengthen observability through monitoring, alerting, dashboards, and operational tooling.
- Participate in on-call rotations, lead incident response, and conduct post-incident reviews.
- Support strategic cloud initiatives including networking, identity, governance, and multi-account architecture.
- Review technical designs and code, maintain runbooks, and mentor other engineers.
- Leverage AI-assisted engineering tools to automate tasks and reduce operational toil.
View Full Description & ApplyYou'll be redirected to the employer's site