Senior Site Reliability Engineer
S
SecurityScorecardCybersecurity
Location: Remote (Canada)Full-TimeSenior
SalaryCAD 197,500 – CAD 225,000 (base plus bonus)
Apply NowOpens the employer's application page
Job Details
- Experience
- 6+ years
- Required Skills
- PythonBashKafkaKubernetesClickhouseGoTerraformHelm
Requirements
- 6+ years in SRE, DevOps, or Infrastructure roles, with significant production Kubernetes experience.
- Hands-on experience integrating AI/LLM tooling into engineering or operational workflows (e.g., MCP servers, AI agents acting on infrastructure).
- Clear grasp of the security and governance considerations of giving AI access to production.
- Proven success building CI/CD pipelines (GitHub Actions, Jenkins, GitLab CI, or similar).
- Strong with Kubernetes internals and managed services like EKS, GKE, or AKS.
- Expertise with Infrastructure as Code (Terraform, Helm, Pulumi) and GitOps.
- Proficient in Python, Bash, or Go.
- Knowledge of observability tooling (Prometheus, Grafana, Datadog, OpenTelemetry).
- Production experience with Kafka, Flink, and ClickHouse.
- Strong communication and cross-team collaboration skills.
Responsibilities
- Design, build, and scale Kubernetes infrastructure for secure, multi-tenant, high-availability applications.
- Build and operate AI tooling infrastructure — stand up MCP servers and establish secure, governed AI access and guardrails for production systems.
- Optimize and maintain CI/CD pipelines, improving reliability, speed, and rollback safety.
- Implement progressive delivery strategies such as blue/green and canary deployments.
- Advance Infrastructure as Code with Terraform, Helm, and Argo CD, defining reusable patterns for the org.
- Operate and optimize streaming and analytics infrastructure: Kafka, Flink, and ClickHouse.
- Build automated testing into the CI/CD lifecycle.
- Improve system observability — define SLOs, alerts, and dashboards.
- Lead incident response and postmortems, focusing on root cause and durable fixes.
- Mentor engineers across teams on Kubernetes, CI/CD, and cloud infrastructure.
View Full Description & ApplyYou'll be redirected to the employer's site