Senior Site Reliability Engineer, Security

New
J
JobgetherCloud Security
Canada, flexible schedule designed to accommodate distributed teams and different time zonesFull-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page

Job Details

Required Skills
AWSGCPKubernetesAzureCI/CDTerraform

Requirements

  • Significant professional experience operating production systems as an SRE, Platform, Infrastructure, or Security Engineer.
  • Deep hands-on experience with Kubernetes and containerized production environments.
  • Strong experience with at least one major cloud provider (AWS, GCP, or Azure) and its security model.
  • Strong understanding of IAM, workload identity, networking, TLS/PKI, encryption, and secrets management.
  • Experience with Infrastructure as Code (e.g., Pulumi, Terraform, or AWS CDK).
  • Experience designing, operating, and securing CI/CD and software delivery pipelines.
  • Strong Linux and container security fundamentals.
  • Experience with observability, monitoring, logging, production operations, and incident response.
  • Strong software engineering capabilities with demonstrated ability to automate infrastructure and security challenges.
  • Sound security judgment and the ability to assess technical risk in distributed systems.
  • Ability to work independently, take ownership of initiatives, and drive projects across teams.
  • Strong communication and technical leadership skills.

Responsibilities

  • Operate and evolve production Kubernetes and cloud infrastructure with a focus on availability, scalability, performance, reliability, and security.
  • Strengthen infrastructure security across Kubernetes, cloud platforms, networking, IAM, workload identity, secrets management, encryption, and service-to-service communication.
  • Secure the software supply chain from source code through production, including CI/CD pipelines, dependencies, container images, build artifacts, provenance, vulnerability scanning, and deployment controls.
  • Build automated security guardrails, policies, and secure defaults into the engineering platform to reduce reliance on manual security processes.
  • Improve vulnerability management practices by identifying, prioritizing, and addressing vulnerabilities across containers, dependencies, infrastructure, and cloud services.
  • Expand security telemetry, logging, monitoring, alerting, and detection capabilities, while developing and exercising effective security incident-response procedures.
  • Participate in the SRE on-call rotation and lead or contribute to production incidents, using post-incident learnings to improve reliability.
  • Partner with engineering teams on architecture and infrastructure decisions, threat modeling, and security considerations throughout the development lifecycle.
View Full Description & ApplyYou'll be redirected to the employer's site
View details
Apply Now