Senior Cloud Operations Engineer

New
D
DragosCybersecurity
United StatesFull-TimeSenior
Salary$165,000.00
Apply NowOpens the employer's application page

Job Details

Experience
4+ years
Required Skills
AWSPythonBashGCPAzureTerraformDatadog

Requirements

  • 4+ years of hands-on cloud operations experience across one or more of Azure, AWS, or GCP.
  • Strong proficiency with Terraform for infrastructure-as-code.
  • Experience operating cloud environments at scale, including fleet management and drift detection.
  • Hands-on experience with multi-tenant cloud architectures and customer-facing environment management.
  • Solid knowledge of cloud networking: VPCs, VNets, peering, transit gateways, DNS, firewalls, and hybrid connectivity.
  • Experience with IAM across AWS, Azure Entra ID, and/or GCP.
  • Proficiency with Datadog for monitoring, alerting, dashboards, and log pipelines.
  • Ability to participate in a PagerDuty on-call rotation.
  • Experience with SRE practices including SLO definition and error budget management.
  • Strong scripting skills in Python and/or Bash.
  • Familiarity with compliance frameworks like FedRAMP, SOC2, and NIST CSF.

Responsibilities

  • Operate, maintain, and improve the Dragos cloud fleet across Azure, AWS, and GCP.
  • Own the full customer environment lifecycle including onboarding, configuration, upgrades, and offboarding.
  • Build and maintain Terraform-based infrastructure-as-code for provisioning and standardization.
  • Manage fleet health, drift detection, patching, and version lifecycle management.
  • Design and enforce multi-tenant isolation patterns, including blast radius containment and cross-account access controls.
  • Configure cloud networking components such as VPCs, VNets, peering, transit gateways, DNS, firewalls, and load balancers.
  • Implement and maintain IAM, secrets management, and compliance posture.
  • Maintain Datadog observability including monitors, dashboards, log pipelines, and SLOs.
  • Participate in an on-call PagerDuty rotation for production incident response.
  • Drive SRE practices, including runbook maintenance and post-incident reviews.
View Full Description & ApplyYou'll be redirected to the employer's site
$165,000.00
Apply Now