Senior Cloud Operations Engineer
New
D
DragosCybersecurity
United StatesFull-TimeSenior
Salary$165,000.00
Apply NowOpens the employer's application page
Job Details
- Experience
- 4+ years
- Required Skills
- AWSPythonBashGCPAzureTerraformDatadog
Requirements
- 4+ years of hands-on cloud operations experience across one or more of Azure, AWS, or GCP.
- Strong proficiency with Terraform for infrastructure-as-code.
- Experience operating cloud environments at scale, including fleet management and drift detection.
- Hands-on experience with multi-tenant cloud architectures and customer-facing environment management.
- Solid knowledge of cloud networking: VPCs, VNets, peering, transit gateways, DNS, firewalls, and hybrid connectivity.
- Experience with IAM across AWS, Azure Entra ID, and/or GCP.
- Proficiency with Datadog for monitoring, alerting, dashboards, and log pipelines.
- Ability to participate in a PagerDuty on-call rotation.
- Experience with SRE practices including SLO definition and error budget management.
- Strong scripting skills in Python and/or Bash.
- Familiarity with compliance frameworks like FedRAMP, SOC2, and NIST CSF.
Responsibilities
- Operate, maintain, and improve the Dragos cloud fleet across Azure, AWS, and GCP.
- Own the full customer environment lifecycle including onboarding, configuration, upgrades, and offboarding.
- Build and maintain Terraform-based infrastructure-as-code for provisioning and standardization.
- Manage fleet health, drift detection, patching, and version lifecycle management.
- Design and enforce multi-tenant isolation patterns, including blast radius containment and cross-account access controls.
- Configure cloud networking components such as VPCs, VNets, peering, transit gateways, DNS, firewalls, and load balancers.
- Implement and maintain IAM, secrets management, and compliance posture.
- Maintain Datadog observability including monitors, dashboards, log pipelines, and SLOs.
- Participate in an on-call PagerDuty rotation for production incident response.
- Drive SRE practices, including runbook maintenance and post-incident reviews.
View Full Description & ApplyYou'll be redirected to the employer's site