Senior Software Engineer (Cloud Infrastructure)
New
Y
YipitDataMarket Research
This role may be performed fully remotely within the United States., East Coast hoursFull-TimeSenior
Salary$168,000 - 193,725
Apply NowOpens the employer's application page
Job Details
- Experience
- 4+ years
- Required Skills
- AWSKubernetesCI/CDDevOpsTerraformDatabricksDatadog
Requirements
- 4+ years of relevant software, infrastructure, SRE, platform engineering, or DevOps experience.
- Hands-on experience using AWS, Kubernetes, Datadog, Databricks, and Terraform in production environments.
- Experience operating business-critical production systems, including observability, incident response, and root-cause analysis.
- Experience improving CI/CD systems, infrastructure as code, deployment workflows, or internal developer platforms.
- Understanding of reliability concepts such as SLIs, SLOs, error budgets, capacity planning, and disaster-recovery objectives.
- Ability to independently own projects, make sound technical tradeoffs, and collaborate across teams.
- Comfort participating in an on-call or incident-response rotation.
- Bachelor’s degree or equivalent practical experience.
Responsibilities
- Design, build, and operate shared cloud infrastructure using AWS, Kubernetes, Terraform, Databricks, Cloudflare, and related cloud-native technologies.
- Deliver SRE and DevOps initiatives that improve reliability, scalability, observability, deployment safety, and operational readiness.
- Build reusable infrastructure modules, automation, and self-service workflows that reduce manual work and improve the developer experience.
- Help define and implement service-level indicators, service-level objectives, monitoring, alerting, and error-budget practices for critical systems.
- Participate in incident response and improve operational outcomes through clear runbooks, effective post-incident reviews, and durable corrective actions.
- Strengthen disaster-recovery readiness through recovery planning, automation, testing, and remediation of identified gaps.
- Improve CI/CD workflows and infrastructure delivery so engineering teams receive faster feedback and can deploy confidently.
- Partner with application, Data Platform, and Data Engineering teams to understand infrastructure needs and help teams operate their workloads effectively.
- Improve cloud efficiency through thoughtful architecture, capacity planning, Kubernetes resource optimization, cost visibility, and automation.
- Contribute to technical standards, architecture decisions, documentation, and the evolution of the team’s sustainable 24/7 operating model.
View Full Description & ApplyYou'll be redirected to the employer's site