Senior DevOps Engineer
T
TechnologyAdviceB2B Technology
India - Remote, 1:30 PM to 10:30 PM ISTContractSenior
Salary1,500 - 2,000 INR per year
Apply NowOpens the employer's application page
Job Details
- Required Skills
- AWSDockerPythonKubernetesCI/CDLinuxTerraformDatadog
Requirements
- Deep experience designing, operating, and troubleshooting highly available production systems in cloud environments (AWS or equivalent).
- Strong proficiency with Kubernetes and container technologies (e.g., Docker).
- Extensive experience with Infrastructure as Code using Terraform, OpenTofu, or similar.
- Strong understanding of CI/CD, GitOps, and modern software delivery workflows (e.g., GitHub Actions, ArgoCD).
- Experience improving production observability with platforms like Datadog.
- Strong scripting or programming ability using Python, Bash, or equivalent for automation.
- Working knowledge of networking, DNS, CDN/WAF, load balancing, and TLS.
- Experience supporting stateful systems such as MySQL, Redis, or Redshift.
- Ability to implement secure infrastructure patterns, access controls, and secrets management.
- Strong Linux and systems troubleshooting capabilities across networks and cloud services.
- Experience leveraging AI-assisted engineering tools for infrastructure development and automation.
- Proven track record of independently driving platform initiatives and resolving complex production incidents.
Responsibilities
- Own and continuously improve cloud infrastructure and platform capabilities supporting applications and data systems.
- Design and maintain secure, scalable infrastructure using Infrastructure as Code (IaC) and cloud-native practices.
- Build automation and self-service platform tooling to eliminate repetitive operational toil.
- Evolve Kubernetes/container infrastructure to standardize deployment, scaling, and recovery.
- Manage and improve CI/CD and GitOps workflows for consistent and safe software delivery.
- Partner with cross-functional engineering teams to address operational friction and increase leverage.
- Enhance observability through monitoring, logging, tracing, alerting, and actionable service health indicators.
- Lead technical investigations for production incidents to ensure systemic improvements.
- Evaluate architecture for cost-efficiency and security, proactively reducing spend and infrastructure risk.
View Full Description & ApplyYou'll be redirected to the employer's site