Senior Manager Site Reliability Engineer
New
G
GoDaddyCloud Operations / SaaS
Remote, India, Must be comfortable working across multiple time zones.Full-TimeManager
Salary5,253,000 - 11,830,000 INR per year
Apply NowOpens the employer's application page
Job Details
- Languages
- Fluent in English.
- Experience
- 12+ years of experience in software engineering, SRE, DevOps, infrastructure, cloud operations, or platform engineering.
- Required Skills
- CI/CDDevOps
Requirements
- 12+ years of experience in software engineering, Site Reliability Engineering, DevOps, infrastructure, cloud operations, or platform engineering.
- 6+ years of engineering leadership experience, including leading distributed, international, or cross-functional teams.
- Experience directing and scaling units of 10–20+ engineers, including senior roles.
- Proven experience building or reforming platform, infrastructure, or SRE organizations.
- Strong background in platform-as-a-product operating models, focusing on developer experience and self-service.
- Experience modernizing large-scale systems, including migration from monolith to cloud-native platforms.
- Deep operational experience with secure, reliable, and scalable systems in high-traffic environments.
- Expertise in modern cloud infrastructure, observability, incident management, automation, and infrastructure as code.
- Proven history of producing quantifiable results in developer productivity, reliability, and cost reduction.
- Fluent in English with excellent communication and influencing skills.
- Experience integrating globally distributed engineering cultures.
Responsibilities
- Build and lead a top-performing SRE and infrastructure team in India.
- Develop a robust operating model for efficient cooperation among global engineering teams.
- Promote platform-as-a-product thinking by developing self-service options for infrastructure, CI/CD, observability, and database operations.
- Drive infrastructure modernization from monolithic to cloud-native architectures.
- Establish reliability standards including SLOs, SLIs, and error budgets to improve availability and reduce MTTR.
- Lead automation and toil reduction initiatives to improve operational efficiency.
- Partner with product, engineering, and security leaders to align platform investments with business outcomes.
View Full Description & ApplyYou'll be redirected to the employer's site