Principal Site Reliability Engineer, Platform Engineering
New
G
GitLabDevSecOps Platform
Location: Remote, Canada; Remote, United Kingdom; Remote, United StatesFull-TimePrincipal
Salary223,200 - 380,400 USD per year
Apply NowOpens the employer's application page
Job Details
- Required Skills
- PythonRubyGo
Requirements
- Deep expertise in Site Reliability, Platform, Infrastructure, or Backend Engineering.
- Experience designing and operating large-scale production systems.
- Hands-on experience with cloud infrastructure, automation, and observability.
- Proficiency with infrastructure as code and modern production engineering practices.
- Strong software engineering fundamentals with experience in Go, Ruby, or Python.
- Expertise in distributed systems and systems design.
- Proven track record of technical leadership across multiple teams.
- Experience leading platform or infrastructure transformations including modernization and scaling.
- Experience improving operational readiness and service ownership.
- Exceptional technical communication and influence skills.
Responsibilities
- Set technical direction for GitLab Dedicated, shaping architecture and platform strategy.
- Lead platform transformations across resilience, failover, tenant orchestration, change management, and self-service tooling.
- Drive scalable, modular architecture that aligns Dedicated with broader Cells strategy.
- Strengthen service ownership and operational maturity across engineering teams.
- Identify and address systemic reliability and scalability risks using production signals and incident patterns.
- Establish reusable platform patterns and automation to reduce operational toil.
- Lead complex technical decisions balancing reliability, security, cost, and maintainability.
- Advance engineering excellence through architectural leadership and mentorship.
View Full Description & ApplyYou'll be redirected to the employer's site