Senior Site Reliability Engineer (SRE) I
New
C
Consensys IncorporatedBlockchain Infrastructure
APAC, APAC hoursFull-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Experience
- 4+ years
- Required Skills
- AWSPythonBashKubernetesGoGrafanaPrometheusCI/CDTerraform
Requirements
- 4+ years of professional experience in DevOps, SRE, or platform engineering.
- Proven experience managing production on-call responsibilities for systems with external SLAs.
- Proficiency in Kubernetes and Infrastructure-as-Code (Terraform or equivalent).
- Experience with GitOps-based deployment workflows.
- In-depth cloud experience, primarily with AWS; Azure experience is a plus.
- Strong expertise in observability practices including Prometheus, Grafana, and SLI/error budget management.
- Proficiency in scripting or automation using Python, Go, or Bash.
- Ability to work autonomously in a distributed team environment.
- Excellent documentation skills for runbooks and incident management.
- Experience operating in regulated environments with change control and audit processes is preferred.
- Prior experience with Ethereum or blockchain node operations (Besu, Geth, Teku) is highly desirable.
- Experience operating JVM services under load is a plus.
Responsibilities
- Manage front-line operational coverage for production networks during APAC hours, including monitoring, triage, escalation, and incident resolution.
- Participate in and design a follow-the-sun on-call rota.
- Execute deployment and maintenance windows during APAC hours, including node upgrades and chain releases.
- Develop and improve runbooks, alert tuning, and observability practices to reduce manual operational load.
- Perform platform build tasks including Kubernetes, GitOps, IaC, and CI/CD pipeline management.
- Act as the primary operational contact for APAC-based clients and partner teams.
View Full Description & ApplyYou'll be redirected to the employer's site