Senior Network Reliability Engineer, Incident Management
New
S
SkyloTelecommunications
Remote, US, Global follow-the-sun on-call rotationFull-TimeSenior
Salary$125,000 – $135,000 base salary + equity
Apply NowOpens the employer's application page
Job Details
- Experience
- 5-10+ years
- Required Skills
- KubernetesJiraGrafanaPrometheusServiceNow
Requirements
- 5-10+ years of experience in telecom/wireless operations, network operations, or NRE in a production 24x7 environment.
- Proven ability to independently manage incident bridge calls and maintain bridge discipline.
- Strong understanding of 5G functional components (AMF, SMF, UPF) and RAN architecture (CUSM, eCPRI, PTP/SyncE).
- Expertise in incident and outage management under high-pressure conditions.
- Hands-on experience with observability platforms such as Grafana, Prometheus, or Loki.
- Kubernetes operational literacy including pod troubleshooting and log analysis.
- Strong written communication skills for producing incident timelines and executive updates.
- Proficiency in ticketing systems like Jira or ServiceNow.
- Experience with on-call alerting tools such as PagerDuty.
- Ability to collaborate across engineering, operations, and external partner teams.
Responsibilities
- Serve as the central command point during all network degradations and service disruptions.
- Manage the end-to-end incident lifecycle for Sev 1-4, including bridge call initiation and vendor engagement.
- Prioritize incidents based on urgency and business impact using OSS telemetry.
- Produce structured incident timelines and documentation within two hours of closure.
- Track network availability KPIs and SLA obligations to MNO partners.
- Participate in a 24x7 global follow-the-sun on-call rotation.
- Coordinate post-incident reviews (PIR) and identify root causes for operational improvements.
View Full Description & ApplyYou'll be redirected to the employer's site