Senior Network Reliability Engineer, Incident Management

New
S
SkyloTelecommunications
Remote, US, Global follow-the-sun on-call rotationFull-TimeSenior
Salary$125,000 – $135,000 base salary + equity
Apply NowOpens the employer's application page

Job Details

Experience
5-10+ years
Required Skills
KubernetesJiraGrafanaPrometheusServiceNow

Requirements

  • 5-10+ years of experience in telecom/wireless operations, network operations, or NRE in a production 24x7 environment.
  • Proven ability to independently manage incident bridge calls and maintain bridge discipline.
  • Strong understanding of 5G functional components (AMF, SMF, UPF) and RAN architecture (CUSM, eCPRI, PTP/SyncE).
  • Expertise in incident and outage management under high-pressure conditions.
  • Hands-on experience with observability platforms such as Grafana, Prometheus, or Loki.
  • Kubernetes operational literacy including pod troubleshooting and log analysis.
  • Strong written communication skills for producing incident timelines and executive updates.
  • Proficiency in ticketing systems like Jira or ServiceNow.
  • Experience with on-call alerting tools such as PagerDuty.
  • Ability to collaborate across engineering, operations, and external partner teams.

Responsibilities

  • Serve as the central command point during all network degradations and service disruptions.
  • Manage the end-to-end incident lifecycle for Sev 1-4, including bridge call initiation and vendor engagement.
  • Prioritize incidents based on urgency and business impact using OSS telemetry.
  • Produce structured incident timelines and documentation within two hours of closure.
  • Track network availability KPIs and SLA obligations to MNO partners.
  • Participate in a 24x7 global follow-the-sun on-call rotation.
  • Coordinate post-incident reviews (PIR) and identify root causes for operational improvements.
View Full Description & ApplyYou'll be redirected to the employer's site
$125,000 – $135,000 base salary + equity
Apply Now