Principal Systems Engineer
New
A
Altera Digital HealthHealthcare Technology
This is a remote position open to candidates within the United States., 24x7 support coverageFull-TimePrincipal
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Experience
- 7+ years
- Required Skills
- PythonKubernetesMicrosoft SQL ServerAzure.NETTerraformServiceNow
Requirements
- 7+ years of experience supporting enterprise applications, infrastructure, or cloud environments.
- Strong experience with APM tools such as LogicMonitor, AppDynamics, Azure Monitor, SentryOne, Dynatrace, Datadog, or New Relic.
- Deep knowledge of Windows Server administration, IIS, .NET applications, Windows Clustering, MSMQ, Event Logs, and PerfMon.
- Strong SQL Server experience, including performance tuning, query optimization, blocking analysis, and Always On Availability Groups.
- Experience with Azure cloud environments and a solid understanding of networking fundamentals (DNS, TCP/IP, load balancing, firewalls).
- Familiarity with ServiceNow (or other ITSM platforms) and ITIL principles.
- Bachelor's Degree in Computer Science, Information Technology, or Engineering is preferred.
- Participation in an on-call rotation to support our 24x7 healthcare environment.
- Availability for occasional after-hours work for activations, upgrades, and major incidents.
Responsibilities
- Maintain and improve the reliability, availability, and performance of our production environments.
- Lead the investigation and resolution of complex application, database, and infrastructure issues.
- Participate in incident management, conduct root cause analysis (RCA), and contribute to post-incident reviews.
- Define and measure Service Level Indicators (SLIs) and Objectives (SLOs) to meet our service commitments.
- Develop proactive monitoring and alerting strategies to identify and resolve issues before they impact customers.
- Automate operational tasks using scripting and Infrastructure-as-Code (IaC) to improve efficiency.
- Partner with engineering and cloud teams to refine deployment, monitoring, and support processes.
- Provide technical leadership during major incidents and act as a key escalation point for critical issues.
View Full Description & ApplyYou'll be redirected to the employer's site