Senior Site Reliability Engineer – Network Observability
New
B
BlueprintTechnology Solutions
RemoteFull-TimeSenior
Salary50 - 55 USD per hour
Apply NowOpens the employer's application page
Job Details
- Experience
- Five to seven years of enterprise experience in systems engineering, network engineering, site reliability engineering, or a related IT infrastructure role.
- Required Skills
- PythonBashAzureLinuxNetworking
Requirements
- Bachelor’s degree in computer science, computer engineering, information technology, or a related technical field, or equivalent professional experience.
- Five to seven years of enterprise experience in systems engineering, network engineering, site reliability engineering, or a related IT infrastructure role.
- At least five years of Linux system administration experience.
- At least three years of network engineering experience.
- At least three years of hands-on Syslog-NG experience.
- Strong understanding of network observability and management technologies, including SNMP, SNMP Traps, NetFlow, and gNMI.
- Strong knowledge of enterprise networking, including routing and switching protocols.
- Hands-on experience with Azure or a comparable cloud platform.
- Experience operating and troubleshooting network monitoring systems in a large enterprise environment.
- Experience with system capacity planning, configuration management, compliance audits, and performance analysis.
- Ability to investigate complex incidents involving infrastructure, operating systems, applications, and network telemetry.
Responsibilities
- Operate and support network observability platforms, with a primary focus on Syslog-NG and Trapd.
- Administer and operate Linux and Windows virtual machines hosted in a cloud environment.
- Automate recurring operational tasks using Bash, PowerShell, or Python.
- Investigate automated alerts, customer-reported incidents, and platform performance issues.
- Troubleshoot complex network observability configurations, software applications, and operating systems.
- Conduct system capacity planning and recommend improvements to support scalability and reliability.
- Assist network and security engineers with analyzing traffic patterns, network telemetry, and resource utilization.
- Implement DevOps practices, including CI/CD pipelines, source control, and infrastructure as code.
- Participate in an on-call rotation and provide timely incident response.
View Full Description & ApplyYou'll be redirected to the employer's site