Senior Automation & Observability Engineer
New
E
EnsonoManaged Services
Remote - United StatesFull-TimeSenior
Salary$113,000 to $147,000 annually
Apply NowOpens the employer's application page
Job Details
- Experience
- 7 to 10+ years
- Required Skills
- PythonGrafanaPrometheusAnsible
Requirements
- 7 to 10+ years of experience in Monitoring, Observability, Infrastructure Operations, SRE, or Platform Engineering.
- Hands-on experience with Grafana, IBM Instana, SolarWinds, Telegraf, Prometheus, and InfluxDB.
- Strong proficiency in Python, PowerShell, Shell Scripting, and VBScript.
- Experience with Infrastructure as Code (IaC) using Ansible or Puppet.
- Familiarity with enterprise logging, telemetry, and event correlation platforms.
- Experience supporting large-scale enterprise environments including Linux, Windows, VMware, and Citrix.
- Strong troubleshooting, incident management, and problem-solving skills.
- Knowledge of ITIL frameworks and operational support processes.
- Ability to participate in 24x7 operations and on-call rotations.
Responsibilities
- Design, implement, and maintain enterprise monitoring and observability solutions.
- Monitor infrastructure, applications, middleware, and IoT services using IBM Instana, Grafana, and Prometheus.
- Configure and troubleshoot Enterprise Logging & Telemetry (ELT) integrations.
- Perform Root Cause Analysis (RCA) and incident management for infrastructure and application issues.
- Develop automation solutions using Python, PowerShell, and Bash to improve operational efficiency.
- Automate server provisioning and configuration management using Ansible or Puppet.
- Maintain SOPs, runbooks, and operational documentation.
- Coordinate with cross-functional teams during major incident bridges and DR exercises.
View Full Description & ApplyYou'll be redirected to the employer's site