Senior Automation & Observability Engineer

New
E
EnsonoManaged Services
Remote - United StatesFull-TimeSenior
Salary$113,000 to $147,000 annually
Apply NowOpens the employer's application page

Job Details

Experience
7 to 10+ years
Required Skills
PythonGrafanaPrometheusAnsible

Requirements

  • 7 to 10+ years of experience in Monitoring, Observability, Infrastructure Operations, SRE, or Platform Engineering.
  • Hands-on experience with Grafana, IBM Instana, SolarWinds, Telegraf, Prometheus, and InfluxDB.
  • Strong proficiency in Python, PowerShell, Shell Scripting, and VBScript.
  • Experience with Infrastructure as Code (IaC) using Ansible or Puppet.
  • Familiarity with enterprise logging, telemetry, and event correlation platforms.
  • Experience supporting large-scale enterprise environments including Linux, Windows, VMware, and Citrix.
  • Strong troubleshooting, incident management, and problem-solving skills.
  • Knowledge of ITIL frameworks and operational support processes.
  • Ability to participate in 24x7 operations and on-call rotations.

Responsibilities

  • Design, implement, and maintain enterprise monitoring and observability solutions.
  • Monitor infrastructure, applications, middleware, and IoT services using IBM Instana, Grafana, and Prometheus.
  • Configure and troubleshoot Enterprise Logging & Telemetry (ELT) integrations.
  • Perform Root Cause Analysis (RCA) and incident management for infrastructure and application issues.
  • Develop automation solutions using Python, PowerShell, and Bash to improve operational efficiency.
  • Automate server provisioning and configuration management using Ansible or Puppet.
  • Maintain SOPs, runbooks, and operational documentation.
  • Coordinate with cross-functional teams during major incident bridges and DR exercises.
View Full Description & ApplyYou'll be redirected to the employer's site
$113,000 to $147,000 annually
Apply Now