Senior Site Reliability Operations Engineer - Finance

New
T
TruelogicFinancial services
Location: LatAm, Operations swing shift from 2 PM to 10:30 PM PST timezone.Full-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page

Job Details

Experience
5+ years in an Operations Center (SRO/NOC) or cloud infrastructure environment
Required Skills
AWSPythonJenkins

Requirements

  • Have 5+ years of experience in an Operations Center (SRO/NOC) or cloud infrastructure environment.
  • Have hands-on experience with full-stack application deployments.
  • Be proficient in Windows and UNIX/Linux administration and troubleshooting.
  • Have experience with scripting, grepping logs, analyzing performance metrics, and managing virtual servers and desktops.
  • Have practical experience with AWS cloud services, including storage, virtual machines, and networking.
  • Have experience with AWS CloudWatch, New Relic, Nagios, and SumoLogic.
  • Have scripting or programming capability in PowerShell, Python, or bash.
  • Have practical experience with Jenkins, GitLab, or similar CI/CD platforms.
  • Have experience with ServiceNow or Jira ITSM ticket platforms and backup solutions such as CommVault, Veeam, or AWS Backup.
  • Have verbal and written communication skills and experience working with technical teams, executive stakeholders, and external vendors.
  • Advanced AWS certifications, AI/ML infrastructure-monitoring experience, ITIL-aligned or enterprise change/incident management experience, and a related bachelor's degree are pluses.

Responsibilities

  • Monitor multi-platform IT infrastructure health using AWS CloudWatch, New Relic, Nagios, and SumoLogic, and refine alert thresholds.
  • Troubleshoot complex issues across Linux/UNIX, Windows, virtual servers, and virtual desktop environments.
  • Coordinate, automate, and execute code deployments using Jenkins, GitLab, or similar CI/CD tools.
  • Collaborate with application developers, third-party vendors, and internal Incident Management.
  • Lead medium- to large-scale infrastructure projects, including migrations, cloud upgrades, and performance tuning.
  • Maintain SOPs in team knowledge bases.
  • Oversee enterprise backup operations using CommVault, Veeam, and AWS Backup.
View Full Description & ApplyYou'll be redirected to the employer's site
View details
Apply Now