Senior Site Reliability Operations Engineer - Finance
New
T
TruelogicFinancial services
Location: LatAm, Operations swing shift from 2 PM to 10:30 PM PST timezone.Full-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years in an Operations Center (SRO/NOC) or cloud infrastructure environment
- Required Skills
- AWSPythonJenkins
Requirements
- Have 5+ years of experience in an Operations Center (SRO/NOC) or cloud infrastructure environment.
- Have hands-on experience with full-stack application deployments.
- Be proficient in Windows and UNIX/Linux administration and troubleshooting.
- Have experience with scripting, grepping logs, analyzing performance metrics, and managing virtual servers and desktops.
- Have practical experience with AWS cloud services, including storage, virtual machines, and networking.
- Have experience with AWS CloudWatch, New Relic, Nagios, and SumoLogic.
- Have scripting or programming capability in PowerShell, Python, or bash.
- Have practical experience with Jenkins, GitLab, or similar CI/CD platforms.
- Have experience with ServiceNow or Jira ITSM ticket platforms and backup solutions such as CommVault, Veeam, or AWS Backup.
- Have verbal and written communication skills and experience working with technical teams, executive stakeholders, and external vendors.
- Advanced AWS certifications, AI/ML infrastructure-monitoring experience, ITIL-aligned or enterprise change/incident management experience, and a related bachelor's degree are pluses.
Responsibilities
- Monitor multi-platform IT infrastructure health using AWS CloudWatch, New Relic, Nagios, and SumoLogic, and refine alert thresholds.
- Troubleshoot complex issues across Linux/UNIX, Windows, virtual servers, and virtual desktop environments.
- Coordinate, automate, and execute code deployments using Jenkins, GitLab, or similar CI/CD tools.
- Collaborate with application developers, third-party vendors, and internal Incident Management.
- Lead medium- to large-scale infrastructure projects, including migrations, cloud upgrades, and performance tuning.
- Maintain SOPs in team knowledge bases.
- Oversee enterprise backup operations using CommVault, Veeam, and AWS Backup.
View Full Description & ApplyYou'll be redirected to the employer's site