Lead Cloud Platform & Dev Ops Engineer
New
U
University of ColoradoHigher Education IT
Applicants must be Colorado residents or able to relocate to Colorado within a month of starting employment with University of Colorado System Administration. This role is eligible to work remotely within Colorado.Full-TimeLead
Salary121,791.28 USD per year
Apply NowOpens the employer's application page
Job Details
- Experience
- Master's degree plus three (3) years of experience or Bachelor's degree plus five (5) years of experience as a Platform Engineer or similar position.
- Required Skills
- AWSPythonJavaMachine LearningGrafanaPrometheusCI/CDDevOps
Requirements
- Master's degree in CS, Computer/Electrical Engineering, or related IT field plus 3 years of experience as a Platform Engineer; OR Bachelor's degree plus 5 years of experience.
- 3 years of experience with programming languages like Python and/or Java, including machine learning frameworks.
- 3 years of experience leading or supporting DevOps initiatives, CI/CD pipelines, and infrastructure-as-code.
- 3 years of experience designing and maintaining public cloud solutions (e.g., AWS EKS, ECS, or Lambda) for production.
- 3 years of experience implementing monitoring, logging, and reliability practices (e.g., Prometheus, Grafana).
- 2 years of experience with artificial intelligence, API development, and integration technologies (e.g., Kafka or RabbitMQ).
- Experience collaborating with security or compliance stakeholders to support technical controls.
- Must be a Colorado resident or able to relocate to Colorado within one month of hire.
Responsibilities
- Design, implement, and maintain cloud platform architecture including security, identity, networking, and policy controls.
- Develop and manage CI/CD pipelines and release processes across development and production environments.
- Drive the adoption of platform engineering practices such as infrastructure-as-code and containerization.
- Support data-driven technologies, including machine learning and intelligent automation workloads.
- Define service performance metrics (SLOs/SLIs) and implement monitoring, observability, and incident response solutions.
- Collaborate with security and compliance teams to maintain platform security and data protection controls.
- Provide technical guidance and documentation to development and operations teams.
View Full Description & ApplyYou'll be redirected to the employer's site