- Operate and maintain all Develocity instances and supporting services.
- Participate in a follow-the-sun on-call rotation, owning incident response and troubleshooting issues across the stack.
- Drive automation across application deployment, upgrades, monitoring, self-healing, and recovery.
- Build and maintain observability for all managed services (logging, metrics, tracing, and alerting).
- Work with engineering teams to build reliability into features from the start.
- Run incident response and retrospectives, and make sure we learn from them.
- Own disaster recovery, backups, and business continuity.
- Communicate with customers during incidents and maintenance windows.
- Optimize performance, resource usage, and costs.
- Help evolve our SaaS operations as we grow.