Senior Site Reliability Engineer, NetBox Delivery
New
N
NetBox LabsNetwork automation
Location: LATAM, Remote; Secondary Locations: US Remote, UK, RemoteFull-TimeSenior
Salary$75K - $195K; $75K – $195K • Offers Equity • Offers Bonus • Multiple Ranges; Remote, LATAM: Remote, LATAM $75K – $85K; Remote, UK: Remote, UK £100K – £110K • Offers Equity • Offers Bonus; Remote, US: Remote, US $180K – $195K • Offers Equity • Offers Bonus
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years in software engineering, platform engineering, or SRE
- Required Skills
- AWSPostgreSQLDjangoKubernetesGrafanaPrometheusTerraformGitHub ActionsHelm
Requirements
- Have 5+ years of experience in software engineering, platform engineering, or SRE.
- Have proven experience writing robust, maintainable code.
- Have production experience with Django and Postgres at scale, including schema design, migration risk, and query performance under real load.
- Have strong container build skills, including base image design and Python dependency management.
- Know supply chain security practices such as vulnerability scanning and image signing.
- Have hands-on experience with AWS, including EC2, VPC, IAM, and RDS.
- Have hands-on experience with Kubernetes and Helm, GitHub Actions, ArgoCD or FluxCD, Terraform, Prometheus, and Grafana, or a close equivalent stack.
- Have hands-on experience building inside an AI-augmented development harness, including Claude Code and workflows for reliable agentic tooling.
- Have a track record of driving work across team boundaries, from writing an RFC to completing a cross-team migration.
- Nice to have familiarity with the NetBox ecosystem or network automation.
- Nice to have open source contributions or project involvement.
- Nice to have experience in a B2B software startup or high-growth organization.
Responsibilities
- Own the NetBox build and release pipeline, from base images to downstream availability on Cloud and Enterprise.
- Build the release handoff between NetBox Core and the Cloud and Enterprise teams.
- Improve NetBox production performance and reliability, from application startup to Postgres performance.
- Build application and release-pipeline observability, including monitoring, alerting, and SLOs.
- Serve as the escalation point for performance and reliability issues and fix underlying causes in NetBox Core when appropriate.
- Strengthen supply chain security and support SOC 2 compliance for the build pipeline.
- Share on-call duties and lead incident response and postmortems for the area.
View Full Description & ApplyYou'll be redirected to the employer's site