Senior Infrastructure Engineer
New
V
Valarian TechnologiesSecurity Technology
If you want to join us in the London office – it’s quite nice! – you’re welcome to; and if you’d prefer to work remotely, that’s fine too.Full-TimeSenior
SalaryA competitive salary
Apply NowOpens the employer's application page
Job Details
- Required Skills
- KubernetesLinuxTerraformHelmNetworking
Requirements
- Strong experience in infrastructure engineering, platform engineering, DevOps, SRE, or production systems engineering.
- Deep hands-on experience operating Kubernetes in production.
- Strong understanding of Linux systems, containers, networking, storage, and distributed infrastructure.
- Strong networking fundamentals, including TCP/IP, DNS, TLS, routing, load balancing, ingress, egress, network policy, and service discovery.
- Experience debugging difficult infrastructure issues across clusters, nodes, pods, networks, storage, and workloads.
- Experience with Kubernetes networking, CNI, service mesh, network policy, or secure service-to-service communication.
- Experience with infrastructure as code and GitOps workflows, especially Terraform, Argo CD, Helm, Kustomize, or similar.
- Experience building reliable infrastructure for production environments.
- Strong operational judgement regarding reliability, resilience, failure domains, and production risk.
- Strong ownership mindset and clear communication skills.
Responsibilities
- Design, build, operate, and improve Kubernetes-based infrastructure for ACRA.
- Own core infrastructure plumbing across networking, storage, workload scheduling, service communication, ingress, egress, DNS, certificates, and cluster-level security.
- Operate and debug production Kubernetes environments across GCP, on-premise, bare-metal, sovereign cloud, air-gapped, and customer-managed deployments.
- Work on multi-cluster Kubernetes environments, cluster networking, network policy, service mesh, and secure service-to-service communication.
- Operate and evolve distributed storage systems including capacity planning, replication, and recovery.
- Build infrastructure automation that improves repeatability, reliability, and operational safety.
- Improve observability across infrastructure layers including metrics, logs, traces, and alerting.
- Investigate and resolve complex production issues across networking, storage, Kubernetes, and Linux.
View Full Description & ApplyYou'll be redirected to the employer's site