Senior Staff Infrastructure Engineer – Kubernetes Platform
New
T
TensorWaveCloud Infrastructure
RemoteFull-TimeStaff
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Experience
- 7+ years of experience in infrastructure, platform engineering, or distributed systems
- Required Skills
- KubernetesLinuxDistributed Systems
Requirements
- 7+ years of experience in infrastructure, platform engineering, or distributed systems.
- Deep experience operating Kubernetes at scale in production environments.
- Strong understanding of Kubernetes internals (API server, scheduler, controller manager, etcd).
- Proven experience scaling Kubernetes across multiple clusters and regions/data centers.
- Strong Linux systems expertise.
- Deep troubleshooting ability across Kubernetes, container runtimes, and networking stacks.
- Experience with CNI plugins (Cilium preferred).
- Strong understanding of networking, traffic patterns, resource isolation, and scheduling.
- Experience in CSP, hyperscale, or large-scale environments is strongly preferred.
- Experience with virtual cluster technologies (vcluster, Kamaji) is a plus.
- Experience supporting GPU workloads in Kubernetes is a plus.
- Familiarity with NUMA-aware scheduling and RDMA high-throughput networking is a plus.
Responsibilities
- Design and evolve Kubernetes control plane architecture across regions.
- Implement multi-tenant cluster models including virtual cluster approaches like vcluster or Kamaji.
- Manage reliability, lifecycle, and incident response for Kubernetes platforms in production.
- Design strategies for multi-region scaling and cluster topology.
- Optimize ingress/egress architectures and pod-to-pod networking using CNI plugins like Cilium.
- Improve observability across control plane components and cluster health.
- Collaborate with DevOps and Infrastructure teams to align platform design with compute and storage capabilities.
View Full Description & ApplyYou'll be redirected to the employer's site