Staff Infrastructure Engineer - Data Streaming
New
J
JobgetherCybersecurity Infrastructure
Based in the United States, U.S. Eastern Time ZoneFull-TimeStaff
Salary$156,000–$215,000 USD annually, with the applicable range varying by candidate location.
Apply NowOpens the employer's application page
Job Details
- Experience
- 8+ years of experience
- Required Skills
- AWSPythonGCPKubernetesApache KafkaGoRedisTerraform
Requirements
- 8+ years of experience in infrastructure, platform, or systems engineering.
- Deep hands-on expertise with self-hosted Kafka and Redis in Kubernetes environments.
- Strong understanding of Kubernetes internals and production best practices for stateless and stateful workloads.
- Experience delivering Database-as-a-Service, Messaging-as-a-Service, or Platform-as-a-Service capabilities.
- Experience working in multi-cloud environments (AWS, GCP, or Azure).
- Strong experience with Infrastructure as Code and GitOps methodologies (Terraform, ArgoCD, or Pulumi).
- Strong scripting or software development capabilities using Python, Go, or a comparable language.
- Solid understanding of CI/CD pipelines and workflow automation (GitHub Actions or Argo Workflows).
- Familiarity with modern deployment approaches including blue-green, canary, and rolling deployments.
Responsibilities
- Lead the architecture, deployment, and operation of distributed data services, including self-hosted Kafka and Redis, across large-scale Kubernetes clusters and multi-cloud environments.
- Build highly automated, self-service infrastructure that enables services to operate consistently across AWS, GCP, and air-gapped on-premises environments.
- Manage data infrastructure supporting more than 5 PB of daily ingestion, optimizing for high throughput, low latency, reliability, scalability, and cost efficiency.
- Consolidate and optimize multi-tenant Kafka environments to improve resilience, operational efficiency, and infrastructure economics.
- Drive lifecycle automation for Kafka and Redis using Infrastructure as Code and GitOps practices, including Terraform and ArgoCD.
- Establish and enforce platform standards for observability, high availability, backup, and disaster recovery.
- Partner with FinOps and engineering stakeholders to continuously identify opportunities to improve performance and reduce costs.
View Full Description & ApplyYou'll be redirected to the employer's site