Senior Software Engineer – Cloud Platform & Operations
New
T
TalentCloud GroupCloud Infrastructure
United StatesFull-TimeSenior
SalaryCompetitive salary, equity, and benefits
Apply NowOpens the employer's application page
Job Details
- Experience
- 5–7+ years
- Required Skills
- AWSPythonGCPKubernetesAzureGoCI/CDDistributed Systems
Requirements
- 5–7+ years in platform engineering, infrastructure, or SRE-focused roles.
- Strong proficiency in Go and Python (expertise in one with willingness to use both).
- Hands-on experience running Kubernetes in production.
- Solid understanding of distributed systems and cloud-native architectures.
- Experience with major cloud providers (AWS, GCP, or Azure).
- Familiarity with CI/CD, infrastructure-as-code, and automation tooling.
- Comfortable participating in on-call rotations and managing production incidents.
- Strong ownership mindset.
- Experience building Kubernetes operators or control-plane components (nice-to-have).
- Background in SaaS, database, or systems-level products (nice-to-have).
- Exposure to Prometheus, Grafana, OpenTelemetry, or similar observability tools (nice-to-have).
- Knowledge of networking, load balancing, or service meshes (nice-to-have).
Responsibilities
- Design, build, and operate foundational cloud platform components.
- Develop and maintain Kubernetes clusters, including custom operators.
- Write production-grade Go and Python code for platform services and automation.
- Improve the stability, performance, and cost efficiency of cloud environments across AWS, GCP, and Azure.
- Strengthen observability through metrics, logging, alerting, and monitoring frameworks.
- Participate in incident response, root-cause analysis, and long-term system hardening.
- Automate operational workflows, integrations, and infrastructure processes.
- Collaborate closely with Platform, Regions & Clusters, and Feature teams to ensure seamless delivery.
View Full Description & ApplyYou'll be redirected to the employer's site