Senior Systems Software Engineer - NV Cloud Functions

New
J
JobgetherCloud Infrastructure
IndiaFull-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page

Job Details

Experience
3+ years
Required Skills
PythonBashJavaKubernetesGoRustLinuxGitLab

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related discipline, or equivalent practical experience.
  • 3+ years of hands-on professional software engineering experience.
  • Expert-level proficiency in at least one systems programming language, such as Go, C, or Rust.
  • Strong understanding of data structures, algorithms, distributed software architecture, and systems engineering principles.
  • Hands-on experience with Kubernetes, container orchestration, and container technologies.
  • Experience automating software delivery and infrastructure workflows using continuous integration and deployment frameworks such as GitLab and ArgoCD.
  • Strong scripting skills in Bash, Python, or a comparable scripting language.
  • Solid understanding of Linux or other Unix-like operating systems and familiarity with system and kernel internals.
  • Strong understanding of performance, security, reliability, and scalability considerations in complex distributed systems.
  • Clear technical communication skills and an interest in contributing to open-source projects.

Responsibilities

  • Design, develop, and ship scalable services using Java, Go, and Rust for a distributed GPU workload platform.
  • Improve the performance, reliability, scalability, and operational behavior of systems that route AI workloads across distributed GPU fleets.
  • Build and optimize cloud-native services capable of supporting inference, streaming, and batch workloads.
  • Develop and maintain control-plane and edge components for a distributed, open-source platform.
  • Automate and optimize build, testing, integration, deployment, and release processes.
  • Work with Kubernetes and container technologies to develop reliable workload orchestration capabilities.
  • Collaborate with engineering teams to integrate the platform with related scheduling and AI infrastructure.
  • Investigate complex distributed systems challenges and balance performance, security, and scalability.
  • Contribute to public open-source repositories through code, design proposals, and documentation.
  • Triage community issues and maintain a high-quality contributor experience.
View Full Description & ApplyYou'll be redirected to the employer's site
View details
Apply Now