Software Engineer, Distributed Data Systems

New
O
OnehouseData Infrastructure
India-RemoteFull-Time
Salary not disclosed
Apply NowOpens the employer's application page

Job Details

Required Skills
JavaKubernetesC++LinuxDistributed Systems

Requirements

  • Strong object-oriented design and coding skills.
  • Proficiency in Java and/or C/C++.
  • Experience working on UNIX or Linux platforms.
  • Deep understanding of distributed multi-tiered systems and algorithms.
  • Practical experience with the inner workings of relational databases.
  • Ability to think abstractly to articulate and solve complex technical challenges.
  • Proven ability to prioritize feature development against technical debt.
  • Experience quickly prototyping optimization solutions.
  • Capability to analyze large and complex data sets.
  • Robust and clear communication skills.

Responsibilities

  • Productionize the next generation of the data tech stack by building software and data features that process ingested data.
  • Accelerate the open source and enterprise flywheel by working on the internals of Apache Hudi's transactional engine.
  • Design new concurrency control and transactional capabilities to maximize throughput.
  • Implement new indexing schemes optimized for incremental data processing and analytical query performance.
  • Design systems to scale and streamline metadata and data access across different query and compute engines.
  • Solve optimization problems to improve performance and cost-efficiency of distributed data processing algorithms on Kubernetes.
  • Leverage data from existing systems to identify inefficiencies, build prototypes, and validate solutions.
  • Collaborate with engineering teams to safely deploy and rollout optimized solutions in production.
View Full Description & ApplyYou'll be redirected to the employer's site
View details
Apply Now