Software Engineer, Distributed Data Systems
New
O
OnehouseData Infrastructure
India-RemoteFull-Time
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Required Skills
- JavaKubernetesC++LinuxDistributed Systems
Requirements
- Strong object-oriented design and coding skills.
- Proficiency in Java and/or C/C++.
- Experience working on UNIX or Linux platforms.
- Deep understanding of distributed multi-tiered systems and algorithms.
- Practical experience with the inner workings of relational databases.
- Ability to think abstractly to articulate and solve complex technical challenges.
- Proven ability to prioritize feature development against technical debt.
- Experience quickly prototyping optimization solutions.
- Capability to analyze large and complex data sets.
- Robust and clear communication skills.
Responsibilities
- Productionize the next generation of the data tech stack by building software and data features that process ingested data.
- Accelerate the open source and enterprise flywheel by working on the internals of Apache Hudi's transactional engine.
- Design new concurrency control and transactional capabilities to maximize throughput.
- Implement new indexing schemes optimized for incremental data processing and analytical query performance.
- Design systems to scale and streamline metadata and data access across different query and compute engines.
- Solve optimization problems to improve performance and cost-efficiency of distributed data processing algorithms on Kubernetes.
- Leverage data from existing systems to identify inefficiencies, build prototypes, and validate solutions.
- Collaborate with engineering teams to safely deploy and rollout optimized solutions in production.
View Full Description & ApplyYou'll be redirected to the employer's site