- Productionize the next generation of the data tech stack by building software and data features that process ingested data.
- Accelerate the open source and enterprise flywheel by working on the internals of Apache Hudi's transactional engine.
- Design new concurrency control and transactional capabilities to maximize throughput.
- Implement new indexing schemes optimized for incremental data processing and analytical query performance.
- Design systems to scale and streamline metadata and data access across different query and compute engines.
- Solve optimization problems to improve performance and cost-efficiency of distributed data processing algorithms on Kubernetes.
- Leverage data from existing systems to identify inefficiencies, build prototypes, and validate solutions.
- Collaborate with engineering teams to safely deploy and rollout optimized solutions in production.
JavaKubernetesC+++2 more