Senior Software Engineer, Storage
New
J
JobgetherSoftware Engineering
USFull-TimeSenior
Salary$196,000–$230,000 USD
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years
- Required Skills
- SQLDistributed Systems
Requirements
- 5+ years of experience building and operating large-scale distributed systems, particularly in storage, data ingestion, backup and restore, or streaming.
- Hands-on experience with database or storage internals, including query execution, storage engines, replication, or consensus.
- Strong ability to understand, own, and contribute deeply to complex codebases, ideally including open-source systems.
- Proven experience maintaining, analyzing, and debugging high-severity production incidents.
- Strong software engineering skills with a focus on clean, testable, and maintainable production code.
- Strong understanding of distributed systems, reliability, performance, and operational best practices.
- Ability to investigate complex technical problems and translate findings into practical, scalable solutions.
- Strong collaboration and communication skills, particularly within a remote working environment.
- High ownership, sound technical judgment, and a consistent track record of delivering complex engineering projects.
- Ability to work effectively across infrastructure, platform, database, and application engineering teams.
Responsibilities
- Design, build, and operate scalable distributed storage and database systems supporting critical online workloads.
- Investigate database internals, query execution, storage engines, and SQL latency or availability issues to identify and resolve complex technical problems.
- Improve database performance, reliability, and stability across storage, query execution, consensus, change data capture, routing, and shard management.
- Partner with database technology vendors on root cause analysis, technical improvements, and upstream contributions to open-source projects.
- Design and implement infrastructure for database adoption, including two-way replication, automated failover and failback, and migration tooling.
- Collaborate across Kernel, SRE, ORM, KV, and SQL interface teams to make database onboarding and migrations reliable and efficient for internal users.
- Validate and safely roll out new database architectures while maintaining transactional correctness across distributed shards.
- Build supporting frameworks and tooling for monitoring, observability, permissions, service discovery, and database operations.
- Analyze and optimize traffic prioritization, connection pooling, latency, concurrency, and overall system performance.
- Take ownership of high-severity production incidents, conduct deep technical investigations, and drive durable improvements to system reliability.
View Full Description & ApplyYou'll be redirected to the employer's site