Senior Backend Software Engineer - Agent Platform
New
J
JobgetherCybersecurity / SaaS
Fully remote work opportunity within the United States.Full-TimeSenior
SalaryBase salary range of $132,000–$182,000 USD
Apply NowOpens the employer's application page
Job Details
- Experience
- 5+ years
- Required Skills
- AWSDockerPostgreSQLPythonGCPJavaKafkaKubernetesGoDistributed Systems
Requirements
- 5+ years of professional backend software engineering experience, with strong expertise in at least one of Java, Go, or Python and the ability to work across the broader technology stack.
- Demonstrated reliability-first mindset, including experience handling complex production incidents, performing root-cause analysis, and implementing preventative engineering improvements.
- Strong experience designing, building, and operating large-scale distributed systems, with a solid understanding of failure modes, performance trade-offs, resilience patterns, and operational excellence.
- Ability to quickly navigate unfamiliar systems and codebases and troubleshoot complex technical problems across multiple service boundaries.
- Hands-on experience with cloud platforms such as AWS or GCP and technologies including Docker, Helm, and Kubernetes.
- Experience with distributed messaging and data technologies such as Kafka, PostgreSQL, Redis, Cassandra, ClickHouse, or comparable platforms.
- Familiarity with service communication technologies such as gRPC, REST, GraphQL, or similar APIs.
- Strong software engineering fundamentals, including testing, documentation, deployment, monitoring, refactoring, and production operations.
- Excellent communication and collaboration skills, with the ability to work effectively with engineering teams, product partners, technical stakeholders, and customer-facing groups.
- Demonstrated ownership, autonomy, curiosity, and sound judgment when driving ambiguous and technically complex problems toward successful outcomes.
- Ability to influence technical direction and mentor other engineers while working effectively across organizational boundaries.
- Experience in enterprise SaaS, cybersecurity, endpoint security, or another high-scale technology environment is highly desirable.
Responsibilities
- Lead rapid response to customer-critical production incidents, diagnosing, triaging, and resolving complex issues spanning multiple services across the agent platform.
- Conduct systematic root-cause analysis and translate incident findings into lasting improvements across reliability, scalability, observability, operability, and system resilience.
- Quickly understand unfamiliar services, architectures, and large codebases while partnering with engineering teams to troubleshoot and resolve cross-service failures.
- Design, develop, test, document, deploy, and operate large-scale distributed systems capable of processing millions of events per second with high availability and low latency.
- Build and evolve backend services responsible for policy, configuration, and command distribution to millions of endpoints worldwide.
- Maintain and improve existing services through refactoring, feature development, architectural enhancements, and ongoing operational improvements.
- Monitor application health, system metrics, and data integrity while strengthening operational visibility and platform stability.
- Translate business and functional requirements into robust, scalable, maintainable, and operationally sound technical solutions.
- Collaborate across engineering and other stakeholder groups to solve complex technical challenges, influence architecture and technical direction, and deliver scalable solutions.
- Evaluate and adopt technologies, engineering practices, and tooling that improve platform reliability, scalability, performance, and developer productivity.
View Full Description & ApplyYou'll be redirected to the employer's site